Urevangelium

ur-eh-van-GAY-lee-um · the original Gospel

The Gospels across their earliest witnesses, and the tradition they carried forward

← Back to the Gospel table/Sources & rules

Editorial record · manifest 2026-08-28.2

Sources and governing rules

Urevangelium preserves each witness or textual tradition in its own form so readers can follow the transmission of the Gospels across millennia. This page names the immediate source behind every live column and the rules that source is allowed to follow.

Current certification statement: the site is an open textual-comparison project under active source collation. A source being present locally does not by itself certify every displayed word. Each status below describes the present source-to-display relationship.

Documentary limit

Chapters without an extant registered Greek New Testament papyrus

These are gaps in the surviving manuscript record, not merely gaps in Urevangelium’s transcription files. The Earliest Papyri column keeps its red lost dots throughout these chapters. Parchment codices and parallel traditions never supply or reconstruct a papyrus cell.

ChaptersSurviving papyrus record
Matthew 6–9Papyrus evidence reaches portions of Matthew 5 and resumes in Matthew 10.
Matthew 15–16P103 preserves Matthew 13:55–14:5; the next surviving registered papyrus material begins in Matthew 17.
Matthew 22Surviving papyrus material reaches Matthew 21 and resumes in Matthew 23.
Mark 3Surviving Mark papyri cover portions of chapters 1–2 and resume with P45 at Mark 4:36.
Mark 10P45 stops at Mark 9:31 and resumes at Mark 11:27.
Mark 13–16P45’s surviving Mark text ends at Mark 12:28. No later Mark chapter survives in a registered New Testament papyrus.
Luke 19–21P75 stops at Luke 18:18 and resumes at Luke 22:4; the intervening leaves are lost.

Column boundary

Vaticanus (GA 03) and Sinaiticus (GA 01) are displayed only in their own manuscript columns. Their surviving text may guide shared row location, but it never replaces a lost dot or becomes an Earliest Papyri reading.

Catalogue records: GA 03 ↗ · GA 01 ↗

Exceptions requiring separate labels

  • Matthew 6: P.Oxy. 5575 is a second-century parallel collection of sayings, not a manuscript of Matthew and not a papyrus-column source. Matthew 6 therefore retains lost dots. Oxford record ↗
  • Mark 16:9–20: Vaticanus and Sinaiticus end at 16:8. Washingtonianus (GA 032) and Alexandrinus (GA 02) are later parchment evidence and must not be represented as papyri.

Column records

Positions 3a and 3b share the Alexandrian toggle in the live table. Dates distinguish the history of a tradition from the date of the particular manuscript or edition actually displayed.

Position 1 · Greek papyrus witnesses

Earliest Papyri

Editorially provisional

What the column displays

A governed composite selecting extant Gospel papyrus evidence word by word; it is not a single reconstructed Greek text.

Dates

Individual papyri range from the second through seventh centuries CE. Each displayed siglum carries its own paleographic date.

Current coverage

65 registered papyri; 2,132 distinct Gospel verses have at least one coverage record.

Status finding

Direct CNTR readings and INTF checks coexist with explicitly provisional coverage stubs.

Immediate source material

CNTR Class 1 Gospel papyrus transcriptionstext

data/sources/earliest-papyrus/P*.txt

Local acquisition is not yet pinned to a commit · CC BY-SA 4.0 · source record ↗

INTF NTVMR diplomatic transcriptionsverification

data/cache/intf/

Cached per GA witness; retrieval revision not yet recorded · External verification source · source record ↗

STEPBible TAGNTalignment

data/sources/greek-shared/TAGNT-Mat-Jhn-CC-BY.txt

Local acquisition is not yet pinned · CC BY 4.0 · source record ↗

Governing rules

  • Keep every papyrus word in manuscript order and align only by a contiguous forward scan.
  • Show every siglum that actually attests the displayed location; multiple agreeing papyri may share a compact cell.
  • Preserve disagreeing readings separately in the data even when the interface later offers a compact view.
  • When attesting papyri disagree, display the reading of the papyrus ranked first by the public papyrus chronology: earliest starting year in the paleographic date range, then lower Gregory-Aland papyrus number as the deterministic display tie-breaker.
  • Attach compact-view sigla only to papyri that attest the selected displayed reading; retain dissenting papyri and readings in provenance.
  • Use lost status for non-extant material; a leading loss requires INTF confirmation.
  • Strip diacritics and expand nomina sacra for comparison only, never as an unrecorded alteration of stored evidence.
  • Coverage stubs may use TAGNT only as a visibly identified provisional reconstruction, never as transcribed papyrus text.

Not permitted

  • Silent TAGNT substitution
  • Automatic correction of genuine variants
  • Reordering manuscript words
  • Treating verse coverage as proof that every word survives

Next certification action: Add per-reading provenance and a visible stub/reconstruction state, then finish CNTR-to-INTF word-level collation.

Position 2 · Sahidic Coptic New Testament

Sahidic

Source verified

What the column displays

Sahidica NT 4.1.0, a normalized electronic Sahidic edition rather than one physical manuscript.

Dates

Sahidic Gospel tradition is ancient; the column does not display a single early codex. Sahidica NT version 4.1.0 (metadata dated 2021-03-31).

Current coverage

48,275 Sahidica word-groups displayed exactly once: 13,857 Matthew, 8,390 Mark, 14,237 Luke, and 11,791 John; zero missing, unexpected, or altered forms, with occurrence provenance on every word-group.

Status finding

All four Gospels are occurrence-complete and diplomatically exact against pinned Sahidica NT 4.1.0. This certifies the source-form inventory, not parallel-row placement: 42,803 placements remain computationally provisional and the current audit identifies 3,174 source-order breaks. Parallel alignment and published-translation alignment remain under review.

Immediate source material

Sahidica NT 4.1.0 via Coptic SCRIPTORIUMtext

data/sources/coptic-tt/*.tt

4.1.0 · Free-electronic-edition permission identified; SCRIPTORIUM academic-use wording requires documented clarification · source record ↗

Crum through KELLIA CCL — lexical aid, not contextual translationgloss

scripts/coptic/kellia-lexicon.xml

CCL v1.2 (2020) · CC BY-SA 4.0

Urevangelium Sahidic English evidence ledgerverification

data/sources/coptic-english/manifest.json; docs/audits/coptic-english-system/

v1; one deterministic decision per Sahidica word-group · Project-generated provenance and admission metadata

CrossWire CopSahHorner 1.5verification

data/sources/crosswire-copsahhorner/

1.5; package hash verified · Public domain module; transcription provenance insufficient for authoritative admission

George W. Horner — proposed published translation authority, subject to Coptic-text applicabilityverification

data/sources/horner-pilot/

The Coptic Version of the New Testament in the Southern Dialect; qualified human transcription pending · Public-use rights and transcription provenance must be recorded before admission

STEPBible TAGNTalignment

data/sources/greek-shared/TAGNT-Mat-Jhn-CC-BY.txt

Local acquisition is not yet pinned · CC BY 4.0

Governing rules

  • Preserve every Sahidica word-group exactly and attach edition, source file, verse, occurrence number, diplomatic form, and SHA-256 provenance.
  • Preserve Sahidica source order and word-group boundaries independently; never reshape the Coptic sequence merely to fit another tradition’s grid.
  • Represent many-to-many correspondence with alignment links or spans while retaining provenance for every Sahidica source unit.
  • Treat John 8 as the second logical chapter embedded in the distributed 43_John_07.tt file; keep John 7:53 and John 8:1–11 explicitly omitted.
  • Describe SCRIPTORIUM segmentation and linguistic annotation as automatic source layers, not manually Coptologist-validated facts.
  • Display Crum/KELLIA only as lexical aid.
  • Admit Horner English verbatim only when Horner’s underlying Coptic is exact or nonlexically equivalent to the corresponding Sahidica span.

Not permitted

  • Calling Sahidica simply Horner
  • Claiming that one normalized edition represents every Sahidic manuscript
  • Presenting automatic SCRIPTORIUM annotations as manually validated
  • Using TAGNT, another tradition, a dictionary, or project-generated wording as Sahidic translation
  • Filling Sahidica omissions from another tradition
  • Extracting Logos content for public redistribution without written permission

Next certification action: Acquire a qualified Horner Coptic-and-English transcription for contextual translation units, while preserving the complete multi-source decision ledger; repair parallel placement separately.

Position 3a · Alexandrian Greek manuscript witness

Vaticanus

Source verified

What the column displays

Codex Vaticanus, GA 03.

Dates

The broader Alexandrian stream predates the codex. Codex: approximately 325 CE.

Current coverage

All 3,779 canonical Gospel verses are classified: 63,511 INTF source tokens displayed as 63,546 lexical words, explicit textual omissions, and physical lacunae. English is certified for 63,543 words (99.995%); 134 certified system-generated lexical glosses are orange and 3 manuscript-event cases remain without English.

Status finding

The four-Gospel column is generated from the pinned INTF original-hand transcription and automatically collated against pinned CNTR GA 03. English is published only from reproducible decision ledgers; system-generated lexical English is orange. This is internal source certification, not a claim of external peer review.

Immediate source material

INTF NTVMR transcription of GA 03text

data/sources/vaticanus/intf/*.xml

NTVMR document 20003, PUBLISHED original-hand TEI; per-Gospel SHA-256 hashes recorded in the certification artifact · CC BY 4.0 · source record ↗

CNTR Class 1 transcription of GA 03verification

data/sources/vaticanus/03.txt

CNTR commit 4c0e9f94117ec3dc4ae40094aec044bb7a416a53; SHA-256 cea945958d065699d3ab42f05d2afa3be54af4551a68e2e0a32090cd9fa0bb7f · CC BY-SA 4.0 · source record ↗

STEPBible TAGNTalignment

data/sources/greek-shared/TAGNT-Mat-Jhn-CC-BY.txt

Local acquisition is not yet pinned · CC BY 4.0

MorphGNT SBLGNT morphologyverification

data/sources/greek-shared/morphgnt/*-morphgnt.txt

Commit aaed91e57c8e4a8dc9a2383e129ca5e75fe6393d; per-file SHA-256 hashes recorded in the secondary-source ledger · Morphological parsing and lemmatization CC BY-SA · source record ↗

PROIEL Greek New Testament Treebankverification

data/sources/greek-shared/proiel/greek-nt.xml

Commit 8e388967a1335ed12335ddc655fe46993ee7d57a; SHA-256 recorded in the secondary-source ledger · CC BY-NC-SA 3.0 · source record ↗

MorphGNT Morphological Lexiconverification

data/sources/greek-shared/morphgnt-lexicon/lexemes.yaml

Commit 0dca2af89f413cbb24f617ddbdc347e9d798ddf3; SHA-256 recorded in the secondary-source ledger · Content CC BY-SA 3.0 · source record ↗

MorphGNT Tischendorf 8 morphologyverification

data/sources/greek-shared/tischendorf-morphgnt/*.txt

Dataset 2.8 at commit 795f2f4f9fe7cb98bf8736b0c5cb59c43aa9c32e; per-file SHA-256 hashes recorded in the secondary-source ledger · Public domain · source record ↗

TBESG / Abbott-Smithgloss

data/sources/greek-shared/TBESG-CC-BY.txt

SHA-256 312f723d7b8ef263bbdfb0451c9b8057125804dfff390b6f8544cff2a84b57f4 · CC BY 4.0

Governing rules

  • Display GA 03, not a critical edition proxy.
  • Normalize case and accents only under the declared display policy.
  • Preserve lacunae, supplied text, uncertainty, corrections, and selected scribal hand.
  • Use TAGNT contextual English only after deterministic alignment to the INTF-controlled GA 03 word; verify its lexical identity against TBESG/Abbott-Smith.
  • For unmatched forms, require exact surface-form lemma agreement between MorphGNT and PROIEL, or a registered contextual/source-native rule with corroborating morphology.
  • Use pinned Tischendorf morphology only as an additional annotation witness; record its shared textual ancestry with PROIEL.
  • Withhold English whenever word alignment or lexical identification is ambiguous.
  • Exclude OCR and AI image transcription from certification.

Not permitted

  • NA28 text presented as Vaticanus
  • Silent resolution of corrections or uncertain letters
  • Filling a Vaticanus omission from another tradition
  • Crossing English meanings from another tradition column
  • Using OCR or AI image transcription to certify English
  • Treating Parker/Heinfetter as a diplomatic representation of GA 03

Next certification action: Retain the three remaining non-lexical or incomplete manuscript forms as explicit red manuscript-status annotations unless stronger Vaticanus-specific evidence emerges.

Position 3b (toggle) · Alexandrian Greek manuscript witness

Sinaiticus

Source verified

What the column displays

Codex Sinaiticus, GA 01, in the pinned CNTR base-reading transcription.

Dates

The broader Alexandrian stream predates the codex. Codex: approximately 350 CE.

Current coverage

All 3,779 Gospel records classified: 63,917/63,917 Greek source tokens displayed in source order; 14 damaged tokens, 2 supplied tokens, 1,746 nomina sacra, 547 explicit textual-omission cells, and 12,594 comparison gaps. Of 80,540 acquired Anderson words, 80,394 have extant GA 01 parents and 146 are deliberately suppressed where Anderson includes wording absent from the certified Greek or exceeds the three-word parent ceiling. Seven missing printed verse labels are resolved by exact Anderson substring boundaries; John 21:25 retains its 23 Greek words without invented English because Anderson provides no unit.

Status finding

Every displayed GA 01 Gospel token is occurrence-certified against the hash-pinned CNTR Class 1 transcription. The selected reading is the original scribe including recorded first-scribe corrections; later correctors a-c are excluded from display but retained in provenance. Anderson 1918 supplies all displayed English verbatim. Deterministic monotonic allocation permits at most three consecutive Anderson words per extant Greek parent and never generates English.

Immediate source material

CNTR Class 1 transcription of GA 01text

data/sources/sinaiticus/01.txt

Commit 4c0e9f94117ec3dc4ae40094aec044bb7a416a53; SHA-256 A0812404A5EF2904A06F91F3A535CA30712F8BC0EE5F5388373C440964BE2C5F · CC BY-SA 4.0 · source record ↗

Henry T. Anderson, The New Testament translated from the Sinaitic Manuscript (1918)gloss

data/sources/sinaiticus-english/anderson-1918

208 hash-recorded responses from the Codex Sinaiticus Project public translation endpoint · Public domain · source record ↗

Codex Sinaiticus Projectverification

External manuscript images and project transcription

Independent visual control remains available by quire, folio, side, column, and line · CC BY-NC 4.0 · source record ↗

Governing rules

  • Display the pinned GA 01 source, never a critical-edition proxy.
  • Select the base reading: original scribe including recorded first-scribe corrections; exclude later a-c correctors from display.
  • Preserve damage, supplied text, abbreviations, correction layers, textual omissions, and comparison gaps as distinct states.
  • Use Anderson 1918 verbatim; the system assigns published words but never translates.
  • Assign zero to three consecutive English words to each extant Greek parent with deterministic monotonic ordering and fixed tie-breaking.
  • Suppress Anderson wording with no extant GA 01 parent rather than masking a manuscript omission.
  • Keep every decision inside its canonical verse.

Not permitted

  • Westcott-Hort or another critical text presented as Sinaiticus
  • Combining later corrector hands without attribution
  • Generated or AI-translated English
  • English-only display rows that conceal a GA 01 omission
  • Cross-verse spans
  • Merged or continuation cells
  • Arrow glyphs

Next certification action: Retain the pinned source and rerun the source and Anderson certificates whenever Sinaiticus cells, source files, or alignment rules change.

Position 4 · Latin Vulgate received tradition

Vulgate

Source verified

What the column displays

Clementine Vulgate, not the Stuttgart critical Vulgate.

Dates

Jerome’s Gospel revision began about 383 CE. Displayed recension: 1592/1598.

Current coverage

All four Gospels: 59,029/59,029 Latin source tokens across 3,779 local Gospel verse divisions; 82,900/82,900 Douay-Rheims words across 3,776 admitted source units; 32,288 displayed English row/span objects, including 14,910 multi-Latin phrase spans.

Status finding

All 59,029 Clementine Gospel tokens are displayed once in source order under a hash-pinned local source audit. All 3,776 selected Douay-Rheims 1899 translation units are source-admitted and internally aligned as 32,288 ordered row or phrase-span objects. Every Latin token and every published English word is accounted for once. This is internal source-constrained alignment, not a claim that the Douay translators published an interlinear and not independent scholarly review.

Immediate source material

Biblia Sacra juxta Vulgatam Clementinamtext

data/sources/vulgate/VulgClementine.txt

1592/1598 received edition; local SHA-256 F2BCC2BF6C7CCEC7258AE096A200F9C685B783A9E0A656232365792EBEC028AC; upstream revision still to be identified · Public domain · source record ↗

Douay-Rheims American Edition (1899; displayed published translation units)verification

data/sources/vulgate-english/challoner-1899/*.usfm; data/sources/vulgate-english/admitted-units.json

eBible engDRA source files dated 2022-11-03; 3,776 internally admitted units, ledger generated 2026-08-19 · Public domain · source record ↗

Original Rheims New Testament (1582; secondary translation witness—not yet displayed)verification

data/sources/vulgate-english/rheims-1582/*.json

janvier-s structured transcription acquired 2026-08-18 · CC0 1.0 · source record ↗

Whitaker's Words (lexical aid only)gloss

data/sources/glosses/whitaker/DICTLINE.GEN

Local acquisition is not yet pinned · Public domain

Lewis and Short, A Latin Dictionary (Perseus TEI; corroborating lexical evidence)verification

data/sources/glosses/lewis-short/vulgate-gospels-evidence.json

PerseusDL/lexica commit 40038e40937fa639639802e73dac15e6c938496b; scoped Gospel extraction · Public domain · source record ↗

Governing rules

  • Retain source-token order.
  • Give additional Latin words their own alignment rows.
  • Use empty cells only where Latin has no corresponding word.
  • Treat Whitaker as a lexical aid requiring contextual review.
  • Keep published English in source-supported translation units; Urevangelium aligns but does not translate.

Not permitted

  • Calling this text Weber–Gryson/Stuttgart
  • Dating the displayed edition to 383 CE
  • Dropping excess source words
  • Presenting Whitaker lexical output as contextual translation
  • Subdividing published English more finely than its source supports
  • Silently harmonizing Challoner and Rheims

Next certification action: Obtain independent scholarly review of the Vulgate source audit, whole-unit English alignment method, and edition-specific adjudication ledger.

Position 5 · Western bilingual manuscript witness

Bezae

Source verified

What the column displays

Codex Bezae Cantabrigiensis, GA 05 / VL 5, Greek and Latin sides.

Dates

The Western textual environment predates the codex. Codex: approximately 400 CE.

Current coverage

All 3,779 Gospel verse files classified: 59,984 Bezae text rows, 3,398 full physical-loss rows, 19 Greek-side loss rows, 94 full textual-omission rows, 236 Greek-side omission rows, 6,425 explicit comparison gaps, and 1,780 unpopulated post-generation display gaps.

Status finding

The site display is internally certified against hash-pinned ITSEE/IGNTP TEI files. Every one of 48,920 visible Greek forms and 52,749 visible Latin forms consumes a unique exact occurrence in its corresponding TEI verse, with zero unsupported or reused occurrences. This is display-scope certification; unused apparatus layers remain outside the claim.

Immediate source material

ITSEE/IGNTP Codex Bezae TEI transcriptionstext

data/sources/bezae/Bezae-Greek.xml; Bezae-Latin.xml

Display certificate pins canonical-text SHA-256: Greek 494725684D6211ACDA3D4ABCB147054CC54C26E22D76B3552170EEA94E8B0256 and Latin E8A392C42938C710129DA854697F80814558BA6A838786297075C3417B27036A · CC BY-NC-SA 3.0 · source record ↗

Governing rules

  • Keep Greek and Latin sides distinct within the same codex witness.
  • Require each displayed form to consume one unique occurrence in the corresponding TEI verse.
  • Treat a shared row as comparative placement, not literal bilingual equivalence or diplomatic sequence.
  • Represent full and side-specific physical loss separately from alignment and unpopulated display gaps.
  • Within an apparatus element, use the first encoded rdg as the displayed base-reading stream while retaining alternatives in the TEI.

Not permitted

  • Filling a physical lacuna from another tradition
  • Treating blank post-generation rows as manuscript omissions
  • Claiming comparison-row order is diplomatic sequence
  • Flattening unused apparatus readings into the certified display

Next certification action: Retain the pinned TEI files and rerun certify:bezae:display whenever Bezae cells or source files change.

Position 6 · Syriac Peshitta received tradition

Peshitta

Source verified

What the column displays

The scrollmapper electronic Syriac Peshitta text at pinned commit ba07bc991644d82b24426b920245eb4422daa769; its exact printed exemplar remains unestablished.

Dates

Peshitta Gospel tradition: approximately fourth to fifth century CE. Pinned electronic revision committed 2024-11-19; no BFBS/Urmia exemplar claim is made.

Current coverage

All 3,779 Gospel verse records; 50,477 Syriac source tokens; 84,133 Murdock English words; 50,464 populated Syriac parents; 13 unavoidable blank parents in ten verses where Murdock has fewer words than the Syriac source; 15 explicit English-only expansion words where the three-word parent capacity is exhausted; 2,726 governed shared-row decisions certified; zero cross-verse units, merged cells, continuation cells, arrow glyphs, absorbable expansions, or certification failures.

Status finding

All four Gospels are exact and occurrence-complete against the hash-pinned electronic source: 50,477/50,477 Syriac tokens. Every one of the 84,133 admitted Murdock words is preserved verbatim and accounted for once. Deterministic parent assignment uses SEDRA IV headwords, pinned ETCBC/SyrNT morphology, independent witness-row evidence, and source-ordered Murdock phrase context with fixed tie-breaking. No source token receives more than three Murdock words. This is internal source and process certification, not independent Syriacist review.

Immediate source material

scrollmapper Peshitta.txttext

data/sources/peshitta/Peshitta.txt

Commit ba07bc991644d82b24426b920245eb4422daa769; SHA-256 6E6E13089148E2D9809103F4B0BBB602D95086C28B37F44B086E800C5690651B · Public domain · source record ↗

ETCBC/syrnt Text-Fabric morphologyalignment

data/sources/peshitta/etcbc-syrnt/tf/0.1

Commit dae3eb6ff62b9b272fb503646796c25d248175ce · MIT · source record ↗

SEDRA IV lexical evidence, Beth Marduthoalignment

data/sources/peshitta/sedra-inserted-token-evidence.json

API responses retained with per-form analyses and source URLs · External lexical evidence; attribution retained · source record ↗

Payne Smith, A Compendious Syriac Dictionarygloss

data/sources/peshitta/payne-smith-proof-verses.tsv

1903 reference; proof-verse evidence only and withheld after the source rebuild pending occurrence remapping · Public domain

Etheridge (1846) and Murdock (1851) Peshitta translationsverification

data/sources/peshitta/etheridge.txt; data/sources/peshitta/murdock.txt

Verse-level contextual witnesses; proportional word extraction prohibited · Public domain

Murdock 1851 Gospel translation unitsgloss

data/sources/peshitta/murdock-gospels.json; data/sources/peshitta/murdock-admitted-units.json

Two-transcription collation acquired 2026-08-21; certificate hashes recorded locally · Public domain

Governing rules

  • Preserve every Syriac source token exactly once, with source-file hash, source occurrence, and stable row identity.
  • Preserve RTL display while allowing a governed shared-row move to expose meaningful cross-tradition correspondence; the move never changes the stored Syriac token sequence.
  • Keep every governing Murdock verse independent and preserve every admitted word verbatim.
  • Assign zero to three contiguous Murdock words to each Syriac parent using fixed evidence priorities and stable tie-breaking.
  • Use SEDRA IV only for atomic lexical headwords and ETCBC/SyrNT for morphology. Cross-tradition placement evidence requires two independent witness families among Greek, Latin, and Coptic; dependent Greek columns count as one family. Murdock remains the sole displayed English authority.
  • Use an explicit English-only expansion only when all honest adjacent parents are at the three-word ceiling; leave a Syriac parent blank only when the published Murdock unit is mathematically shorter than the source-token inventory.
  • Rerun source, English, row-parent, and governed-alignment certificates after any corpus change.

Not permitted

  • Claiming BFBS 1905 without an exemplar chain
  • Treating statistical support or phrase context as dictionary equivalence
  • Using proportional English distribution or AI-generated translation
  • Borrowing displayed wording from another tradition
  • Normalizing samekh forms in the archival display
  • Dropping or duplicating source or English words
  • Merged cells, continuation cells, arrows, or cross-verse English spans

Next certification action: Seek independent Syriacist review of the 2,726 governed shared-row decisions and the 28 mathematically explicit edge cases while preserving the pinned source and deterministic certificates.

Position 7 · Byzantine Greek textual tradition

Byzantine

Source verified

What the column displays

The Robinson–Pierpont 2018 Byzantine Textform electronic edition, not one medieval manuscript.

Dates

Byzantine-type readings emerge earlier; the mature tradition spans later centuries. Displayed edition: RP2018, byztxt v3.3.2.

Current coverage

3,778 RP2018 Gospel verse records and 66,130 tokens; 66,130/66,130 English placements admitted; Luke 17:36 is explicitly omitted because RP2018 has no verse record there.

Status finding

All four Gospels are rebuilt exclusively from the hash-pinned RP2018 v3.3.2 CSVs. Every one of 66,130 source tokens is displayed once in source order, and every token has English admitted by the Byzantine-specific evidence chain. Direct contextual TAGNT English is distinguished from orange project-adjudicated lexical output. This is internal source and process certification, not independent scholarly review.

Immediate source material

Robinson–Pierpont 2018 Byzantine Textform via byztxttext

data/sources/byzantine/{MAT,MAR,LUK,JOH}.csv

v3.3.2; commit 27a45ff1b7be6c17ccbfeac414f3f55732ae8e28; per-file SHA-256 hashes recorded in the certification ledger · Unlicense / public domain · source record ↗

STEPBible TAGNTalignment

data/sources/greek-shared/TAGNT-Mat-Jhn-CC-BY.txt

Local acquisition is not yet pinned · CC BY 4.0

STEPBible TBESG / Abbott-Smithgloss

data/sources/greek-shared/TBESG-CC-BY.txt

SHA-256 recorded in docs/audits/byzantine-english-shadow.json · CC BY 4.0

MorphGNT, PROIEL, and MorphGNT lexiconverification

data/sources/greek-shared/{morphgnt,proiel,morphgnt-lexicon}

Per-file SHA-256 hashes recorded in the English certification ledger · Source licenses retained in each local source directory

Governing rules

  • Use the pinned Byzantine CSV as the sole text authority.
  • Admit contextual TAGNT English only for an explicitly Byzantine-aligned token with RP2018 identity evidence.
  • Display project-adjudicated TBESG or MorphGNT-lexicon output in orange.
  • Require exact RP2018 surface and morphology identity before any English admission.
  • Treat the column as an edited textform rather than a physical witness.
  • Represent a verse absent from RP2018 as omitted rather than filling it from another Greek tradition.

Not permitted

  • Silent generic TAGNT fallback
  • Calling the edited textform one physical manuscript
  • Applying physical-manuscript lacuna rules
  • Borrowing readings or English from another Greek column
  • Displaying unresolved generated English as source translation

Next certification action: Seek independent specialist review of the published ledger and adjudication rules; preserve the current pinned source versions until a separately audited migration.

Rules shared by every column

  • Every displayed text token must trace to a permitted text source for that column.
  • Alignment tools and lexicons may place or explain words; they may not silently become witness text.
  • Physical loss, canonical absence, alignment emptiness, editorial supply, and unavailable transcription are different states.
  • Normalization must be reversible and documented; the source form remains the archival authority.
  • Draft computational alignment is labeled as draft until reviewed by a qualified reader.