Tech House Church — the evidence
techhousechurch.org · open working draft

The Jesus Evidence
Graph

One hundred and twenty-five teachings, weighed against every witness that carries them — and then discounted for the witnesses that were only copying each other.

125teaching units
366attestations
17witness works
107works cited
245citation links
24redaction pairs
11version readings
0page-verified cites
The rule

Abundance is not evidence

Four gospels saying the same thing is not four votes. Matthew and Luke both had Mark open in front of them. A ninth-century manuscript is not an independent memory. A Syriac version is a control on wording, not a witness to what Jesus said. Fifty English translations are one translation, fifty times.

So each teaching gets one score, and it is built by subtraction. Start from how many sources carry it, then take away everything that isn’t independent: material copied from an earlier gospel, later reception, translations used as controls, unstable text. What survives is the number.

The graph

Every teaching, wired to its witnesses

Three ways in. The graph shows how teachings connect to the works that carry them. The list is the same set, ranked and searchable. Threads is different in kind — it asks how the teachings and the events hold together as a body of thought, and where they don’t. Switch with the control top right. Click anything to open everything the database holds on that teaching — where it appears, which streams carry it, how its score was computed, and what argues against it. Each small node is a teaching; each ring is a work that carries it. Size is how many connections a node has — for a teaching, how many works carry it; for a work, how many teachings it holds. Colour is the evidence score. Those two deliberately no longer agree: a fat node is a well-connected one, not a well-attested one, and telling those apart is the whole argument of this page. Hover or tap to read it, click to open it fully.

Aramaic anchors are a separate class, added in this pass: the twelve places where an Aramaic word survives untranslated inside the Greek — Talitha koum, Ephphatha, Abba, Korban. They are evidence of an Aramaic-speaking layer in the tradition, so they are kept out of the teaching ranking entirely. Their score means “a Semitic element is genuinely preserved here”, never “Jesus said this.”

Lane
9–10 strongest 7–8 strong 6 moderate 4–5 weak 1–3 very weak Witness work Size = number of connections · colour = evidence score · edge weight = independence of that witness
Every place the corpus argues with itself. These are the same tensions that run through the threads, pulled out and ranked, because they are the hardest thing on this page to find and the most useful thing on it to read. Nothing here is a claim that the tradition is incoherent — only that these pairs have not been reconciled, by us or by the sources.

The arithmetic can say where a contradiction would score high — both sides well attested, sitting close in the text — but not whether one is there. The last search of that kind named 27 candidate pairs and 8 turned out to hold a real clash. The other 19 were teachings that simply sit near each other, and filling this view with them would have been the easiest possible way to make the corpus look more broken than it is.
Show at least everything
doubt one side and it goes away hard to escape
This view is interpretation, not evidence. Everything else on this page is a claim about what the sources say. This is a claim about what the teachings and the events mean together — which one illustrates another, which one only makes sense if another is true, and which ones pull against each other. None of it touches any score. Read it as an argument you are invited to disagree with. Each link has to say what new idea it brings; one that only restates its parent does not belong on the page, so brings in is the toll every level pays. The gold links are the exception in kind: an untranslated Aramaic word sitting inside a teaching is a fact about how the text was carried, not a reading of it. Which teaching it belongs with is still a judgement, which is why it lives here.

Indentation is local. A nested row relates to the row directly above it, not to the teaching at the top of the thread. Follow four levels down and you may be a long way from where you started — the graph contains paths twenty steps long, and a long path through a connected graph is a tour, not an argument. That is why threads stop at three levels here.
what follows from it what contradicts it an Aramaic word it still carries
Creative association 4
0 · the text put these together 10 · we did
Start from

Hover a node to inspect it.

The ladder

What survives the discount

One number per teaching: the evidence score. It is worked out by counting how many independent sources carry the teaching, discounting the ones that were copying each other, and subtracting for bundled units, unstable text, and later theological expansion.

Nothing here is a guess. Every score can be taken apart into the evidence that produced it — click any row and you get the whole derivation, witness by witness.

Out of 10, and nothing reaches 10. The best-attested teaching in the corpus scores 9. That is not modesty — it is what the evidence supports, and a 10 would have to mean something no first-century saying can claim.

And the five colours are softer than they look. A check built for the contradiction scale was turned on this one, and the result belongs here rather than in a working file. This database records, for every teaching, the gap between what the formula computes and what a person scored by hand — and the median gap is 0.98 points out of 10. Move every score by that much and 89% of teachings land in a different colour band. At three bands it would be 60%; at two, 42%. So the bands are a reading aid, not a finding: trust the number and the derivation under it, and treat the colour as a rough sort. Whether to paint fewer bands is a decision about this page, not about the evidence, and it has not been made yet.

The scale is deliberately coarse. Across 125 teachings the formula produces only 37 distinct values, so a finer number would be inventing precision. It is also absolute: a teaching’s score never changes because some other teaching was added or removed. And it is not a probability — a 9 does not mean a 90% chance he said it. It means the evidence is about as strong as evidence here gets.

Teachings sharing a number are tied. Six of these score 9, and there is no honest way to rank one above another — the evidence does not separate them. The bar is drawn from the finer underlying figure so you can see which way it leans, but do not read much into a difference the number itself refuses to make.

The discount

Where confidence drains away

The teachings that lose the most. Nearly all of them are the beloved parables that survive in one stream — Luke's special material. They are not less true; they are less independently attested, and the graph refuses to pretend otherwise.

Raw attestation Reliability-weighted
The redaction

What the later writers actually did

Until now the discount rested on a general principle: Matthew and Luke had Mark open, so their agreement isn’t independent. This pass checked it case by case — twenty-four pairs, each asking what the later witness demonstrably does to the earlier one.

Fourteen came out likely dependent. Exactly one came out likely independent. That is the discount, measured rather than assumed. No source text is stored anywhere in this database — these observations describe the changes rather than reproducing the texts.

The streams

Who is actually testifying

Attestation counts by work, with the dates the database holds. Matthew and Luke lead on volume — which is exactly why volume alone can't be trusted: most of it is Mark, re-told.

WorkDate (CE)StatusAttestationsMean independence
The hole

What wasn’t here at all

Ten cycles asked whether the entries were correct. None asked whether they were complete. They weren’t: the Golden Rule was missing entirely from a database of 123 teachings, and so was the leaven — the paired opposite of the mustard seed, which was present and ranked.

The hole is systematic, not random. It is almost entirely Luke’s special parables: the lost coin, the rich fool, the persistent widow, the barren fig tree, the two debtors, counting the cost. Because Luke-only material scores low under the canonical formula, none of this changes the top of the ladder — which is exactly why nobody noticed. A database can be silently incomplete in the places its own scoring cares least about.

The Golden Rule and the leaven are now in, both landing mid-ladder (#47 and #52) on two streams each. The remaining eight are recorded rather than invented in bulk at the end of a push.

The objection

Attested, or just widespread?

Here is the strongest argument against the ladder above, and the formula cannot see it: multiple attestation may record a saying’s circulation, not its origin. Proverbs and apocalyptic commonplaces travel on their own and collect famous attributions along the way. A saying in Mark, Q, Paul and Thomas might be everywhere because it was common property.

It bites hardest exactly where it hurts most. Four of the canonical top ten carry this risk — including the top three.

Nothing here is scored. I went looking for the scholarship — the rabbinic “uprooter of mountains” parallel, the Greek proverb behind “a prophet is not without honour” — and the search returned devotional writing, not scholarship. Under this project’s own rule an unverified claim may not move a number, so every row below has affects_score = 0. The register is a queue of open questions, not a verdict.

Each row separates two things that are easy to blur: what the primary text actually shows, which is checkable now — and the inference drawn from it, which is contested and needs a named source.

The audit

Where the number comes from

The first scores in this project were guesses — numbers typed in from a quick read, before any evidence was counted. They have been removed from every ranking and display. They are kept in the database as a record of how the project started, and nothing else.

What replaced them is a written formula over stored evidence only: streams weighted by quality, diminishing returns on later witnesses, penalties for bundling and unstable text. You can ask why any teaching scores what it does, and the database answers.

The first version failed its own audit, which is the whole argument for writing it down. It upgraded the longer ending of Mark from 0.22 to 0.44, because the label “Late Mark ending” matched mark before it matched late. A known late addition got promoted into the earliest stream by a string-matching bug. It now scores 0.14. A number typed in by hand could never have exposed that, because there was nothing to inspect.

Then the method’s fifth criterion — late legendary expansion — was encoded as 27 cited flags, and the first attempt made things worse. Seven teachings were driven to the very bottom of the scale, and all seven had already been penalised for having no independent source. The same fact was being counted against them twice, and the bottom of the ladder collapsed into a single value where real differences had been.

Scaling the penalty instead of subtracting it fixed that. The failure is kept on this page because a formula that can be inspected can be caught being wrong — which is the entire argument for writing one down.

FormulaChangeflattened to the floordistinct values, bottom 20

The flags themselves, each cited, six of them carrying a source that argues the other way:

Three scores was two too many

Auditing this turned up something worse than a stale number. Three scoring systems had accumulated — the original guessed numbers, an independence weighting buried in an export script, and this formula — and they produced three different top-six lists with no teaching common to all three. Asked which teachings are most probably his, the database gave three answers depending on which column you read.

One is now canonical. Nothing was deleted.

125

scores, all derived

Every one recomputable from stored evidence. None estimated.

2

of five criteria encoded

Attestation and late expansion. Coherence, contextual fit and genre are not in the formula.

3

formula versions, in public

Including the one that got worse. The failures are on this page, not tidied away.

Be clear about what this measures. Attestation and late expansion — roughly two of the method’s five criteria. Coherence with the earliest layer, fit with first-century Jewish Palestine, and genre are not encoded, because no honest way to compute them from stored evidence has been found yet. Until there is, they stay out rather than being faked.

So this is a ranking by how well a teaching is attested, not a verdict on what he certainly said. The difference matters and the page will not blur it.

One word

Abba

Of everything here, one Aramaic word carries more evidential weight than any parable. Not because it says the most — because it survived a journey that should have destroyed it. Greek-speaking congregations in Rome and Galatia, people who knew no Aramaic, kept praying an Aramaic word and then immediately translating it. You don’t do that with vocabulary. You do that with something you were given.

It appears three times in the entire New Testament, always as the same fixed formula, always in prayer:

Mark 14:36JesusGethsemane Romans 8:15believersprayer of adoption Galatians 4:6the Spiritcrying in believers’ hearts

Looking at it closely cost us something. In cycle 1 this page called Mark and Paul two independent streams. That was overstated — an identical fixed formula in two authors is better explained by one inherited liturgical unit than by two separate memories. The anchor’s score came down from 0.95 to 0.90, and the correction is stored in the database rather than written over the old claim.

Two other things long said on the strength of this word are simply wrong. It does not mean Daddy. And Jesus was not the first to pray this way.

What survives is narrower than the popular claim and, I think, better: strong evidence that the earliest Jesus movement prayed in Aramaic using a form its Greek-speaking members refused to let go of, and good — if inferential — evidence that the form goes back to him.

The bundles

The number at the top of the ladder was wrong

Attestation counts works per unit. When a unit bundles several different sayings, witnesses to different sayings get summed — and the unit displays a strength no single saying in it possesses.

Wealth is dangerous scores 0.92 on five works. But Mark’s rich young man, Matthew’s treasures-and-two-masters, and Luke’s rich fool are not parallels. They are different sayings. Nothing in that unit is attested five times.

This inflates twice over: apparent agreement, between witnesses that were never discussing the same saying, and apparent independence, from streams that never independently attested anything.

works summed across the nine highest-scoring units genuinely independent streams, once bundling and Markan copying are removed

Seven of the nine come down to a single independent stream. That is not a reason to disbelieve them — one early stream is still evidence, and Allison’s counter-argument is recorded alongside: a recurring theme may be better attested than any one saying. But it is the difference between “five witnesses agree” and “one source, copied four times.”

One unit is a category error rather than a bundle. Parables as Jesus’ characteristic speech cites whole chapters and is a claim about the form of his teaching, not a teaching whose wording can be compared. It has been flagged, not moved — reclassifying it changes the headline ranking, and that is a call for Wenzl, not for me.

The versions

Watching a prayer grow

Syriac earns its own pass for one reason: it is the only Aramaic-family translation tradition. For a Syriac scribe the Aramaic words preserved in the Greek — Talitha koum, Abba, Korban — aren’t foreign vocabulary. They are native. No Latin or Coptic witness can tell us anything about that.

But the sharpest thing these witnesses show is the ending of the Lord’s Prayer, caught mid-growth:

Didache… for yours is the power and the glory forever Old Syriac (Curetonian)… for yours is the kingdom and the glory, forever, Amen Peshitta… for yours is the kingdom and the power and the glory

Three witnesses, three endings, each fuller than the last — and the familiar one is absent from the earliest Greek manuscripts entirely. Nobody was lying. Liturgy accretes as congregations say it aloud.

This is the project’s whole thesis in a single variant. A reading can be ancient, widespread, beloved, and still be an addition. Counting witnesses would have made this more certain, not less.

Variant siteWitnessDateDoesReading

The apparatus

Every claim, traced to a source

A hundred and four works, chosen for what they license rather than for weight of name. Each carries two fields that matter more than its title: what it may legitimately be cited for, and where it is dated, contested, or routinely over-read.

No page numbers are recorded anywhere in this database. A page citation requires the book actually open, and none of these have been. Recording one from memory would be the exact failure this apparatus exists to prevent.

The discounts

Even the rules are cited

Every weight this project applies is a scholarly judgement. If the weights can't be traced, the ranking is just an opinion with decimals attached. Five of these twelve rules carry a documented counter-position — including one that would break the model if it holds.

The honest part

What isn't built yet

Part of this list has since been filled in — the source text went in, and you can now read it on the page. What is left is mostly the part no machine can do for us. Until it is done, the graph can tell you a teaching ranks strongly — but not let you audit, down to a scholar and a page, why.

354

passages now carry text

Of 366 — done since this list was written. Public-domain English throughout, with the Nestle 1904 Greek on 309 of them. Twelve remain textless, two of those correctly: Q is a reconstruction and has no manuscript to quote.

0

page-verified citations

All 245 citations point at works, none at pages. Turning them into page-level references needs the books physically opened. It cannot be done from memory, and that is the point.

0

modern_translations

No reader-facing comparison layer. Deliberate: rights posture has to be explicit before anything goes in.

12

Syriac questions open

Twelve substrate-control questions queued against Kiraz’s Comparative Edition. The sharpest: if the children/stones pun is real, a Syriac translator should reproduce it automatically.

Next

The moves, in order of impact

From working/next-database-actions-ranked.md. The order is deliberate: nothing downstream is trustworthy until the citation backbone exists.

Build the per-witness wording tables

Right now the database knows which witnesses carry a teaching, not what they actually say. Until then, dependency discounting is an estimate rather than a measurement.

Pilot on two teachings

Sabbath for human good and Kingdom/reign of God is near — one mixes close Markan wording with Lukan expansion, the other is the conceptual spine of the whole book.

Split the broad units into comparable pieces

"Wealth is dangerous" bundles several separate sayings. Bundles manufacture fake agreement and fake independence at the same time.

Move the graph relationships into the database

So the visual map is a product of the data, not a one-off export sitting beside it.

Verify page citations, one work at a time

The apparatus currently cites works. Turning those into page-level citations is what would make this genuinely audit-grade — and it cannot be done from memory.

Open

The one job that needs a person

This is a working draft, not a finished thing, and it is meant to be worked on by more than one person. There is a specific job it needs, and it is the one job the machine that built this cannot do.

Every citation here names a book. None names a page. All 245 of them. If you wanted to check whether Barr has been represented fairly, you would have to read his whole article to find the sentence we are leaning on.

Nobody has opened these books. Writing a page number from memory would look more scholarly and would be a fabrication — so the field was left empty and the count published instead. It has read 0 since the first day.

Adopt a citation. If you have a university library card, JSTOR, or an Internet Archive account, you can close a piece of this in about thirty seconds per entry: find the passage, note the page, send it back. Twenty citations from one person with access would move this further than another week of machine work.

Verify a page → 11 works, the full worklist is live Start with the ones that license a confidence upgrade: Barr 1988 (JTS 39, 28–47), Schelbert 2011, Fitzmyer 1979, Williams 2004.
Answer a Syriac question 12 open, needs Kiraz’s Comparative Edition The sharpest: if the children/stones pun is real, a native Aramaic-family translator should reproduce it automatically. If the Syriac shows none, that argument weakens badly.
Source an objection 8 rows still unsourced The strongest argument against this ladder is that multiple attestation may measure a saying’s circulation, not its origin. It needs real specialist sources.
Source one of the nine disputed links 9 assert a dispute, 0 carry a citation Nine links here say a point is argued over among scholars, and none names who argues it. Until this week there was no way to fix that — the database had no path for a source to attach to one of these links. It has one now and it is empty. If you know the literature on the two resurrection strands, or on whether the vineyard parable’s ending is later allegory, that is one citation and it closes one row.
Judge whether a thread link earns its place 236 links, none ever deleted Machines can check a link’s form — that it does not restate its own reasoning, or its siblings, or name a chapter that is not there. All three of those checks run here and all three come back clean. Whether a link is worth keeping is a different question, and four attempts to automate it produced false positives rather than findings. It needs somebody willing to say “this one is thin”.
Argue with it every claim stores its counter-argument If something here is wrong, the fix is a source, not an opinion. Contributions follow the same rules the project follows.

The rules are short and they are not negotiable, because they are the only reason any of this is worth reading: no page cited from memory. Ancient translations control wording, they never vote. Where scholarship is contested, the opposing source is stored beside the claim.

Database christian-scriptures.sqlite · integrity ok Canonical score formula v1.3 · recomputable from stored evidence Open working draft · last updated 2026-09-07 No book drafted — method and database only 18 of 106 sources independently verified