Well, maybe we haven't. Um... The Library — Why We Have the Manuscripts
The Library  ·  a supplemental page

Why we have
the manuscripts

One million and thirty-four thousand pages of the oldest writing humanity managed to keep. Nineteen collections. Held, not borrowed.
1,034,177
manuscript page images held in our own vault
measured 2026-08-15 · 19 shelves
19
separate collections, from five traditions
each shelf sized individually
1.53TB
of page imagery — the manuscripts alone
1,530,832,097,044 bytes
20,760/hr
still arriving, right now, around the clock
65 GB per hour · measured live

What this actually is

Not a subscription. Not an index of things held elsewhere. The page images themselves, on our own shelves, ours to read, ours to keep, and ours to teach a machine with.
The scale, made real

A person could not read it in a lifetime of trying

A million pages is a number nobody can feel. So here it is in hours: reading one page every minute, eight hours a day, without a day off — turning every page in this library exactly once takes nearly six years.

That is one pass. No study, no cross-referencing, no translation. Just turning the pages.

5.9 years
of continuous eight-hour days to turn every page once
1,034,177 pages ÷ 60/hr ÷ 8 hr/day = 2,154 days ÷ 365
Two-thirds of
the Eiffel Tower
if the leaves were stacked as physical folios
1,034,177 × 0.2 mm ≈ 207 m · Eiffel Tower = 330 m
0.2 mm = typical parchment leaf — stated assumption, not a measurement
Why it is ours and not rented

Nobody can take it back

Every one of these pages sits in storage we control. Not an API key that expires. Not a licence that gets revoked. Not a viewer that a library can switch off. If every institution on this list closed its doors tomorrow, the pages would still be here.

That mattered more than it sounds. Institutions restrict, reorganise and retire their digital collections constantly, and the terms change without notice. What is downloaded is settled. What is merely linked is on loan.

The shelves

Nineteen collections, grouped by the tradition they come out of. Every count below was measured directly from the vault, shelf by shelf, not estimated.

The Hebrew and Jewish world

349,319 pages  ·  177 GB
Overwhelmingly the Cairo Genizah — the storeroom of a Cairo synagogue where, for roughly a thousand years, nothing bearing the name of God was thrown away. Letters, contracts, children's writing exercises, medical notes, and scripture, all preserved by accident of reverence.
GenizahThe Cairo Genizah core — everyday documents and sacred text from the 9th century forward321,461125.4 GB
Cambridge GenizahThe Taylor–Schechter holdings at Cambridge12,4824.2 GB
JTS GenizahJewish Theological Seminary Genizah fragments11,10145.9 GB
Hebrew manuscriptsHebrew codices outside the Genizah collections4,2751.4 GB

The Christian west

504,704 pages  ·  1.29 TB
The monastic and cathedral libraries of Europe — the institutions that copied scripture and classical learning by hand, generation after generation, and are the reason any of it survived at all.
Vatican LibraryThe Vatican Apostolic Library — one of the oldest libraries on earth185,433179.5 GB
e-codicesThe medieval manuscripts of Switzerland — the highest-resolution imagery we hold126,411699.8 GB
Bavarian State LibraryMunich — one of Europe's great manuscript holdings70,159116.8 GB
Parker LibraryCorpus Christi, Cambridge — the Anglo-Saxon survival, including some of the oldest books in England45,063201.8 GB
BodleianOxford32,09673.6 GB
TreasuresCurated flagship items across institutions22,8907.5 GB
CambridgeCambridge University Library, general manuscript holdings18,0534.9 GB
WaltersThe Walters Art Museum — illuminated manuscripts4,5993.0 GB

The east

152,035 pages  ·  47.6 GB
The traditions that were writing, and preserving what they wrote, at the same time as everyone else — and which almost every western collection leaves out.
JapaneseJapanese manuscript and early printed holdings57,62924.8 GB
SanskritThe Cambridge Sanskrit collection — Indian religious and scientific texts53,19510.4 GB
ChineseChinese manuscript holdings24,9697.3 GB
IslamicArabic and Persian manuscripts — science, medicine, philosophy, scripture16,2425.1 GB

Science and medicine

16,679 pages  ·  4.8 GB
Small in page count, and the shelf most people would not expect us to hold.
Newton papersIsaac Newton's own manuscripts — the notebooks, not the published books16,3704.5 GB
Historical medicineEarly medical texts3090.3 GB

The American record

11,440 pages  ·  14.6 GB
Library of CongressManuscript holdings from the Library of Congress11,44014.6 GB

Why we went and got them

Four reasons, and they compound. None of them is "because it is interesting."
One — the source, not the translation

Every answer traces back to what was actually written

Every layer between a reader and an original text is a place where a human decision got made — a translation choice, a doctrinal preference, an editorial cut. Most people never see those layers, because they only ever meet the finished product.

Holding the original manuscripts means a question can be taken all the way down. Not to what a commentary says the text says. To the page.

God said put me at the center — the original manuscripts without man's interpretation.Carter · August 3
Two — it is built into the machine

The system cannot finish a piece of work without going back to them

This is not an intention. It is a gate. Every task the system completes has to pass a final check that ties it back to the original manuscripts before it can be called done — a mechanism that no session, and no person, can step around.

That gate was built because intentions decay and mechanisms do not.

Every single thing we do goes back to the original manuscripts and the construction of the Kingdom of God and His Spirit… we need a mechanism that is mechanical that you or none of the sessions can get beyond unless that's part of it.Carter · August 7
Three — what we teach it with

Anyone can train on the internet. Almost nobody can train on this.

The frontier systems are all fed roughly the same thing: the public web, scraped. They are converging, and they are starting to sound like each other, because they are eating each other's output.

A library of a million original pages that predate the internet by centuries is not a bigger pile of the same material. It is different material — and it is the part that has already been tested by time. Texts that survived a thousand years of copying survived because people kept deciding they were worth the labour.

If we just took our brain and we got all the books… everything that's ancient forward that we can get our hands on and the original manuscripts, and designing it according to the Kingdom of God.Carter · August 1
Four — because they are disappearing

Preservation is not a side effect. It was a motive.

Physical manuscripts degrade. Institutions lose funding, reorganise, restrict access, and quietly take digital collections offline. Material that is freely reachable this year is routinely gone the next.

Every page pulled into our own vault is a page that stops depending on somebody else's budget. The collection is ordered oldest-first for exactly that reason — the most ancient material is the most fragile and the least replaceable.

What it leads to

The library is not the destination. It is the fuel, and it is already in the tank.
The next step

A measure of outcome the manuscripts define

Money coming out the other side of this system is easy to count, and everybody counts it. The harder question is whether the work is any good by a standard that does not move — and that is what a library like this makes possible.

Not our opinion of whether something was worthwhile. A judgement traceable to the oldest written record humanity kept.

We can have a king to watch the kingdom outcome based upon the original manuscripts.Carter · S1501

Still arriving — stated plainly

The collection is not finished, and this page does not pretend it is.

  • 375,112 further page images are being re-acquired after the loss of the old server. At the current measured rate that is under a day of running.
  • Cambridge Sanskrit is mid-collection — roughly half its manifests fed so far.
  • Whole traditions are not represented yet — the Sinai monastery, the Oxyrhynchus papyri, and the great national libraries beyond those listed above.
  • Of everything that has been digitised on earth, what we hold is a sliver. That is the honest framing, and it is also the opportunity.