The Jawaharlal Nehru Archive
Where scholars proofread and enrich thousands of historical documents — correcting AI-processed text, refining metadata, and resolving the people, places, and institutions of Nehru's world — before they reach the public archive.
A guided tour
A six-minute walkthrough, from creating an account to exporting a finished corpus.
What the platform does
Every document begins as AI-processed text and TEI. Editors correct it against the original, and the platform preserves the AI's version untouched for a faithful round trip.
People, places, and institutions are tagged and linked to shared, Wikidata-backed records — merged, disambiguated, and reused across the entire corpus.
Claim, proofread, certify, and review — every change versioned, every issue tracked, and nothing certified while questions remain open.
The editor's desk
A focused editing surface built for careful, reversible scholarly work.
Claiming a document assigns it to you and locks it, so two researchers never overwrite one another.
Select any span of text and link it to a dictionary entry; the editor keeps every annotation aligned as the text changes.
Compare any two versions — including against the original AI output — and revert safely. The original is never lost.
“At the stroke of the midnight hour, when the world sleeps, India will awake to life and freedom.”
Built for the published archive
The finished corpus flows back out in exactly the shape the public Nehru Archive consumes.
Shared dictionaries keep names and identifiers consistent across thousands of documents.
Certified documents flow to a reviewer queue for a final scholarly check before release.
Proofread documents export in the exact JSON and TEI the public site consumes — unedited files byte-for-byte identical.
Help bring the papers of Jawaharlal Nehru to the world — one carefully edited document at a time.