DT DIGILABS PRODUCT
DT Cipher
Advanced OCR, translation, and lost-language recovery for historic text.
Unlocking the Words Within
DT Cipher unlocks the knowledge trapped inside historic documents, transforming handwriting, faded text, and difficult period typefaces into searchable digital content. Built for material conventional OCR struggles to read, it can transcribe, translate, and interpret complex historic scripts—making previously inaccessible text discoverable and usable while preserving the original document alongside it.
What It Does
- Next-generation OCR built for historic material — cursive handwriting, faded ink, esoteric period typefaces, and damaged or degraded documents
- Multi-language output — including translation and recovery of lost or endangered languages
- Controlled, topic-specific vocabularies built around your collection’s actual subject matter to improve accuracy
- Delivers searchable text alongside your images — in the formats your systems already use: PDF, PDF/A, METS/ALTO sidecar XML, and plain text
Built For
- Libraries and archives with historic correspondence and manuscripts
- Government and legal institutions with historic records
- Religious and cultural institutions with sacred or liturgical texts in historic scripts
Responsible AI – DT Cipher’s transcriptions and translations are AI-assisted first drafts, built to be reviewed by a qualified linguist or archivist before publication — especially for lost or endangered languages, where accuracy is never optional.
Pairs Well With
- Digitization & Capture — start with a fresh capture, then transcribe and translate
- DT Registrar — combine searchable text with object and subject metadata for full collection search
Featured Work
Cursive Hebrew Manuscript Recovery
DT DigiLabs applied DT Cipher to a collection of cursive Hebrew manuscripts — work that required more than OCR alone; it required a genuine understanding of a difficult, easily-misread historic script.