Recover
Return to household-level records rather than relying only on published aggregates.
Digital history project
Tamuz uses handwritten text recognition and computational methods to return to the original household schedules of the 1897 Census of the Russian Empire. This makes it possible to reconsider published census data and analyse information that disappeared from the official statistical tables.
Learn about the projectThe published volumes of the 1897 census remain an essential source for historical research. Yet publication transformed millions of individual records into aggregated categories and statistical tables.
Tamuz works with the surviving handwritten census sheets themselves. By recognising, correcting, and structuring these records, we can test published figures, reconstruct how categories were produced, and study combinations of variables that were never included in the printed census.
Return to household-level records rather than relying only on published aggregates.
Analyse age, occupation, birthplace, language, religion, kinship, and residence together.
Compare original records with published statistics and question established interpretations.
The project trained and evaluated a handwritten text recognition model for Russian-language census schedules in Transkribus. The workflow combines model training with manual correction, layout-aware transcription, and Python-based data processing.
Published census tables answer questions defined by the statistical authorities of the late Russian Empire. The household schedules allow researchers to formulate new questions and combine variables at the level of people, families, and places.
Tamuz therefore treats digitisation not simply as access work, but as a method of historical reinterpretation.
Contact us about research collaboration, census collections, handwritten text recognition, data methods, presentations, or supporting the project.