🔬 Impresso Datalab Notebooks
-
Updated
Sep 14, 2026 - Jupyter Notebook
🔬 Impresso Datalab Notebooks
A Toponym Resolution Pipeline for Digitised Historical Newspapers
This repository provides underlying code and materials for the paper 'Station to Station: Linking and Enriching Historical British Railway Data'.
An experimental infrastructure for computational critique of historical writing using RAG architectures, argumentation analysis, rhetorical mapping, and large language models.
Official repository of WikiTextGraph
Teaching materials for the workshop "Network Analysis in the Humanities" organised by DigiTS (Center for Digital Text Scholarship) at the University of Tartu.
Training classifier models to predict genres and subgenres on album cover data.
RISHI-Q — computational comparative history of physics. Flagship: Vaiśeṣika ākāśa–śabda sound-medium ontology vs Greek/Chinese/Buddhist controls & Maxwell EM (9/9 · 6/6 · 0/5 · R2 unique).
Training classifier models on acoustic metadata to predict genres and subgenres.
Scripts and archived outputs for the computational analysis in my essay on Miike Takashi's First Love (2019). No media redistributed.
Word–color association from large-scale online image data (CIELch). Code for the PLOS ONE manuscript.
Replication package for PHTS Theory v3.0 — a two-tier structural framework for the Voynich Manuscript. Submitted to Cryptologia, 2026. Contains Python scripts, data files, and pre-computed results.
A machine-checked formalization of the Fuxi 64-hexagram system. 40 claims verified; 7 corrected, including a clustering coefficient that is 0 not 5/12, and a divination kernel whose stationary distribution is not uniform.
Impresso Python Library to interact with the Impresso Public API
Information dynamics in Korean flash fiction — surprisal, coherence, and semantic-shift trajectories. Code for the Physica A manuscript.
5D semantic-dynamics framework separating Korean prose poetry from flash fiction. Code for the Scientific Reports manuscript.
Chunk-level FAISS + Solar-10.7B RAG pipeline over 2,900+ Korean flash fiction texts.
Computational vocabulary analysis across Buddhist Vajrayana, Shakta Tantra and Baul Bengali texts · TF-IDF char n-grams · cosine similarity · Tara to Kali lexical migration · Ramprasad Sen OCR · Charyapada · 75 texts across 8 tradition layers
Reproducible GHSA study of Fludd's 1617 Utriusque Cosmi Historia: a typed relational graph encodes the monochord's proportional grammar; six invariants are subjected to a seeded N=100 perturbation regime and a stability surface is measured. E0–E3 evidence never collapsed; the experiment is the authority.
To associate your repository with the computational-humanities topic, visit your repo's landing page and select "manage topics."