Reproducible Universal-Dependencies pipeline and XLM-R parser for Katharevousa Greek. Paper: arXiv:2605.22978.
-
Updated
May 30, 2026 - Python
Reproducible Universal-Dependencies pipeline and XLM-R parser for Katharevousa Greek. Paper: arXiv:2605.22978.
Computational vocabulary analysis across Buddhist Vajrayana, Shakta Tantra and Baul Bengali texts · TF-IDF char n-grams · cosine similarity · Tara to Kali lexical migration · Ramprasad Sen OCR · Charyapada · 75 texts across 8 tradition layers
ProMeTEXT develops corpora, methods, and tools for the segmentation and multilingual alignment of medieval texts, with a focus on 13th–16th-century romance traditions.
Old Icelandic spaCy POS-Tagger and Lemmatizer
A Gradio interface for exploring multilingual alignments of medieval textual traditions produced with Aquilign.
Experimental computational framework for semantic exploration of the Voynich Manuscript using transformer embeddings, medieval corpora, semantic clustering and digital humanities methodologies.
Computational stylistics of post-junta Greek parliamentary questions in Katharevousa (1976–1977). Code accompanying our DSH paper.
Add a description, image, and links to the historical-nlp topic page so that developers can more easily learn about it.
To associate your repository with the historical-nlp topic, visit your repo's landing page and select "manage topics."