Skip to main content
Have a personal or library account? Click to login
Verse, Variation, and Meaning: Annotating Similarity in Greek Marginal Texts Cover

Verse, Variation, and Meaning: Annotating Similarity in Greek Marginal Texts

Open Access
|Aug 2026

Abstract

This paper introduces a manually annotated dataset for verse-level semantic textual similarity in Byzantine Greek book epigrams. The benchmark captures graded similarity judgements as it was annotated by making use of pairwise comparison. It is furthermore designed to evaluate computational models of meaning in historical languages. Alongside the benchmark, we present a series of experiments that compare static word embeddings, contextual language models, and a semantically informed edit distance approach. Our results show that lemmatised CBOW embeddings perform slightly better than more complex transformer-based models, and that edit distance methods can be effectively adapted to capture semantic proximity. The study also highlights key challenges, including annotation asymmetry and the interpretative limits of partial data. By combining philological expertise with computational experimentation, this work lays the foundation for scalable similarity detection in Ancient Greek and contributes a reusable resource to both digital philology and ancient language processing.

DOI: https://doi.org/10.5334/johd.587 | Journal eISSN: 2059-481X
Language: English
Page range: 113 - 113
Submitted on: May 6, 2026
Accepted on: Jun 25, 2026
Published on: Aug 18, 2026
Published by: Ubiquity Press
In partnership with: Paradigm Publishing Services

© 2026 Colin Swaelens, Maxime Deforche, Ilse De Vos, Els Lefever, published by Ubiquity Press
This work is licensed under the Creative Commons Attribution 4.0 License.