
From Online Resource to Reusable Dataset: The Dataset of Byzantine Book Epigrams
Abstract
This paper presents a structured, machine-readable dataset of approximately 13,000 Byzantine book epigrams, mostly dating from the 11th to 15th centuries and written in Byzantine Greek. The dataset is derived from the Database of Byzantine Book Epigrams, an ongoing Ghent University project that provides a web interface for exploring the epigrams but does not support large-scale exports or cross-referencing of metadata. Our pipeline extracts and normalizes information from Elasticsearch and PostgreSQL, stores it in SQLite, and exports the results to Zenodo. By structuring the data, it becomes easier to use in computational and AI-driven research and allows intuitive linking of metadata with information from other sources.
© 2026 Paulien Lemay, Klaas Bentein, Frederic Lamsens, Joren Six, Kristoffel Demoen, Els Lefever, published by Ubiquity Press
This work is licensed under the Creative Commons Attribution 4.0 License.