Skip to main content
MenuEtruscan

Open Datasets

The published corpus lives on Zenodo under a persistent DOI. Per-inscription TEI is served live from each inscription page. There are no nightly dump endpoints on this site.

Primary

Published corpus (Zenodo)

Versioned archive of the 5,932-inscription unified corpus plus software. Concept DOI always points at the latest deposit; each release also has a version-specific DOI for reproducibility.

Source

GitHub releases

Source code, tagged releases, and release notes. Pin a tag when you need the exact pipeline that produced a published snapshot.

Models

Hugging Face models

Model weights and derived artefacts published under the openEtruscan org: classifiers, embeddings, and evaluation bundles where available.

Per-inscription TEI

Every inscription page exports EpiDoc TEI XML on demand. Open an inscription, or call the export endpoint directly:

GET /api/inscription/{id}/tei
// application/tei+xml ยท EpiDoc subset

There is no bulk TEI zip on this site. For the full corpus as a single archive, use the Zenodo deposit above. For one record at a time, TEI is the supported live path.

Licence & citation

Corpus data (transliterations, normalisations, metadata, and, where available, translations from the source editions) is published under Creative Commons Attribution 4.0 (CC BY 4.0), the licence on the Zenodo deposit. You may copy, modify, and redistribute without asking permission, provided you credit the source.

Please cite the Zenodo concept DOI when you publish work that depends on this corpus. BibTeX and RIS blocks live on the citation page.

Full-corpus machine dumps (JSON/CSV/RDF) are not served from openetruscan.com. Serverless routes are the wrong place for a multi-megabyte nightly dump; Zenodo is the citable, versioned archive.