HomeDatasetsKen-Z/Latin-Audio
L

Ken-Z/Latin-Audio

Text To Speech · Ken-Z· 1.8K
cc-by-4.0 15 GB

Vox Classica is a Latin speech corpus of ~73 hours of audio, segmented into short audio clips by sentence. Vox Classica is a large-scale, ML-ready dataset of human-read Classical Latin. It was designed to address the absence of a publicly available human-read Latin corpus large enough for model training.

Open in MLForge Sign up free Desktop app
# download instantly
mlforge datasets pull Ken-Z/Latin-Audio

Dataset details

Task
Text To Speech
Language
la
License
cc-by-4.0
Size
15 GB
Rows / images
23.9K
Creator
Ken-Z
Downloads
1.8K
Source
huggingface_datasets
Updated
2026-05-14

About Ken-Z/Latin-Audio

Vox Classica is a Latin speech corpus of ~73 hours of audio, segmented into short audio clips by sentence. Vox Classica is a large-scale, ML-ready dataset of human-read Classical Latin. It was designed to address the absence of a publicly available human-read Latin corpus large enough for model training.