Team Ai
Datasetpublic

reesjon9/Latin-Audio

Dataset Summary Vox Classica is a Latin speech corpus of ~73 hours of audio, segmented into short audio clips by sentence. Vox Classica is a large-scale, ML-ready dataset of human-read Classical Latin. It was designed to address the absence of a publicly available human-read Latin corpus large enough for model training. Alignment and curation: Kaiyuan Zhao Language: Latin (Classical) Uses This dataset is built for training and evaluating speech processing models… See the full description on the dataset page: https://huggingface.co/datasets/reesjon9/Latin-Audio.

sourceHugging Facecc-by-4.0updated 10mo agoView on Hugging Face
0likes124downloads
1 commits on main
b388c9010mo ago

Duplicate from Ken-Z/Latin-Audio

reesjon9, Ken-Z