CoolFace
Datasetpublic

Ken-Z/Latin-Audio

Dataset Summary Vox Classica is a Latin speech corpus of ~73 hours of audio, segmented into short audio clips by sentence. Vox Classica is a large-scale, ML-ready dataset of human-read Classical Latin. It was designed to address the absence of a publicly available human-read Latin corpus large enough for model training. Alignment and curation: Kaiyuan Zhao Language: Latin (Classical) Uses This dataset is built for training and evaluating speech processing models… See the full description on the dataset page: https://huggingface.co/datasets/Ken-Z/Latin-Audio.

sourceHugging Facecc-by-4.0updated 2mo agoView on Hugging Face
8likes3.2kdownloads

Ken-Z/Latin-Audio · main · files are served by the source, never re-hosted here