CoolFace
Datasetpublic

SparkAudio/voxbox

VoxBox This dataset is a curated collection of bilingual speech corpora annotated clean transcriptions and rich metadata incluing age, gender, and emotion. Dataset Structure . ├── audios/ │ └── aishell-3/ # Audio files (organised by sub-corpus) │ └── ... └── metadata/ ├── aishell-3.jsonl ├── casia.jsonl ├── commonvoice_cn.jsonl ├── ... └── wenetspeech4tts.jsonl # JSONL metadata files Each JSONL file… See the full description on the dataset page: https://huggingface.co/datasets/SparkAudio/voxbox.

sourceHugging Facecc-by-nc-sa-4.0updated 1y agoView on Hugging Face
76likes45kdownloads
settings

This repository belongs to SparkAudio on Hugging Face.

CoolFace never edits a repository it does not host. Visibility, licence, collaborators and gating are all managed at the source.

namevoxbox
visibilitypublic
licencecc-by-nc-sa-4.0
gatedno
ownerSparkAudio
Account settings
SparkAudio/voxbox · CoolFace