datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
Belarusian-Speech-Dataset
🎧 Belarusian Speech Dataset
The Belarusian Speech Dataset is a high-quality speech audio dataset designed to provide structured and diverse audio data for AI systems focused on speech technologies and language preservation. It includes 118 hours of audio data across 571 files, delivered in MP3 and WAV formats, with a total size of 192 MB. This well-balanced audio dataset ensures reliable voice data, with 55% female and 45% male speakers, and an age distribution spanning from 18 to… See the full description on the dataset page: https://huggingface.co/datasets/Speech-data/Belarusian-Speech-Dataset.BELLE-eval-S2S
BELLE-eval-S2S
💡 Dataset Description
BELLE-eval-S2S is a Chinese evaluation dataset for speech-to-speech conversational tasks. It contains 250 Chinese audio samples with corresponding text annotations and is intended for model evaluation rather than training.
🔗 Source
Original text source: the test set from LianjiaTech/BELLE
This dataset is built by filtering 250 samples from the original test set and synthesizing them into speech audio
📖 Data… See the full description on the dataset page: https://huggingface.co/datasets/ICTNLP/BELLE-eval-S2S.
