datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
nEMO
nEMO: Dataset of Emotional Speech in Polish
Dataset Description
nEMO is a simulated dataset of emotional speech in the Polish language. The corpus contains over 3 hours of samples recorded with the participation of nine actors portraying six emotional states: anger, fear, happiness, sadness, surprise, and a neutral state. The text material used was carefully selected to represent the phonetics of the Polish language. The corpus is available for free under the Creative… See the full description on the dataset page: https://huggingface.co/datasets/amu-cai/nEMO.MG_NEMO
NeMo Tarred Dataset
Generated from MG.
Train rows: 166399 · Test rows: 1668
Shards: 16 · Codec: wav · Sample rate: 16000 Hz mono
Primary text: text · target_lang: ta-IN
is_tarred: true
tarred_audio_filepaths: .../audio__OP_0..15_CL_.tar
manifest_filepath: .../train_manifest.json
MRTS_NEMO
NeMo Tarred Dataset
Generated from MRTS.
Train rows: 76080 · Test rows: 736
Shards: 16 · Codec: wav · Sample rate: 16000 Hz mono
Primary text: text · target_lang: ta-IN
is_tarred: true
tarred_audio_filepaths: .../audio__OP_0..15_CL_.tar
manifest_filepath: .../train_manifest.json
