datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
homorich-negara-gooya-pron-regen-audiogoodforft
goodforft
This is a merged speech dataset containing 863 audio segments from 4 source datasets.
Dataset Information
Total Segments: 863
Speakers: 4
Languages: en
Emotions: angry, happy, neutral
Original Datasets: 4
Dataset Structure
Each example contains:
audio: Audio file (WAV format, original sampling rate preserved)
text: Transcription of the audio
speaker_id: Unique speaker identifier (made unique across all merged datasets)
emotion: Detected emotion… See the full description on the dataset page: https://huggingface.co/datasets/Codyfederer/goodforft.homorich-negara-gooya-grapheme-regen-audio
