datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
asr-farsi-youtube-chunked-30-seconds
How To Use
from datasets import load_dataset
train = load_dataset('pourmand1376/asr-farsi-youtube-chunked-30-seconds', split='train+val')
test =load_dataset('pourmand1376/asr-farsi-youtube-chunked-30-seconds', split='test')
+300 Hours ASR dataset generated from this kaggle dataset
second_americas_nlp_2022
Second AmericasNLP 2022
Dataset Summary
This dataset contains the speech data released as part of the Second Workshop on Natural Language Processing for Indigenous Languages of the Americas (AmericasNLP 2022). It provides audio recordings and corresponding transcriptions for several Indigenous languages of the Americas together with Spanish, supporting research on multilingual and low-resource Automatic Speech Recognition (ASR).
This Hugging Face version has been… See the full description on the dataset page: https://huggingface.co/datasets/ivangtorre/second_americas_nlp_2022.yoruba-second-sbpn-demucs-20260826
yoruba-second-sbpn-demucs-20260826
This dataset combines six independently aligned source archives. Each row embeds its selected MP3 in the audio Parquet column. SBPN-derived word timestamps are observational and do not control chunk edges or the Demucs vote. Accepted hard-word verbalizations are projected back to the original written forms; pronunciation_alignment_dictionary_json records the winning spoken form. Non-music tags are preserved using the existing same-speaker… See the full description on the dataset page: https://huggingface.co/datasets/Kppwdfgu1/yoruba-second-sbpn-demucs-20260826.
