CoolFace
18 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01tech4humans /Audio-Transcription-Models-Comparison-PT-BR Audio Transcription Models Comparison A dataset dedicated to comparing the performance of modern Speech-to-Text (STT) models, focusing exclusively on Brazilian Portuguese. About the Dataset This dataset was created to store and compare transcription results from different Artificial Intelligence models in challenging scenarios. Unlike generic benchmarks, this project focuses on the reality of usage in Brazil, covering: Regionalism: Local vocabulary, accents, and… See the full description on the dataset page: https://huggingface.co/datasets/tech4humans/Audio-Transcription-Models-Comparison-PT-BR.audioautomatic-speech-recognitionn<1K3 likes138 downloads8mo agoHugging Face02kabirsumaiya /Bangla-Youtube-audio-transcription-datasetaudio1K<n<10K0 likes97 downloads6d agoHugging Face03Ko-Yin-Maung /mig-burmese-audio-transcription 👨‍💻 Burmese Audio Transcription Dataset Myanmar (Burmese) audio transcription အတွက် ပြုစုထားသော dataset ဖြစ်ပါတယ်။ Speech to Text, Text to Speech (TTS) နဲ့ ASR လုပ်ငန်းစဉ်များအတွက် တစ်ထောင့်တစ်နေရာက အထောက်အကူပြုနိုင်လိမ့်မယ်လို့ မျှော်လင့်မိပါတယ်။ Samples ပေါင်း 2822 ဝန်းကျင်ခန့် ရှိတာကြောင့် project အသေးလေးတွေအတွက် စမ်းကြည့်နေလို့ ရပါပြီ။ နောက်ပိုင်းမှာလည်း တတ်နိုင်သလောက် ဖြည့်စွတ်ပေးသွားပါမယ်။ Audio ဖိုင်တွေကိုတော့ Ramblings by Hein, Knowledge Worm နဲ့ youtube audio book များမှ… See the full description on the dataset page: https://huggingface.co/datasets/Ko-Yin-Maung/mig-burmese-audio-transcription.audio1K<n<10K3 likes92 downloads1y agoHugging Face04AIxBlock /Eng-Filipino-Accented-audio-with-human-transcription-call-center-topicThis dataset contains 103+ hours of spontaneous English conversations spoken in a Filipino accent, recorded in a studio environment to ensure crystal-clear audio quality. The conversations are designed as role-play scenarios between agents and customers across a variety of call center domains. 🗣️ Speech Style: Natural, unscripted role-playing between native Filipino-accented English speakers, simulating real-world customer interactions. 🎧 Audio Format: High-quality stereo WAV files, recorded… See the full description on the dataset page: https://huggingface.co/datasets/AIxBlock/Eng-Filipino-Accented-audio-with-human-transcription-call-center-topic.audioautomatic-speech-recognitionn<1K5 likes54 downloads1y agoHugging Face05Nawsantki /mig-burmese-audio-transcription 👨‍💻 Burmese Audio Transcription Dataset Myanmar (Burmese) audio transcription အတွက် ပြုစုထားသော dataset ဖြစ်ပါတယ်။ Speech to Text, Text to Speech (TTS) နဲ့ ASR လုပ်ငန်းစဉ်များအတွက် တစ်ထောင့်တစ်နေရာက အထောက်အကူပြုနိုင်လိမ့်မယ်လို့ မျှော်လင့်မိပါတယ်။ Samples ပေါင်း 2822 ဝန်းကျင်ခန့် ရှိတာကြောင့် project အသေးလေးတွေအတွက် စမ်းကြည့်နေလို့ ရပါပြီ။ နောက်ပိုင်းမှာလည်း တတ်နိုင်သလောက် ဖြည့်စွတ်ပေးသွားပါမယ်။ Audio ဖိုင်တွေကိုတော့ Ramblings by Hein, Knowledge Worm နဲ့ youtube audio book… See the full description on the dataset page: https://huggingface.co/datasets/Nawsantki/mig-burmese-audio-transcription.audio1K<n<10K0 likes50 downloads16d agoHugging Face06hackerlim7 /mig-burmese-audio-transcription 👨‍💻 Burmese Audio Transcription Dataset Myanmar (Burmese) audio transcription အတွက် ပြုစုထားသော dataset ဖြစ်ပါတယ်။ Speech to Text, Text to Speech (TTS) နဲ့ ASR လုပ်ငန်းစဉ်များအတွက် တစ်ထောင့်တစ်နေရာက အထောက်အကူပြုနိုင်လိမ့်မယ်လို့ မျှော်လင့်မိပါတယ်။ Samples ပေါင်း 2822 ဝန်းကျင်ခန့် ရှိတာကြောင့် project အသေးလေးတွေအတွက် စမ်းကြည့်နေလို့ ရပါပြီ။ နောက်ပိုင်းမှာလည်း တတ်နိုင်သလောက် ဖြည့်စွတ်ပေးသွားပါမယ်။ Audio ဖိုင်တွေကိုတော့ Ramblings by Hein, Knowledge Worm နဲ့ youtube audio book… See the full description on the dataset page: https://huggingface.co/datasets/hackerlim7/mig-burmese-audio-transcription.audio1K<n<10K0 likes47 downloads16d agoHugging Face07fabhaus /masri_audio_transcriptionaudio1K<n<10K1 likes15 downloads2y agoHugging Face08LautaroOcho /Argentinian-audio-transcriptions Argentinian audio transcriptions dataset The dataset choosen contains audio recordings of Argentinian speaker along with their corresponding transcription. It was obtain from this link. It consists of 3,921 recordings from female speakers and 1,818 recordings from male speakers. These recordings were downsampled to a 16 kHz sample rate. The spectrogram folder contains the spectrograms of the recordings, as well as spectrograms of augmented versions of those recordings. For more… See the full description on the dataset page: https://huggingface.co/datasets/LautaroOcho/Argentinian-audio-transcriptions.audiotranslation1K<n<10K0 likes14 downloads1y agoHugging Face09amgadhasan /longform-audio-transcriptionaudion<1K0 likes13 downloads2y agoHugging Face10Datasmartly /audio-transcription-sample4-yasgated Dataset Card for "audio-transcription-sample4-yas" More Information needed audion<1K0 likes8 downloads1y agoHugging Face11Siddhant00 /AudioTranscriptionDatasetaudion<1K0 likes6 downloads1y agoHugging Face12Datasmartly /audio-transcription-sample3-yasgated Dataset Card for "audio-transcription-sample3-yas" More Information needed audion<1K0 likes6 downloads1y agoHugging Face13phong3103 /audio-transcription-dataset Dataset Card for "audio-transcription-dataset" More Information needed audion<1K0 likes5 downloads1y agoHugging Face14aya1smartly /audio-transcription-sample1 Dataset Card for "audio-transcription-sample1" More Information needed audion<1K0 likes3 downloads1y agoHugging Face15Datasmartly /audio-transcription-sample1gated Dataset Card for "audio-transcription-sample1" More Information needed audion<1K0 likes3 downloads1y agoHugging Face16Datasmartly /audio-transcription-sample2gated Dataset Card for "audio-transcription-sample2" More Information needed audion<1K0 likes2 downloads1y agoHugging Face17Datasmartly /audio-transcription-aya_2gated Dataset Card for "audio-transcription-aya_2" More Information needed audion<1K0 likes2 downloads1y agoHugging Face18Aspata /eurolife-audio-transcriptionsaudion<1K0 likes1 downloads7mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.