Audio-Transcription
face-audio-transcriptions-segmentsAudio-Transcription-Models-Comparison-PT-BR
Audio Transcription Models Comparison
A dataset dedicated to comparing the performance of modern Speech-to-Text (STT) models, focusing exclusively on Brazilian Portuguese.
About the Dataset
This dataset was created to store and compare transcription results from different Artificial Intelligence models in challenging scenarios. Unlike generic benchmarks, this project focuses on the reality of usage in Brazil, covering:
Regionalism: Local vocabulary, accents, and… See the full description on the dataset page: https://huggingface.co/datasets/tech4humans/Audio-Transcription-Models-Comparison-PT-BR.Bangla-Youtube-audio-transcription-datasetmig-burmese-audio-transcription
👨💻 Burmese Audio Transcription Dataset
Myanmar (Burmese) audio transcription အတွက် ပြုစုထားသော dataset ဖြစ်ပါတယ်။
Speech to Text, Text to Speech (TTS) နဲ့ ASR လုပ်ငန်းစဉ်များအတွက် တစ်ထောင့်တစ်နေရာက အထောက်အကူပြုနိုင်လိမ့်မယ်လို့ မျှော်လင့်မိပါတယ်။
Samples ပေါင်း 2822 ဝန်းကျင်ခန့် ရှိတာကြောင့် project အသေးလေးတွေအတွက် စမ်းကြည့်နေလို့ ရပါပြီ။
နောက်ပိုင်းမှာလည်း တတ်နိုင်သလောက် ဖြည့်စွတ်ပေးသွားပါမယ်။
Audio ဖိုင်တွေကိုတော့ Ramblings by Hein, Knowledge Worm နဲ့ youtube audio book များမှ… See the full description on the dataset page: https://huggingface.co/datasets/Ko-Yin-Maung/mig-burmese-audio-transcription.CHiME6_formatted_transcriptionsEng-Filipino-Accented-audio-with-human-transcription-call-center-topicThis dataset contains 103+ hours of spontaneous English conversations spoken in a Filipino accent, recorded in a studio environment to ensure crystal-clear audio quality. The conversations are designed as role-play scenarios between agents and customers across a variety of call center domains.
🗣️ Speech Style: Natural, unscripted role-playing between native Filipino-accented English speakers, simulating real-world customer interactions.
🎧 Audio Format: High-quality stereo WAV files, recorded… See the full description on the dataset page: https://huggingface.co/datasets/AIxBlock/Eng-Filipino-Accented-audio-with-human-transcription-call-center-topic.
