CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01karl-wang /VocalVerse-datasetaudio2 likes630 downloads9mo agoHugging Face02yeeko /Elaina_WanderingWitch_audio_JA 伊蕾娜 语音数据集 (Elaina Voice Audio Dataset) 角色介绍 伊蕾娜(Elaina / イレイナ) 是来自 《魔女之旅》(Majo no Tabitabi / 魔女の旅々) 的主角。 她是一位银发琉璃瞳的旅人魔女,带着温柔的微笑造访各地,见证人间百态。性格冷静又带点俏皮,声音柔和动听,由 本渡楓(Kaede Hondo) 配音。 "这位戴着魔女证明的胸针,飘扬着灰色秀发,那美貌与才能的光辉,让太阳都不禁眯起双眼的美女,到底是谁呢?没错,就是我。"— 伊蕾娜 数据集概述 本数据集收录了 伊蕾娜 日语配音 的音频切片及对应文本。 基本信息 项目 内容 角色 伊蕾娜 (Elaina) 作品 魔女之旅 (Majo no Tabitabi) 配音演员 本渡楓 (Kaede Hondo) 配音语言 日语(JA) 音频来源 B站 / Bilibili 数据格式 Parquet + MP3/WAV 数据预览… See the full description on the dataset page: https://huggingface.co/datasets/yeeko/Elaina_WanderingWitch_audio_JA.audio1K<n<10K1 likes297 downloads6mo agoHugging Face03KishoOoOo /comfyui-wan22-assetsaudion<1K1 likes276 downloads24d agoHugging Face04wannaphong /thai-dialect-isan-dataset Dataset Card for Thai Dialect Isan Speech Corpus Dataset Description This dataset contains audio recordings of Isan (Northeastern Thai) speech, paired with rich transcriptions and demographic metadata. It is designed to support Automatic Speech Recognition (ASR), dialect study, and text normalization tasks for the Isan language. The dataset features spontaneous responses to specific questions, covering two domains (General and Finance), recorded by speakers from different… See the full description on the dataset page: https://huggingface.co/datasets/wannaphong/thai-dialect-isan-dataset.textautomatic-speech-recognition10K<n<100K0 likes122 downloads6mo agoHugging Face05wanasash /enwaucymraegThe training and development set sentences are taken from CoVoST and have been compared to all validated sentences in the Welsh Common Voice data to ensure none of the already recorded sentences will be used here. Then all sentences containing personal names have been extracted and replaced with a randomly generated name using the Faker library and a custom Welsh names list. The sentences were then recorded by 26 volunteers from North-West Wales, 15 women, 10 men and one non-binary person.… See the full description on the dataset page: https://huggingface.co/datasets/wanasash/enwaucymraeg.audioautomatic-speech-recognition1K<n<10K0 likes85 downloads2y agoHugging Face06wannaphong /thai-ser 🇹🇭 THAI-SER Dataset 🎭 [📝 Paper (preprint)] Published by: AI Research Institute of Thailand (AIResearch) In collaboration with: Vidyasirimedhi Institute of Science and Technology (VISTEC) Digital Economy Promotion Agency (depa) Department of Computer Engineering, Faculty of Engineering, Chulalongkorn University Department of Dramatic Arts, Faculty of Arts, Chulalongkorn University Sponsored by: Advanced Info Services Public Company Limited (AIS), and Siam Commercial… See the full description on the dataset page: https://huggingface.co/datasets/wannaphong/thai-ser.audioaudio-classification10K<n<100K0 likes58 downloads8mo agoHugging Face07karl-wang /MuChin1khttps://github.com/CarlWangChina/MuChin audio1K<n<10K2 likes52 downloads1y agoHugging Face08wanasash /whisper-large-v2-eval-cvModel: openai/whisper-large-v2 Test Set: DewiBrynJones/commonvoice_18_0_cy Split: test WER: 41.304748 CER: 16.138398 audio1K<n<10K0 likes49 downloads2y agoHugging Face09opendatalab /WanJuanSiLu-Multimodal-5Languages WanJuan·SiLu Multimodal Multilingual Corpus 🌏Dataset Introduction The newly upgraded "Wanjuan·Silk Road Multimodal Corpus" brings the following three core improvements: The number of languages has been significantly expanded: Based on the five open-source languages ​​of "Wanjuan·Silk Road", namely Arabic, Russian, Korean, Vietnamese, and Thai, "Wanjuan·Silk Road Multimodal" has added three scarce corpus data of Serbian, Hungarian, and Czech, and uses the above… See the full description on the dataset page: https://huggingface.co/datasets/opendatalab/WanJuanSiLu-Multimodal-5Languages.audio100K<n<1M4 likes48 downloads1y agoHugging Face10karl-wang /SaMoyeSVCaudion<1K1 likes41 downloads1y agoHugging Face11wanasash /whisper-large-v3-ec-eval-cvModel: wanasash/whisper-large-v3-ec Test Set: DewiBrynJones/commonvoice_18_0_cy Split: test WER: 38.786166 CER: 11.659708 audio1K<n<10K0 likes34 downloads2y agoHugging Face12wanasash /corpus-siarad-test-setaudio1K<n<10K0 likes34 downloads2y agoHugging Face13svjack /Wang_Leehom_Music_Class_audio_sampleaudio1K<n<10K0 likes32 downloads1y agoHugging Face14dianavdavidson /Vaani-assamese-wancho-nepali-lg-English-no-transcript0audio10K<n<100K0 likes31 downloads4mo agoHugging Face15WandererGuy /fleurs_demoaudio1K<n<10K0 likes30 downloads1y agoHugging Face16jacobavalanchel /wanglihong-matchedaudio10K<n<100K0 likes28 downloads5mo agoHugging Face17wanasash /whisper-large-v2-ec-eval-cvModel: wanasash/whisper-large-v2-ec Test Set: DewiBrynJones/commonvoice_18_0_cy Split: test WER: 38.194402 CER: 11.612194 audio1K<n<10K0 likes25 downloads2y agoHugging Face18dianavdavidson /Vaani-wancho-nepali-majority-lg-English-with-transcriptaudion<1K0 likes24 downloads4mo agoHugging Face19wanasash /whisper-large-v3-ec-eval-ecModel: wanasash/whisper-large-v3-ec Test Set: wanasash/enwaucymraeg Split: test WER: 28.331177 CER: 7.949406 audion<1K0 likes21 downloads2y agoHugging Face20eduhk-compling /11537436_WangYiLinLinda Dataset Description This dataset contains 3 hours of clear and high-quality Mandarin audio at a 44.1kHz sampling rate,sourced from a language laboratory, along with precise textual annotations. The recording was conducted in a manner of 10 sentences per group. Subsequently, the audio was edited and selected to ensure that the volume of each sentence is between 0.3 and 0.7 and that the silent intervals before and after each sentence are between 100 and 200 milliseconds, in order to… See the full description on the dataset page: https://huggingface.co/datasets/eduhk-compling/11537436_WangYiLinLinda.audion<1K0 likes21 downloads8mo agoHugging Face21wandererupak /nepali_asr_evaluation_dataaudio1K<n<10K0 likes21 downloads4mo agoHugging Face22wanglynn /081000audio1K<n<10K0 likes20 downloads1y agoHugging Face23wandererupak /n-demo-ultimateaudio1K<n<10K0 likes15 downloads7mo agoHugging Face24wanghaikuan /sichuanaudio1K<n<10K0 likes12 downloads2y agoHugging Face25wangyueyiiiiiii /audiomc Audio MultiChallenge: A Multi-Turn Evaluation of Spoken Dialogue Systems on Natural Human Interaction Audio MultiChallenge is an open-source benchmark to evaluate E2E spoken dialogue systems under natural multi-turn interaction patterns. Building on the text-based MultiChallenge framework, which evaluates Inference Memory, Instruction Retention, and Self Coherence, we introduce a new axis Voice Editing that tests robustness to mid-utterance speech repairs and backtracking. We… See the full description on the dataset page: https://huggingface.co/datasets/wangyueyiiiiiii/audiomc.audioaudio-text-to-textn<1K0 likes12 downloads5mo agoHugging Face26wangkevin02 /SASLM-demo-audioaudion<1K0 likes11 downloads6mo agoHugging Face27wanasash /whisper-large-v2-ec-eval-ecModel: wanasash/whisper-large-v2-ec Test Set: wanasash/enwaucymraeg Split: test WER: 27.899957 CER: 8.233039 audion<1K0 likes10 downloads2y agoHugging Face28wanasash /whisper-large-v2-eval-ecModel: openai/whisper-large-v2 Test Set: wanasash/enwaucymraeg Split: test WER: 46.959897 CER: 17.439632 audion<1K0 likes10 downloads2y agoHugging Face29svjack /SparkTTS_Wang_Leehom_Ad_wavaudion<1K0 likes10 downloads1y agoHugging Face30karl-wang /MuChin-v2-6066gatedhttps://github.com/CarlWangChina/MuChin-V2-6066 audio1K<n<10K1 likes9 downloads1y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.