CoolFace
18 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01sulabhkatiyar /ne-asr-dataset-nag-aug NE ASR Augmented Dataset -- Nagamese (nag) Augmented automatic speech recognition dataset for Nagamese (nag), a Assamese-based creole language spoken in Nagaland, India. Source Augmented from sulabhkatiyar/ne-asr-nag (original transcribed speech data from the ARTPARK-IISc Vaani project). Language Information Property Value Language Nagamese ISO 639-3 nag Family Assamese-based creole Region Nagaland, India Tonal No Tier D (23.76h… See the full description on the dataset page: https://huggingface.co/datasets/sulabhkatiyar/ne-asr-dataset-nag-aug.audioautomatic-speech-recognition10K<n<100K0 likes687 downloads4mo agoHugging Face02sulabhkatiyar /ne-asr-dataset-nag Nagamese (nag) — ASR dataset A small Nagamese (nag) speech-to-text dataset for automatic speech recognition (ASR) of a low-resource North-East India language. Each example pairs a short audio clip with its Romanized (Latin-script) transcript. Source Derived from the ARTPARK-IISc Vaani project (https://vaani.iisc.ac.in/) Splits Split Samples train 12,862 validation 1,532 test 1,717 Data fields Each example has:… See the full description on the dataset page: https://huggingface.co/datasets/sulabhkatiyar/ne-asr-dataset-nag.audioautomatic-speech-recognition10K<n<100K0 likes172 downloads1mo agoHugging Face03ezfiez /nagatoro-sound # Nagatoro Hayase Voice Dataset (140 Clean Clips) This dataset contains 140 high-quality, pre-processed clean voice clips of the anime character Nagatoro Hayase (Ijiranaide, Nagatoro-san / Don't Toy with Me, Miss Nagatoro). It is specifically curated and optimized for AI voice training pipelines, voice conversion models, and audio deep learning experiments. Dataset Details Character: Nagatoro Hayase (長瀞 早瀬) Language: Japanese (JA) File Format: .wav (High Quality) Total… See the full description on the dataset page: https://huggingface.co/datasets/ezfiez/nagatoro-sound.audioaudio-to-audion<1K2 likes108 downloads3mo agoHugging Face04NagaYu /bleep-spans Bleep spans — synthetic sensitive-speech regions with frame-accurate labels Where sensitive information is spoken, and what kind it is — never what was said. Every recording is synthetic. No real telephone call, clinical recording, or any other real speech was used, recorded, or derived from at any stage. 🤗 Model: NagaYu/bleep-0.09b 🎛️ Demo: NagaYu/bleep What a row contains utt_id, voice_key, condition, duration, subsets, and three parallel arrays —… See the full description on the dataset page: https://huggingface.co/datasets/NagaYu/bleep-spans.audioaudio-classification1K<n<10K0 likes57 downloads7d agoHugging Face05sulabhkatiyar /ne-tts-nag NE-TTS Nagamese (nag) Cleaned TTS dataset for Nagamese (nag), a North East Indian language. Derived from the Vaani dataset with SNR filtering, LUFS normalization, and text cleaning. Stats Metric Value Total clips 14,796 Total hours 21.9h High SNR (>=20dB) 9,688 clips Medium SNR (15-20dB) 5,108 clips Sample rate 16kHz Audio format WAV, 16-bit PCM Schema Column Type Description audio Audio 16kHz WAV audio text… See the full description on the dataset page: https://huggingface.co/datasets/sulabhkatiyar/ne-tts-nag.audiotext-to-speech10K<n<100K0 likes32 downloads4mo agoHugging Face06abar-uwc /vaani-rajasthan_nagaur-cleanedaudio1K<n<10K0 likes31 downloads1y agoHugging Face07sulabhkatiyar /ne-tts-f5-nag NE-TTS F5 Nagamese (nag) F5-TTS training dataset for Nagamese (nag). Contains 9,688 clips at 24kHz (SNR >= 20dB) from the cleaned NE-TTS dataset, formatted for F5-TTS training. Stats Metric Value Clips 9,688 Hours 14.5h Sample rate 24kHz SNR filter >= 20dB Source ne-tts-nag Schema Column Type Description audio Audio 24kHz WAV audio text string Cleaned transcript Note Audio is upsampled from… See the full description on the dataset page: https://huggingface.co/datasets/sulabhkatiyar/ne-tts-f5-nag.audiotext-to-speech1K<n<10K0 likes31 downloads4mo agoHugging Face08dianavdavidson /Vaani-nagamese-majority-lg-English-no-transcript0audio10K<n<100K1 likes27 downloads4mo agoHugging Face09abar-uwc /vaani-maharashtra_nagpur-cleanedaudio1K<n<10K0 likes25 downloads1y agoHugging Face10sulabhkatiyar /ne-tts-mms-nag NE-TTS MMS-VITS Nagamese (nag) MMS-VITS fine-tuning subset for Nagamese (nag). Contains 150 high-quality clips selected from the cleaned NE-TTS dataset (SNR >= 20dB), formatted for MMS-VITS fine-tuning. Stats Metric Value Clips 150 Sample rate 22050Hz SNR filter SNR >= 20dB Source ne-tts-nag Schema Column Type Description audio Audio 22050Hz WAV audio text string Cleaned transcript Usage Use… See the full description on the dataset page: https://huggingface.co/datasets/sulabhkatiyar/ne-tts-mms-nag.audiotext-to-speechn<1K0 likes17 downloads4mo agoHugging Face11dianavdavidson /Vaani-nagamese-majority-lg-English-with-transcriptaudio1K<n<10K0 likes16 downloads4mo agoHugging Face12NagaSaiAbhinay /whisperkit_testsAll files are from: earnings22 Rencoded to 24kbps MP3 using: ffmpeg -i 4446796.wav -vn -map_metadata -1 -ac 1 -c:a libmp3lame -b:a 24k -application voip -y 4446796.mp3 audion<1K0 likes14 downloads2y agoHugging Face13NagaSaiAbhinay /XTTS_testaudio1K<n<10K0 likes13 downloads2y agoHugging Face14kkyo /Nagisinaudion<1K0 likes7 downloads3y agoHugging Face15CoffXD /Nagatoro-voiceaudion<1K0 likes6 downloads3y agoHugging Face16ruchirsahni /Vaani_Nagaur_tran_hin_audioaudion<1K0 likes6 downloads2y agoHugging Face17ruchirsahni /Vaani_Nagpur_tran_hin_audioaudio1K<n<10K0 likes5 downloads2y agoHugging Face18NagareBoshi /trainaudion<1K0 likes3 downloads2y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.