CoolFace
3 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01deepghs /girlsfrontline_voices_jp JP Voice-Text Dataset for Girls Front Line Waifus This is the JP voice-text dataset for girls front line playable characters. Very useful for fine-tuning or evaluating ASR/ASV models. Only the voices with strictly one voice actor is maintained here to reduce the noise of this dataset. 12508 records, 20.9 hours in total. Average duration is approximately 6.01s. id char_id voice_actor_name voice_title voice_text time sample_rate file_size filename mimetype file_url… See the full description on the dataset page: https://huggingface.co/datasets/deepghs/girlsfrontline_voices_jp.tabularautomatic-speech-recognition10K<n<100K9 likes60 downloads2y agoHugging Face02gabrielclark3330 /transcribed_auspicious_anime_girl_audio Transcribed Auspicious Anime Girl Audio A single-speaker English voice-over dataset containing 75 Arlecchino clips (18.24 minutes) with embedded audio and transcripts. Audio and transcripts were collected from the Genshin Impact Wiki Arlecchino voice-over page, revision 2123801. Only English voice-over files with nonempty transcripts are included. Columns audio: embedded audio bytes and source filename text: transcript source_file: original filename… See the full description on the dataset page: https://huggingface.co/datasets/gabrielclark3330/transcribed_auspicious_anime_girl_audio.audioautomatic-speech-recognitionn<1K0 likes17 downloads2mo agoHugging Face03Aananda-giri /openSLR-Nepali OpenSLR Nepali Speech Dataset (Preprocessed) Dataset Description This is a preprocessed version of the Nepali speech dataset from OpenSLR, ready for training speech models including Automatic Speech Recognition (ASR) and Text-to-Speech (TTS). Dataset Statistics Total Audio Files: 118,231 Total Duration: 57.34 hours Sample Rate: 16kHz Channels: Mono Format: WAV Preprocessing Applied Text Preprocessing: Text cleaning and… See the full description on the dataset page: https://huggingface.co/datasets/Aananda-giri/openSLR-Nepali.audioautomatic-speech-recognition10K<n<100K0 likes14 downloads10mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.