CoolFace
7 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01nvidia /hifitts-2 HiFiTTS-2: A Large-Scale High Bandwidth Speech Dataset Dataset Description This repository contains the metadata for HiFiTTS-2, a large scale speech dataset derived from LibriVox audiobooks. For more details, please refer to our paper. The dataset contains metadata for approximately 36.7k hours of audio from 5k speakers that can be downloaded from LibriVox at a 48 kHz sampling rate. The metadata contains estimated bandwidth, which can be used to infer the original… See the full description on the dataset page: https://huggingface.co/datasets/nvidia/hifitts-2.tabular10M<n<100M34 likes1.1k downloads10mo agoHugging Face02kyutai /hifitts2-aligned HiFiTTS-2 word alignments Word-level forced alignments for the HiFiTTS-2 corpus (44 kHz subset, resampled to 24 kHz), as used to train pocket-tts models. Like HiFiTTS-2 itself, this dataset contains no audio — only pointers and annotations. The audio is downloaded from LibriVox and cut locally. Contents train/train_aligned-*.jsonl.gz — the full aligned training manifest eval_aligned.jsonl.gz — a 1000-utterance held-out split scripts/download_audio.py — fetches… See the full description on the dataset page: https://huggingface.co/datasets/kyutai/hifitts2-aligned.tabulartext-to-speech10M<n<100M2 likes144 downloads1mo agoHugging Face03humair025 /hifi-tts-annotatedtabular100K<n<1M0 likes6 downloads7mo agoHugging Face04AdoCleanCode /hifitts2_audio_edit_mfa_v3gatedtabular1K<n<10K0 likes2 downloads11mo agoHugging Face05AdoCleanCode /hifitts2_audio_edit_mfa_v1gatedtabular1K<n<10K0 likes1 downloads11mo agoHugging Face06AdoCleanCode /hifitts2_audio_edit_mfa_v2gatedtabular1K<n<10K0 likes1 downloads11mo agoHugging Face07AdoCleanCode /hifitts2_audio_edit_mfa_v5gatedtabular10K<n<100K0 likes1 downloads11mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.