datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
CSS10-Multilingual-LJSpeech
CSS10-Multilingual-LJSpeech
Multilingual speech dataset combining LJSpeech (English) + CSS10 (10 languages) in a consistent LJSpeech format.
Dataset Description
This dataset merges:
LJSpeech: High-quality English speech dataset
CSS10: A collection of single-speaker speech datasets for 10 languages
All audio files are provided in a consistent format suitable for TTS training.
Features
Each sample contains:
audio: Waveform audio sampled at 22,050 Hz
text:… See the full description on the dataset page: https://huggingface.co/datasets/davidguzmanr/CSS10-Multilingual-LJSpeech.preprocessed_jsut_jsss_css10_common_voice_11
Dataset Card for "preprocessed_jsut_jsss_css10_common_voice_11"
More Information needed
preprocessed_jsut_jsss_css10_fleurs_common_voice_11
Dataset Card for "preprocessed_jsut_jsss_css10_fleurs_common_voice_11"
More Information needed
preprocessed_jsut_jsss_css10
Dataset Card for "preprocessed_jsut_jsss_css10"
More Information needed
css10-ja-ljspeech-audit
CSS10 Japanese LJSpeech — aggregate audit
This one-row audit describes ayousanz/css10-ja-ljspeech at revision
149edaf267ff8048c19ca8324fce07ea7423cd14. It excludes transcript text, utterance
IDs, audio paths, hashes, and audio payloads.
The metadata has 6,841 rows and exactly matches 6,841 ZIP audio members. There are two
empty-text rows, one language-review row, and 23 repeated-text groups. Bounded WAV-header
checks show 22.05 kHz mono 32-bit IEEE-float audio; size-derived… See the full description on the dataset page: https://huggingface.co/datasets/ayousanz/css10-ja-ljspeech-audit.
