datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
MoeGirlPedia_wikitext_raw_archiveGlad to see models and datasets were inspired from this dataset, thanks to all who are using this dataset in their training materials.
Feel free to re-upload the contents to places like the Internet Archive (Please follow the license and keep these files as-is) to help preserve this digital asset.
Looking forward to see more models and synthetic datasets trained from this raw archive, good luck!
Note: Due to the content censorship system introduced by MGP on 2024/03/29, it is unclear that… See the full description on the dataset page: https://huggingface.co/datasets/milashkaarshif/MoeGirlPedia_wikitext_raw_archive.ai-modelsmoe-speech-5speakers-ljspeech
moe-speech-5speakers-ljspeech
Japanese multi-speaker TTS dataset in LJSpeech format.
Overview
Speakers: 5
Total utterances: 19,617
Format: LJSpeech (wavs/ + metadata.csv)
Sample rate: 22050 Hz
Language: Japanese
Speaker Statistics
Speaker ID
Utterances
940de876
4,675
2cf01874
4,632
bbd90363
3,747
1a5a3db8
3,295
4e2f4ba6
3,268
Files
metadata.csv - Full metadata (id|speaker|text)
metadata_{speaker_id}.csv - Per-speaker… See the full description on the dataset page: https://huggingface.co/datasets/ayousanz/moe-speech-5speakers-ljspeech.
