datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
nsfw_tts_datasetA high-quality audio dataset designed for training and fine-tuning NSFW TTS models, including 30 characters, over 1000 hours of audio, and rich emotion/sound annotations.
Sample format: WAV (audio) + TXT (annotations), including emotion_label, sound_label and text.
Annotations: 6000+ emotion labels (intimate, breathy, teasing, etc.) and 760+ sound labels (moan, sigh, laugh, etc.) in the full version.
Audio sample is as follows:
[intimate, breathy, pleased] Oh, <moan> it feels so good when your… See the full description on the dataset page: https://huggingface.co/datasets/DMC-ykfx33/nsfw_tts_dataset.nsfw_tts_dataset_30speakers
Speaker
Clips
Duration
wav
text
Caspian
3,137
5.31 hours
[possessive, intimate] "<heavy breathing> <moan> You're mine now. <sigh>"
Darius
10,480
16.85 hours
[aroused, intimate] "<moan> fuck, this feels so good."
Declan
4,599
7.78 hours
[intimate, possessive, sinister] "You are mine now, darling. <moan> Completely mine. <moan>"
Elara
23,562
43.78 hours
[intimate, breathless, longing] <moan> I'm so close. Oh, God. Coming.
Elias
16,821
31.36 hours
[intimate, desirous]… See the full description on the dataset page: https://huggingface.co/datasets/DMC-ykfx33/nsfw_tts_dataset_30speakers.
