datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
QWEN3-TTS-Voice-Clone-100-Japanese-Female-ITA-Corpus-EmotionITA-Corpus Emotion Dataset (100 Japanese Female Voices)
彼のあだ名は言い得て妙だよね
11:A lower-pitched female voice with a strong core
ヒューズが飛んだ
100:A slightly quirky female voice that leaves a strong impression
Overview
This dataset contains 100 female voices generated with Qwen3-TTS.
Format: 24kHz mono WAV
Source: Link to designed voices
About ITA-Corpus Emotion
The text is based on the ITA-Corpus Emotion, a public domain dataset containing 100… See the full description on the dataset page: https://huggingface.co/datasets/Akjava/QWEN3-TTS-Voice-Clone-100-Japanese-Female-ITA-Corpus-Emotion.neutral-batch-voice-cloneNaiLong-Voice-Clone
奶龙语音克隆数据集
完整项目与 Demo 效果可参见 GitHub
如果这个数据集对你有帮助,欢迎在 GitHub 上点个 Star ⭐ 支持一下!
数据集介绍
数据集按处理阶段分为以下四部分:
1. raw_audio (原始采样)
处理方式:使用 Audacity 直接对视频素材进行录音,格式为 44.1kHz, 16-bit, Stereo。
说明:包含背景音、特效及多角色对话的非结构化原片素材,是整个流水线的起点。
2. vocal_only (人声分离)
处理方式:从 raw_audio 中使用 UVR5 的 MDX-Net 模型剥离背景音乐与噪音。
说明:利用 MDX-Net 模型提取出干净的人声轨道,为后续切片提供高信噪比素材。
3. sliced_vocal (自动化切片)
处理方式:基于停顿检测、音色突变及总时长控制,将 vocal_only 自动化切分为一系列短音频。… See the full description on the dataset page: https://huggingface.co/datasets/pengyichen/NaiLong-Voice-Clone.voice_clone_trainvoice-clone-v1voice-clone-v2voice_clone_sample_24000hzmy-voice-clone-dataset-speecht5my-voice-clone-dataset-shortmyvoicevoice_clone_samplemy-voice-clone-datasetVoiceClonevoice_clone_taskguzman-clone-voice-reference
