CoolFace
Datasetpublic

pengyichen/NaiLong-Voice-Clone

奶龙语音克隆数据集 完整项目与 Demo 效果可参见 GitHub 如果这个数据集对你有帮助,欢迎在 GitHub 上点个 Star ⭐ 支持一下! 数据集介绍 数据集按处理阶段分为以下四部分: 1. raw_audio (原始采样) 处理方式:使用 Audacity 直接对视频素材进行录音,格式为 44.1kHz, 16-bit, Stereo。 说明:包含背景音、特效及多角色对话的非结构化原片素材,是整个流水线的起点。 2. vocal_only (人声分离) 处理方式:从 raw_audio 中使用 UVR5 的 MDX-Net 模型剥离背景音乐与噪音。 说明:利用 MDX-Net 模型提取出干净的人声轨道,为后续切片提供高信噪比素材。 3. sliced_vocal (自动化切片) 处理方式:基于停顿检测、音色突变及总时长控制,将 vocal_only 自动化切分为一系列短音频。… See the full description on the dataset page: https://huggingface.co/datasets/pengyichen/NaiLong-Voice-Clone.

sourceHugging Facecc-by-nc-sa-4.0updated 6mo agoView on Hugging Face
1likes37downloads
discussions and pull requests

Conversations for this repository live on Hugging Face.

CoolFace shows imported repositories read-only. Posting into someone else’s repository from here would need an authorised integration and the account holder’s consent, so the link goes to the source instead.

Open discussions on Hugging Face
pengyichen/NaiLong-Voice-Clone · CoolFace