mustafakali/UyZh-FolkSpeech
UyZh-FolkSpeech:维吾尔语-汉语平行短句与常用词条数据集(四说话人录音全覆盖) UyZh-FolkSpeech is a Uyghur-Chinese parallel dataset with four-speaker audio coverage for each entry. 相关链接 代码与项目主页(GitHub):https://github.com/kalimustafa/UyZh-FolkSpeech 论文与材料(GitHub 仓库内 paper/):https://github.com/kalimustafa/UyZh-FolkSpeech/tree/main/paper 数据集概述 UyZh-FolkSpeech 面向低资源场景的维吾尔语-汉语语言技术研究与工程落地,覆盖新疆地区常见的民间谚语与日常口语表达,并补充常用词与短语条目。数据可用于: 机器翻译(Uyghur ↔ Chinese) 跨语言检索与对话系统 语音增强 NLP、自动语音识别(ASR)… See the full description on the dataset page: https://huggingface.co/datasets/mustafakali/UyZh-FolkSpeech.
This repository belongs to mustafakali on Hugging Face.
CoolFace never edits a repository it does not host. Visibility, licence, collaborators and gating are all managed at the source.
