CoolFace
Datasetpublic

MagicDataTech/magicdata-dialect-northeastern-chinese-tts-lite

MagicData-Dialect-Northeastern Chinese-TTS-Lite MAGIC DATA OPEN-SOURCE LICENSE Dataset Overview Item Information Dataset Type N/A Language Chinese Dialect Speech Style Scripted Content N/A Audio Parameters 48 kHz, 16 bits File Format WAV (PCM) Recording Equipment microphone Recording Environment quiet indoor environment License Creative Commons Attribution-NonCommercial-NoDerivatives 4.0 International License… See the full description on the dataset page: https://huggingface.co/datasets/MagicDataTech/magicdata-dialect-northeastern-chinese-tts-lite.

sourceHugging Faceapache-2.0updated 4mo agoView on Hugging Face
0likes190downloads
Dataset Card

MagicData-Dialect-Northeastern Chinese-TTS-Lite

MAGIC DATA OPEN-SOURCE LICENSE

Dataset Overview

ItemInformation
Dataset TypeN/A
LanguageChinese Dialect
Speech StyleScripted
ContentN/A
Audio Parameters48 kHz, 16 bits
File FormatWAV (PCM)
Recording Equipmentmicrophone
Recording Environmentquiet indoor environment
LicenseCreative Commons Attribution-NonCommercial-NoDerivatives 4.0 International License

Statistics

DialectCityCodeDurationSentencesSpeaker
Northeastern ChineseSipingNED10 minutes75 sentences1 female, 30 years old

Dataset Introduction

MagicData-Dialect-Northeastern Chinese-TTS-Lite is an open-source Northeastern Chinese TTS subset of the MagicData-Dialect-TTS-Lite collection released by Magic Data. It focuses on authentic Northeastern Chinese speech and is designed for research scenarios such as dialect speech synthesis, acoustic analysis, and model evaluation.

The dataset contains approximately 10 minutes of speech data, recorded by one native Northeastern Chinese speaker from Siping. The speaker was born and raised in the local region, and the recordings preserve authentic local accent, intonation, and expression habits.

Dataset Features

  1. 1.Native speaker with authentic accent
  2. 2.The speaker was born and raised in the local region until adulthood.
  3. 3.The speaker’s family and main social environment use the local dialect.
  1. 1.Daily-life content coverage
  2. 2.Weather, food, family conversations, numbers, time, and dates
  3. 3.A small amount of emotional expression
  4. 4.No complex technical terms, news reading, or poetry recitation, in order to avoid style deviation
  1. 1.Clean recording environment
  2. 2.Quiet indoor environment
  3. 3.48 kHz / 16-bit / mono WAV
  1. 1.Moderate sentence length, suitable for TTS modeling
  2. 2.Each sentence is around 5–20 seconds, with an average length of about 10 seconds
  3. 3.Natural punctuation-based segmentation, with no forced truncation

Annotation Guidelines

  1. 1.Chinese character transcription: Standard Chinese characters are used, while dialect-specific words are preserved and restored, such as “咋整” in Northeastern Chinese.
  2. 2.Number annotation: Numbers are written in Chinese character form.
  3. 3.Standardization rule: The original dialect sentences are preserved and are not forcibly “translated” into Mandarin.

Annotation Example

Original Northeastern Chinese sentence: 这事儿咋整啊? Annotated text: 这事儿咋整啊?

Open-source File Structure

dialect-tts-lite/ ├── 东北 │ ├── ProsodyLabeling │ │ ├── txt │ ├── wav │ │ ├── wav/ # 75 audio files

Usage Recommendations

Suitable for

  • Zero-shot / few-shot baseline testing for multi-dialect TTS models
  • Acoustic analysis of dialectal phonetic features
  • Comparative experiments in academic research

Not suitable for

  • Directly training production-level dialect TTS products, as the dataset is non-commercial and limited in scale
  • Evaluating extreme scenarios, such as noisy environments, far-field recording, or children’s voices

If you are interested in a larger-scale commercial version, please contact us.

Open-source License

This dataset is for non-commercial use only under the CC BY-NC-ND 4.0 license. It is suitable for academic research, personal development, and model evaluation.

Business Contact

📧 business@magicdatatech.com