voice-design
QWEN3-TTS-Voice-Design-100-Japanese-Female-Designed-Voices100 Japanese Female Designed Voices by Qwen3-TTS-12Hz-1.7B-VoiceDesign
Note: Contains frequent misreadings. Correct reading data is not provided.
AI Generation: The 100 styles were automatically generated by AI, so there may be some overlaps or duplicates.
voice is designed by Japanese Prompt(see styles_jp.txt)
Dataset: 300 audio clips (100 styles × 3 iterations).
Structure: design1–design3 represent each iteration. Each output is unique.
Fixes: Replaced one instance of a "complete error"… See the full description on the dataset page: https://huggingface.co/datasets/Akjava/QWEN3-TTS-Voice-Design-100-Japanese-Female-Designed-Voices.TTS-Voice-Design-Benchmark
TTS Voice Design Benchmark
🏆 Leaderboard | 🛠️ Evaluation Suite
TTS Voice Design is a high-quality benchmark of 1,000 character voice-design
tasks spanning a broad range of media genres and real-world creative use
cases. It evaluates whether a text-to-speech model can turn an open-ended
character profile into a distinctive, appropriate, and usable voice.
Unlike benchmarks built around a fixed set of speakers or isolated acoustic
attributes, this dataset covers complete… See the full description on the dataset page: https://huggingface.co/datasets/BreezeBlue/TTS-Voice-Design-Benchmark.QWEN3-TTS-Voice-Design-100-Japanese-Female-Designed-Voices100 Japanese Female Designed Voices by Qwen3-TTS-12Hz-1.7B-VoiceDesign
Note: Contains frequent misreadings. Correct reading data is not provided.
AI Generation: The 100 styles were automatically generated by AI, so there may be some overlaps or duplicates.
voice is designed by Japanese Prompt(see styles_jp.txt)
Dataset: 300 audio clips (100 styles × 3 iterations).
Structure: design1–design3 represent each iteration. Each output is unique.
Fixes: Replaced one instance of a "complete error"… See the full description on the dataset page: https://huggingface.co/datasets/196vm3/QWEN3-TTS-Voice-Design-100-Japanese-Female-Designed-Voices.voice-design-bench-50voice-design-bench-50-dnsmos
Model
Mean
Min
Max
Qwen
3.262
2.480
3.619
Echo
3.217
2.158
3.588
Omni
3.190
0.988
3.629
VoiceDesign3
VoiceDesign
Vietnamese TTS dataset synthesized
