SynDataLab-JA-Refs/irodori-refs-10k-v2
Irodori TTS Reference Voices v2 (10K) 10,000 reference voices generated with Aratako/Irodori-TTS-500M-v2-VoiceDesign (no_ref=True) using a richer caption space than v1: 8 axes (gender × age × pitch × tone × speed × distance × emotion × quality) with per-voice unique caption combinations, plus gender alternation, an incompatibility filter (no contradictory "whisper + speak loudly" combos), and a similarity-rejection window so consecutive voices stay distinct. Each ref's text… See the full description on the dataset page: https://huggingface.co/datasets/SynDataLab-JA-Refs/irodori-refs-10k-v2.
056
../
train-00000-of-00010.parquetdownload
train-00001-of-00010.parquetdownload
train-00002-of-00010.parquetdownload
train-00003-of-00010.parquetdownload
train-00004-of-00010.parquetdownload
train-00005-of-00010.parquetdownload
train-00006-of-00010.parquetdownload
train-00007-of-00010.parquetdownload
train-00008-of-00010.parquetdownload
train-00009-of-00010.parquetdownload
