CoolFace
Datasetpublic

TigreGotico/tts-train-synthetic-miro_hi-IN

tts-train-synthetic-miro_hi-IN This is a single-speaker synthetic speech dataset for Hindi. It carries the Miro voice, a male voice. The audio was synthesized with text-to-speech and adapted to the Miro speaker identity, to train a Miro voice model for Hindi. The dataset contains 1100 recordings, with a metadata.csv transcript file and one WAV file per line. Related links Dataset collection: https://huggingface.co/collections/TigreGotico/synthetic-tts-datasets… See the full description on the dataset page: https://huggingface.co/datasets/TigreGotico/tts-train-synthetic-miro_hi-IN.

sourceHugging Facecc-by-nc-nd-4.0updated 2mo agoView on Hugging Face
0likes40downloads
Dataset Card

tts-train-synthetic-miro_hi-IN

This is a single-speaker synthetic speech dataset for Hindi. It carries the Miro voice, a male voice. The audio was synthesized with text-to-speech and adapted to the Miro speaker identity, to train a Miro voice model for Hindi. The dataset contains 1100 recordings, with a metadata.csv transcript file and one WAV file per line.

Related links

  • —Dataset collection: https://huggingface.co/collections/TigreGotico/synthetic-tts-datasets
  • —phoonnx: https://github.com/TigreGotico/phoonnx
  • —voiceclonnx: https://github.com/TigreGotico/voiceclonnx
  • —TigreGotico: https://tigregotico.pt

Ownership and licensing

Miro and Dii are the recorded voices of two real people. The voices, this dataset, and any model trained on it belong to TigreGotico Lda (https://tigregotico.pt).

This dataset is licensed under Creative Commons Attribution-NonCommercial-NoDerivatives 4.0 (CC BY-NC-ND 4.0). You may use and share this dataset for non-commercial purposes only. Give attribution to TigreGotico Lda. Do not modify, remix, or build derivative datasets or voices from this data.

For commercial use, for derivative datasets, or for any other license of the Miro or Dii voice identity, contact TigreGotico Lda.