CoolFace
Datasetpublic

ErikMkrtchyan/Hy-Generated-audio-data-2

Hy-Generated Audio Data 2 This dataset provides Armenian speech data consisting of generated audio clips and is addition to this dataset. The generated split contains 137,419 high-quality clips synthesized using a fine-tuned F5-TTS model, covering 404 equal distribution of synthetic voices. 📊 Dataset Statistics Split # Clips Duration (hours) generated 137,419 173.76 Total duration: ~173 hours 🛠️ Loading the Dataset from datasets… See the full description on the dataset page: https://huggingface.co/datasets/ErikMkrtchyan/Hy-Generated-audio-data-2.

sourceHugging Facecc0-1.0updated 1y agoView on Hugging Face
1likes267downloads
Dataset Card

Hy-Generated Audio Data 2

This dataset provides Armenian speech data consisting of generated audio clips and is addition to this dataset.

  • —The generated split contains 137,419 high-quality clips synthesized using a fine-tuned F5-TTS model, covering 404 equal distribution of synthetic voices.

📊 Dataset Statistics

Split# ClipsDuration (hours)
generated137,419173.76

Total duration: ~173 hours

🛠️ Loading the Dataset

python
from datasets import load_dataset

dataset = load_dataset("ErikMkrtchyan/Hy-Generated-audio-data-2")