datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
snow-mountainThe Snow Mountain dataset contains the audio recordings (in .mp3 format) and the corresponding text of The Bible
in 11 Indian languages. The recordings were done in a studio setting by native speakers. Each language has a single
speaker in the dataset. Most of these languages are geographically concentrated in the Northern part of India around
the state of Himachal Pradesh. Being related to Hindi they all use the Devanagari script for transcription.BridgeDataV2-audio
📚 Dataset Summary
This dataset contains paired audio-text data generated using:
Text prompts from the BridgeData V2 dataset.
Synthetic speech generated using Coqui-TTS.
Total 21676 Audio data.
Approximately 100 style of speakers.
61528 seconds ≈ 17 hours.
📁 Dataset Structure
Each entry in the dataset contains:
first_line: Textual instruction or caption from BridgeData V2.
audio: Corresponding TTS-generated .wav file.
📜 Licenses and Usage… See the full description on the dataset page: https://huggingface.co/datasets/iknow-lab/BridgeDataV2-audio.
