datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
IndicTTS_Telugu
Telugu Indic TTS Dataset
This dataset is derived from the Indic TTS Database project, specifically using the Telugu monolingual recordings from both male and female speakers. The dataset contains high-quality speech recordings with corresponding text transcriptions, making it suitable for text-to-speech (TTS) research and development.
Dataset Details
Language: Telugu
Total Duration: ~8.74 hours (Male: 4.47 hours, Female: 4.27 hours)
Audio Format: WAV
Sampling Rate:… See the full description on the dataset page: https://huggingface.co/datasets/SPRINGLab/IndicTTS_Telugu.telugu-tech-custom-voice
🎙️ Telugu Tech Custom Voice Dataset
A high-quality, clean single-speaker Telugu Speech & Voice dataset tailored for training and fine-tuning neural Text-to-Speech (TTS) models (e.g. Coqui XTTS v2, Piper TTS, VITS, Bark) and Automatic Speech Recognition (ASR).
📊 Dataset Statistics
Total Clips: 455 audio files (.wav)
Total Audio Duration: 1 Hour 12 Minutes 48.5 Seconds (4,368.5 seconds)
Total Dataset Size: ~1.20 GB
Language: Telugu (te) with technical terms /… See the full description on the dataset page: https://huggingface.co/datasets/Noothi/telugu-tech-custom-voice.Telugu_ASR_corpus
Dataset Card for "Telugu_ASR_corpus"
More Information needed
telugu-asrtelugu-tech-indicf5-custom-voice
🎙️ Telugu Tech IndicF5 Custom Voice Dataset
A 100% verified, clean, single-speaker Telugu Speech & Voice dataset specially formatted and phonetically cleaned for training and fine-tuning ai4bharat/IndicF5 and neural Text-to-Speech (TTS) models.
All English technical terms, numbers, acronyms, and ASR mishearings have been converted into native Telugu phonetic script, cleaned of noise/brackets, and validated for optimal IndicF5 fine-tuning performance.
📊 Dataset… See the full description on the dataset page: https://huggingface.co/datasets/Noothi/telugu-tech-indicf5-custom-voice.syspin-telugu-ttstelugu-speech-datasetnoisy_telugutelugu-indicf5-evaluationtelugu-voices-rawtelugu_OpenSLRtelugu_whisper_asrTelugu_whisper_ASR_datasettelugu_asr_630hrTelugu-Audio-Corpusindic-tts-telugutts_synthetic_te-IN_Telugutelugu-movies-speechtelugu-raw-audioTelugu_Call_Center_Audio_Dataset_Dual_ChannelDataset Description:
This dataset is a large-scale collection of 27,787 hours of processed Telugu (TE) dual-channel call center audio recordings, containing 3,569,083 hours of processed call center audio recordings across 54 languages, designed to support the development and training of advanced speech AI and conversational AI systems.
It consists of real-world customer and agent speech recordings collected from call center environments. The dataset is organized in a dual-channel format, where… See the full description on the dataset page: https://huggingface.co/datasets/InfoBayAI/Telugu_Call_Center_Audio_Dataset_Dual_Channel.TTS_Telugu
Telugu Indic TTS Dataset
This dataset is derived from the Indic TTS Database project, specifically using the Telugu monolingual recordings from both male and female speakers. The dataset contains high-quality speech recordings with corresponding text transcriptions, making it suitable for text-to-speech (TTS) research and development.
Dataset Details
Language: Telugu
Total Duration: ~8.74 hours (Male: 4.47 hours, Female: 4.27 hours)
Audio Format: WAV
Sampling… See the full description on the dataset page: https://huggingface.co/datasets/tmtanu/TTS_Telugu.Telugu-Call-Center-Audio-Dataset-Single-ChannelDataset Description:
This dataset is a large-scale collection of 27,787 hours of processed Telugu (TE) single-channel call center audio recordings, containing 3,569,083 hours of processed call center audio recordings across 54 languages, designed to support the development and training of advanced speech AI and conversational AI systems.
The dataset captures authentic speech characteristics such as tone variation, pauses, silence patterns, and natural speaking behaviour commonly observed in… See the full description on the dataset page: https://huggingface.co/datasets/InfoBayAI/Telugu-Call-Center-Audio-Dataset-Single-Channel.telugu_v2_evaltelugu-asr-speech-datatelugu_v3_evalVaani-telugu-lg-English-no-transcript1Telugu_Podcast_Audio_Dataset_Dual_Channel
Dataset Description
This dataset is a large-scale collection of 5,964 hours of processed Telugu dual-channel podcast audio recordings, containing 57,569 hours of processed podcast audio recordings across 12 languages, designed to support the development and training of advanced speech AI, automatic speech recognition (ASR), speaker understanding, audio analytics, and multilingual language technologies.
It captures real-world podcast conversations across diverse topics and… See the full description on the dataset page: https://huggingface.co/datasets/InfoBayAI/Telugu_Podcast_Audio_Dataset_Dual_Channel.ai4bharat_telugu_datasetVaani-telugu-lg-English-no-transcript0telugu-speech-dataset
