InfoBayAI/Telugu_Podcast_Audio_Dataset_Dual_Channel
Dataset Description This dataset is a large-scale collection of 5,964 hours of processed Telugu dual-channel podcast audio recordings, containing 57,569 hours of processed podcast audio recordings across 12 languages, designed to support the development and training of advanced speech AI, automatic speech recognition (ASR), speaker understanding, audio analytics, and multilingual language technologies. It captures real-world podcast conversations across diverse topics and… See the full description on the dataset page: https://huggingface.co/datasets/InfoBayAI/Telugu_Podcast_Audio_Dataset_Dual_Channel.
017
No card is published for this repository, or it could not be fetched from Hugging Face right now.
