CoolFace
Datasetpublicgated

InfoBayAI/Tamil_Podcast_Audio_Dataset_Dual_Channel

Dataset Description This dataset is a large-scale collection of 3,315 hours of processed Tamil dual-channel podcast audio recordings, containing 57,569 hours of processed podcast audio recordings across 12 languages, designed to support the development and training of advanced speech AI and conversational AI systems. It captures real-world podcast conversations across diverse topics and formats. The dataset is organized in a dual-channel format, where corresponding speaker audio… See the full description on the dataset page: https://huggingface.co/datasets/InfoBayAI/Tamil_Podcast_Audio_Dataset_Dual_Channel.

sourceHugging Facecc-by-4.0updated 11d agoView on Hugging Face
0likes18downloads
.gitattributesDownload Raw Back to root

This repository is gated, so its file contents are only served once you have accepted the publisher's terms at Hugging Face. Open it at the source above.