CoolFace
Datasetpublicgated

InfoBayAI/Bengali_Podcast_Audio_Dataset_Dual_Channel

Dataset Description This dataset is a large-scale collection of 7,798 hours of processed Bengali dual-channel podcast audio recordings, containing 57,569 hours of processed podcast audio recordings across 12 languages, designed to support the development and training of advanced speech AI and conversational AI systems. It captures real-world podcast conversations across diverse topics and formats. The dataset is organized in a dual-channel format, where corresponding speaker… See the full description on the dataset page: https://huggingface.co/datasets/InfoBayAI/Bengali_Podcast_Audio_Dataset_Dual_Channel.

sourceHugging Facecc-by-4.0updated 10d agoView on Hugging Face
0likes20downloads
discussions and pull requests

Conversations for this repository live on Hugging Face.

CoolFace shows imported repositories read-only. Posting into someone else’s repository from here would need an authorised integration and the account holder’s consent, so the link goes to the source instead.

Open discussions on Hugging Face
InfoBayAI/Bengali_Podcast_Audio_Dataset_Dual_Channel · CoolFace