ylacombe/podcast_fillers_by_license
Some Podcasts Podcasts are taken from the PodcastFillers dataset. The PodcastFillers dataset consists of 199 full-length podcast episodes in English with manually annotated filler words and automatically generated transcripts. The podcast audio recordings, sourced from SoundCloud, are CC-licensed, gender-balanced, and total 145 hours of audio from over 350 speakers. [!TIP] This dataset doesn't upload the PodcastFillers annotations, which are under a non-commercial license. See… See the full description on the dataset page: https://huggingface.co/datasets/ylacombe/podcast_fillers_by_license.
0160
Update README.md
Upload dataset
initial commit
