CoolFace
Datasetpublic

ylacombe/podcast_fillers_by_license

Some Podcasts Podcasts are taken from the PodcastFillers dataset. The PodcastFillers dataset consists of 199 full-length podcast episodes in English with manually annotated filler words and automatically generated transcripts. The podcast audio recordings, sourced from SoundCloud, are CC-licensed, gender-balanced, and total 145 hours of audio from over 350 speakers. [!TIP] This dataset doesn't upload the PodcastFillers annotations, which are under a non-commercial license. See… See the full description on the dataset page: https://huggingface.co/datasets/ylacombe/podcast_fillers_by_license.

sourceHugging Faceccupdated 2y agoView on Hugging Face
0likes160downloads
3 commits on main
6fad2f02y ago

Update README.md

ylacombe
fb6aa232y ago

Upload dataset

ylacombe
654d2cd2y ago

initial commit

ylacombe