CoolFace
13 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01YomnaGharib /dahih-tts2-demucs-cleanedaudio10K<n<100K1 likes1.6k downloads4mo agoHugging Face02Kppwdfgu1 /yoruba-second-sbpn-demucs-20260826gated yoruba-second-sbpn-demucs-20260826 This dataset combines six independently aligned source archives. Each row embeds its selected MP3 in the audio Parquet column. SBPN-derived word timestamps are observational and do not control chunk edges or the Demucs vote. Accepted hard-word verbalizations are projected back to the original written forms; pronunciation_alignment_dictionary_json records the winning spoken form. Non-music tags are preserved using the existing same-speaker… See the full description on the dataset page: https://huggingface.co/datasets/Kppwdfgu1/yoruba-second-sbpn-demucs-20260826.audioautomatic-speech-recognition1K<n<10K0 likes33 downloads28d agoHugging Face03Cybrpgs /mac-m4pro-fresh-diarization-demucs-20260902gated gdrive-sbpn-fresh-diarization-demucs-mac-m4pro-20260902 This dataset combines six independently aligned source archives. Each row embeds its selected MP3 in the audio Parquet column. SBPN-derived word timestamps are observational and do not control chunk edges or the Demucs vote. Accepted hard-word verbalizations are projected back to the original written forms; pronunciation_alignment_dictionary_json records the winning spoken form. Non-music tags are preserved using the… See the full description on the dataset page: https://huggingface.co/datasets/Cybrpgs/mac-m4pro-fresh-diarization-demucs-20260902.audioautomatic-speech-recognition1K<n<10K0 likes26 downloads22d agoHugging Face04Kppwdfgu1 /yoruba-bolanle-sbpn-demucs-20260825gated yoruba-bolanle-sbpn-demucs-20260825 This dataset combines six independently aligned source archives. Each row embeds its selected MP3 in the audio Parquet column. SBPN-derived word timestamps are observational and do not control chunk edges or the Demucs vote. Accepted hard-word verbalizations are projected back to the original written forms; pronunciation_alignment_dictionary_json records the winning spoken form. Non-music tags are preserved using the existing same-speaker… See the full description on the dataset page: https://huggingface.co/datasets/Kppwdfgu1/yoruba-bolanle-sbpn-demucs-20260825.audioautomatic-speech-recognitionn<1K0 likes19 downloads1mo agoHugging Face05Kppwdfgu1 /yoruba-datasold-sbpn-demucs-20260825gated yoruba-datasold-sbpn-demucs-20260825 This dataset combines six independently aligned source archives. Each row embeds its selected MP3 in the audio Parquet column. SBPN-derived word timestamps are observational and do not control chunk edges or the Demucs vote. Accepted hard-word verbalizations are projected back to the original written forms; pronunciation_alignment_dictionary_json records the winning spoken form. Non-music tags are preserved using the existing same-speaker… See the full description on the dataset page: https://huggingface.co/datasets/Kppwdfgu1/yoruba-datasold-sbpn-demucs-20260825.audioautomatic-speech-recognitionn<1K0 likes18 downloads1mo agoHugging Face06Kppwdfgu1 /two-drive-sbpn-demucs-20260825gated two-drive-sbpn-demucs-20260825 This dataset combines six independently aligned source archives. Each row embeds its selected MP3 in the audio Parquet column. SBPN-derived word timestamps are observational and do not control chunk edges or the Demucs vote. Accepted hard-word verbalizations are projected back to the original written forms; pronunciation_alignment_dictionary_json records the winning spoken form. Non-music tags are preserved using the existing same-speaker… See the full description on the dataset page: https://huggingface.co/datasets/Kppwdfgu1/two-drive-sbpn-demucs-20260825.audioautomatic-speech-recognition1K<n<10K0 likes15 downloads1mo agoHugging Face07Kppwdfgu1 /gdrive-sbpn-fresh-diarization-demucs-optimized-terminal-l4-20260814gated gdrive-sbpn-fresh-diarization-demucs-optimized-terminal-l4-20260814 This dataset combines six independently aligned source archives. Each row embeds its selected MP3 in the audio Parquet column. SBPN-derived word timestamps are observational and do not control chunk edges or the Demucs vote. Accepted hard-word verbalizations are projected back to the original written forms; pronunciation_alignment_dictionary_json records the winning spoken form. Non-music tags are preserved… See the full description on the dataset page: https://huggingface.co/datasets/Kppwdfgu1/gdrive-sbpn-fresh-diarization-demucs-optimized-terminal-l4-20260814.audioautomatic-speech-recognition1K<n<10K0 likes7 downloads1mo agoHugging Face08redsky17 /hindi-tts-digital-commentary-no-demucsaudion<1K0 likes6 downloads10mo agoHugging Face09Kppwdfgu1 /gdrive-sbpn-tagged-demucs-chunks-20260812gated gdrive-sbpn-tagged-demucs-chunks-20260812 This dataset combines six independently aligned source archives. Each row embeds its selected MP3 in the audio Parquet column. SBPN-derived word timestamps are observational and do not control chunk edges or the Demucs vote. Accepted hard-word verbalizations are projected back to the original written forms; pronunciation_alignment_dictionary_json records the winning spoken form. Non-music tags are preserved using the existing… See the full description on the dataset page: https://huggingface.co/datasets/Kppwdfgu1/gdrive-sbpn-tagged-demucs-chunks-20260812.audioautomatic-speech-recognition1K<n<10K0 likes6 downloads1mo agoHugging Face10Kppwdfgu1 /gdrive-sbpn-supplied-diarization-demucs-terminal-l4-20260814gated gdrive-sbpn-supplied-diarization-demucs-terminal-l4-20260814 This dataset combines six independently aligned source archives. Each row embeds its selected MP3 in the audio Parquet column. SBPN-derived word timestamps are observational and do not control chunk edges or the Demucs vote. Accepted hard-word verbalizations are projected back to the original written forms; pronunciation_alignment_dictionary_json records the winning spoken form. Non-music tags are preserved using the… See the full description on the dataset page: https://huggingface.co/datasets/Kppwdfgu1/gdrive-sbpn-supplied-diarization-demucs-terminal-l4-20260814.audioautomatic-speech-recognition1K<n<10K0 likes6 downloads1mo agoHugging Face11Kppwdfgu1 /gdrive-sbpn-fresh-diarization-demucs-h100-20260815gated gdrive-sbpn-fresh-diarization-demucs-h100-20260815 This dataset combines six independently aligned source archives. Each row embeds its selected MP3 in the audio Parquet column. SBPN-derived word timestamps are observational and do not control chunk edges or the Demucs vote. Accepted hard-word verbalizations are projected back to the original written forms; pronunciation_alignment_dictionary_json records the winning spoken form. Non-music tags are preserved using the existing… See the full description on the dataset page: https://huggingface.co/datasets/Kppwdfgu1/gdrive-sbpn-fresh-diarization-demucs-h100-20260815.audioautomatic-speech-recognition1K<n<10K0 likes6 downloads1mo agoHugging Face12Kppwdfgu1 /sbpn-music-demucs-side-by-side-20260722gated Original vs Demucs side-by-side review Each row has two playable audio columns: original_audio: the chunk cut from the original full recording. demucs_audio: the same interval cut from the saved full-recording Demucs vocals MP3. The rows are the chunks selected by the production music detector. Use the two players to audit whether Demucs improves speech quality and removes background music. No chunk-level source separation was performed. Audio is mono 48 kHz, 96 kbps MP3.… See the full description on the dataset page: https://huggingface.co/datasets/Kppwdfgu1/sbpn-music-demucs-side-by-side-20260722.audion<1K0 likes4 downloads2mo agoHugging Face13codingninja /demucs_tasarivgatedaudio1K<n<10K0 likes1 downloads2y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.