datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
dahih-tts2-demucs-cleanedyoruba-second-sbpn-demucs-20260826
yoruba-second-sbpn-demucs-20260826
This dataset combines six independently aligned source archives. Each row embeds its selected MP3 in the audio Parquet column. SBPN-derived word timestamps are observational and do not control chunk edges or the Demucs vote. Accepted hard-word verbalizations are projected back to the original written forms; pronunciation_alignment_dictionary_json records the winning spoken form. Non-music tags are preserved using the existing same-speaker… See the full description on the dataset page: https://huggingface.co/datasets/Kppwdfgu1/yoruba-second-sbpn-demucs-20260826.mac-m4pro-fresh-diarization-demucs-20260902
gdrive-sbpn-fresh-diarization-demucs-mac-m4pro-20260902
This dataset combines six independently aligned source archives. Each row embeds its selected MP3 in the audio Parquet column. SBPN-derived word timestamps are observational and do not control chunk edges or the Demucs vote. Accepted hard-word verbalizations are projected back to the original written forms; pronunciation_alignment_dictionary_json records the winning spoken form. Non-music tags are preserved using the… See the full description on the dataset page: https://huggingface.co/datasets/Cybrpgs/mac-m4pro-fresh-diarization-demucs-20260902.yoruba-bolanle-sbpn-demucs-20260825
yoruba-bolanle-sbpn-demucs-20260825
This dataset combines six independently aligned source archives. Each row embeds its selected MP3 in the audio Parquet column. SBPN-derived word timestamps are observational and do not control chunk edges or the Demucs vote. Accepted hard-word verbalizations are projected back to the original written forms; pronunciation_alignment_dictionary_json records the winning spoken form. Non-music tags are preserved using the existing same-speaker… See the full description on the dataset page: https://huggingface.co/datasets/Kppwdfgu1/yoruba-bolanle-sbpn-demucs-20260825.yoruba-datasold-sbpn-demucs-20260825
yoruba-datasold-sbpn-demucs-20260825
This dataset combines six independently aligned source archives. Each row embeds its selected MP3 in the audio Parquet column. SBPN-derived word timestamps are observational and do not control chunk edges or the Demucs vote. Accepted hard-word verbalizations are projected back to the original written forms; pronunciation_alignment_dictionary_json records the winning spoken form. Non-music tags are preserved using the existing same-speaker… See the full description on the dataset page: https://huggingface.co/datasets/Kppwdfgu1/yoruba-datasold-sbpn-demucs-20260825.two-drive-sbpn-demucs-20260825
two-drive-sbpn-demucs-20260825
This dataset combines six independently aligned source archives. Each row embeds its selected MP3 in the audio Parquet column. SBPN-derived word timestamps are observational and do not control chunk edges or the Demucs vote. Accepted hard-word verbalizations are projected back to the original written forms; pronunciation_alignment_dictionary_json records the winning spoken form. Non-music tags are preserved using the existing same-speaker… See the full description on the dataset page: https://huggingface.co/datasets/Kppwdfgu1/two-drive-sbpn-demucs-20260825.gdrive-sbpn-fresh-diarization-demucs-optimized-terminal-l4-20260814
gdrive-sbpn-fresh-diarization-demucs-optimized-terminal-l4-20260814
This dataset combines six independently aligned source archives. Each row embeds its selected MP3 in the audio Parquet column. SBPN-derived word timestamps are observational and do not control chunk edges or the Demucs vote. Accepted hard-word verbalizations are projected back to the original written forms; pronunciation_alignment_dictionary_json records the winning spoken form. Non-music tags are preserved… See the full description on the dataset page: https://huggingface.co/datasets/Kppwdfgu1/gdrive-sbpn-fresh-diarization-demucs-optimized-terminal-l4-20260814.hindi-tts-digital-commentary-no-demucsgdrive-sbpn-tagged-demucs-chunks-20260812
gdrive-sbpn-tagged-demucs-chunks-20260812
This dataset combines six independently aligned source archives. Each row embeds its selected MP3 in the audio Parquet column. SBPN-derived word timestamps are observational and do not control chunk edges or the Demucs vote. Accepted hard-word verbalizations are projected back to the original written forms; pronunciation_alignment_dictionary_json records the winning spoken form. Non-music tags are preserved using the existing… See the full description on the dataset page: https://huggingface.co/datasets/Kppwdfgu1/gdrive-sbpn-tagged-demucs-chunks-20260812.gdrive-sbpn-supplied-diarization-demucs-terminal-l4-20260814
gdrive-sbpn-supplied-diarization-demucs-terminal-l4-20260814
This dataset combines six independently aligned source archives. Each row embeds its selected MP3 in the audio Parquet column. SBPN-derived word timestamps are observational and do not control chunk edges or the Demucs vote. Accepted hard-word verbalizations are projected back to the original written forms; pronunciation_alignment_dictionary_json records the winning spoken form. Non-music tags are preserved using the… See the full description on the dataset page: https://huggingface.co/datasets/Kppwdfgu1/gdrive-sbpn-supplied-diarization-demucs-terminal-l4-20260814.gdrive-sbpn-fresh-diarization-demucs-h100-20260815
gdrive-sbpn-fresh-diarization-demucs-h100-20260815
This dataset combines six independently aligned source archives. Each row embeds its selected MP3 in the audio Parquet column. SBPN-derived word timestamps are observational and do not control chunk edges or the Demucs vote. Accepted hard-word verbalizations are projected back to the original written forms; pronunciation_alignment_dictionary_json records the winning spoken form. Non-music tags are preserved using the existing… See the full description on the dataset page: https://huggingface.co/datasets/Kppwdfgu1/gdrive-sbpn-fresh-diarization-demucs-h100-20260815.sbpn-music-demucs-side-by-side-20260722
Original vs Demucs side-by-side review
Each row has two playable audio columns:
original_audio: the chunk cut from the original full recording.
demucs_audio: the same interval cut from the saved full-recording Demucs
vocals MP3.
The rows are the chunks selected by the production music detector. Use the two
players to audit whether Demucs improves speech quality and removes background
music. No chunk-level source separation was performed. Audio is mono 48 kHz,
96 kbps MP3.… See the full description on the dataset page: https://huggingface.co/datasets/Kppwdfgu1/sbpn-music-demucs-side-by-side-20260722.demucs_tasariv
