CoolFace
4 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01UIT-ViToSA /ViToSA-2.0gated ViToSA 2.0: A Multi-Task Approach Towards Robust Vietnamese Audio-Based Toxic Span Detection This is the official repository for the ViToSA 2.0 dataset and model framework, introduced in the paper A Multi-Task Approach Towards Robust Vietnamese Audio-Based Toxic Span Detection, accepted at ICASSP 2026.The dataset and multi-task framework were developed by researchers from the University of Information Technology, VNU-HCM. Citation Information If you use this… See the full description on the dataset page: https://huggingface.co/datasets/UIT-ViToSA/ViToSA-2.0.textautomatic-speech-recognition10K<n<100K0 likes43 downloads7d agoHugging Face02UIT-ViToSA /ViToSA-1.0gated ViToSA: Audio-Based Toxic Spans Detection on Vietnamese Speech Utterances This is the official repository for the ViToSA 1.0 dataset, introduced in the paper ViToSA: Audio-Based Toxic Spans Detection on Vietnamese Speech Utterances, accepted at Interspeech 2025.The dataset was developed by researchers from the University of Information Technology, VNU-HCM. Citation Information If you use this dataset, please cite the following paper:… See the full description on the dataset page: https://huggingface.co/datasets/UIT-ViToSA/ViToSA-1.0.audioautomatic-speech-recognition10K<n<100K3 likes32 downloads11mo agoHugging Face03fosters /uilyam_folkner_pah_verbeny_output_original Пах вербены — арыгінальнае аўдыё Аўтар / Author: Уільям ФолкнерМова / Language: Беларуская (Belarusian) Арыгінальнае аўдыё без апрацоўкі, захаванае ў зыходнай якасці. Частка калекцыі Ministerskija — корпус беларускіх аўдыёкніг. Апрацаваная версія (сегменты ~15 с, выраўнаваная транскрыпцыя): uilyam_folkner_pah_verbeny_output Доўгасць аўдыё 1h22m Радкоў у датасеце 366 Структура Кожны радок змяшчае: audio — арыгінальны аўдыёзапіс text —… See the full description on the dataset page: https://huggingface.co/datasets/fosters/uilyam_folkner_pah_verbeny_output_original.audioautomatic-speech-recognitionn<1K0 likes13 downloads4mo agoHugging Face04fosters /uilyam_folkner_pah_verbeny_all Пах вербены Аўтар / Author: Уільям ФолкнерМова / Language: Беларуская (Belarusian) Аўдыё нарэзана з арыгінальнага запісу ў зыходнай частаце дыскрэтызацыі (native), мона, фрагменты да 30 секунд. Частка калекцыі Belarusian Audiobooks (native). Радкоў у датасеце 416 Працягласць 1 гадз 19 хв Частата дыскрэтызацыі 44100 Hz Каналы мона Даўжыня фрагмента да 30 с Структура Кожны радок змяшчае: audio — аўдыёфрагмент (native SR, мона, ≤30 с) text… See the full description on the dataset page: https://huggingface.co/datasets/fosters/uilyam_folkner_pah_verbeny_all.audioautomatic-speech-recognitionn<1K0 likes12 downloads3mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.