datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
ViToSA-2.0
ViToSA 2.0: A Multi-Task Approach Towards Robust Vietnamese Audio-Based Toxic Span Detection
This is the official repository for the ViToSA 2.0 dataset and model framework, introduced in the paper A Multi-Task Approach Towards Robust Vietnamese Audio-Based Toxic Span Detection, accepted at ICASSP 2026.The dataset and multi-task framework were developed by researchers from the University of Information Technology, VNU-HCM.
Citation Information
If you use this… See the full description on the dataset page: https://huggingface.co/datasets/UIT-ViToSA/ViToSA-2.0.ViToSA-1.0
ViToSA: Audio-Based Toxic Spans Detection on Vietnamese Speech Utterances
This is the official repository for the ViToSA 1.0 dataset, introduced in the paper ViToSA: Audio-Based Toxic Spans Detection on Vietnamese Speech Utterances, accepted at Interspeech 2025.The dataset was developed by researchers from the University of Information Technology, VNU-HCM.
Citation Information
If you use this dataset, please cite the following paper:… See the full description on the dataset page: https://huggingface.co/datasets/UIT-ViToSA/ViToSA-1.0.uilyam_folkner_pah_verbeny_output_original
Пах вербены — арыгінальнае аўдыё
Аўтар / Author: Уільям ФолкнерМова / Language: Беларуская (Belarusian)
Арыгінальнае аўдыё без апрацоўкі, захаванае ў зыходнай якасці.
Частка калекцыі Ministerskija —
корпус беларускіх аўдыёкніг.
Апрацаваная версія (сегменты ~15 с, выраўнаваная транскрыпцыя):
uilyam_folkner_pah_verbeny_output
Доўгасць аўдыё
1h22m
Радкоў у датасеце
366
Структура
Кожны радок змяшчае:
audio — арыгінальны аўдыёзапіс
text —… See the full description on the dataset page: https://huggingface.co/datasets/fosters/uilyam_folkner_pah_verbeny_output_original.uilyam_folkner_pah_verbeny_all
Пах вербены
Аўтар / Author: Уільям ФолкнерМова / Language: Беларуская (Belarusian)
Аўдыё нарэзана з арыгінальнага запісу ў зыходнай частаце дыскрэтызацыі (native), мона, фрагменты да 30 секунд.
Частка калекцыі Belarusian Audiobooks (native).
Радкоў у датасеце
416
Працягласць
1 гадз 19 хв
Частата дыскрэтызацыі
44100 Hz
Каналы
мона
Даўжыня фрагмента
да 30 с
Структура
Кожны радок змяшчае:
audio — аўдыёфрагмент (native SR, мона, ≤30 с)
text… See the full description on the dataset page: https://huggingface.co/datasets/fosters/uilyam_folkner_pah_verbeny_all.
