datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
HOIVG-Bench
OmniShow: Unifying Multimodal Conditions for Human-Object Interaction Video Generation
Donghao Zhou1,*, Guisheng Liu2,*, Hao Yang2, Jiatong Li2,†, Jingyu Lin3, Xiaohu Huang4,
Yichen Liu2, Xin Gao2, Cunjian Chen3, Shilei Wen2,§, Chi-Wing Fu1, Pheng-Ann Heng1,§
1The Chinese University of Hong Kong, 2ByteDance, 3Monash University, 4The University of Hong Kong
*Equal contribution, †Project lead, §Corresponding author
🌍 Useful Links
Project Page:… See the full description on the dataset page: https://huggingface.co/datasets/donghao-zhou/HOIVG-Bench.ePark_ju_xing_pian_gao_zhong_sentence_patterns_senior_high
FormosanBank publication status
This audio is associated with XML published in the public FormosanBank corpus and uses the same license recorded in that XML: CC BY-NC-SA 4.0. View the published XML. Publication approval is recorded on the corresponding FormosanBank Basecamp card.
FormosanBank/ePark_ju_xing_pian_gao_zhong_sentence_patterns_senior_high
Commercial AI Use is prohibited without prior written permission. See the FormosanBank Terms of Use and AI Use… See the full description on the dataset page: https://huggingface.co/datasets/FormosanBank/ePark_ju_xing_pian_gao_zhong_sentence_patterns_senior_high.makise-kurisu-vn-voicelines
Makise Kurisu VN dialogue
Transcribed using Whisper Large-V2 from this video.
Clips were separated via pydub, so some text may be incorrect. I have not cleaned it up at all.
Intended for TTS model training.
I do not own any of the content.
zhoujielun-matchedAVQA-Audio-Rubrics
AVQA Audio-Reasoning Rubrics
Project Page | Paper | Code
Audio-grounded, binary-evaluable evaluation rubrics for the full
AVQA training set, generated for
process-level reward modeling in audio reasoning RL (e.g. GRPO / RLHF with
rubric-as-reward).
Each training question is annotated with 5 rubrics, one per evaluation
facet, that judge the quality of an audio-reasoning response — not just final
answer correctness. The rubrics are designed to be scored Yes/No by an
LLM judge that… See the full description on the dataset page: https://huggingface.co/datasets/umd-zhou-lab/AVQA-Audio-Rubrics.ePark_ju_xing_pian_guo_zhong_sentence_patterns_junior_high
FormosanBank publication status
This audio is associated with XML published in the public FormosanBank corpus and uses the same license recorded in that XML: CC BY-NC-SA 4.0. View the published XML. Publication approval is recorded on the corresponding FormosanBank Basecamp card.
FormosanBank/ePark_ju_xing_pian_guo_zhong_sentence_patterns_junior_high
Commercial AI Use is prohibited without prior written permission. See the FormosanBank Terms of Use and AI Use… See the full description on the dataset page: https://huggingface.co/datasets/FormosanBank/ePark_ju_xing_pian_guo_zhong_sentence_patterns_junior_high.OmniShow_example_dataset
OmniShow: Unifying Multimodal Conditions for Human-Object Interaction Video Generation
Donghao Zhou1,*, Guisheng Liu2,*, Hao Yang2, Jiatong Li2,†, Jingyu Lin3, Xiaohu Huang4,
Yichen Liu2, Xin Gao2, Cunjian Chen3, Shilei Wen2,§, Chi-Wing Fu1, Pheng-Ann Heng1,§
1The Chinese University of Hong Kong, 2ByteDance, 3Monash University, 4The University of Hong Kong
*Equal contribution, †Project lead, §Corresponding author
🌍 Useful Links
Project Page:… See the full description on the dataset page: https://huggingface.co/datasets/donghao-zhou/OmniShow_example_dataset.zhongligenshin_zhongli_audiozhongyuanDay_if_sentient_beings_SPLITED_ZHONGLI_adult_CARD_0
zhorzh-simenon-dziauchyna-i-biaskhvostyia-parsiuchki
Дзяўчына і бясхвостыя парсючкі
Metadata
Author: Жорж Сімэнон
Title: Дзяўчына і бясхвостыя парсючкі
Narrator:
Source Group: Аўдыёкнігі
Source:
Notes
The original audio files are preserved as-is:
no conversion;
no re-encoding;
no filename changes inside each split folder, except removing one common top-level archive folder when present.
To avoid Hugging Face Dataset Viewer scan-size errors, the dataset is split into smaller folders.
Target… See the full description on the dataset page: https://huggingface.co/datasets/archivartaunik/zhorzh-simenon-dziauchyna-i-biaskhvostyia-parsiuchki.Day_if_sentient_beings_SPLITED_ZHONGLI_adult_CARD_1
Day_if_sentient_beings_SPLITED_ZHONGLI_adult_CARD_2
MathSpeechzhorzh-simenon-megre-i-chalavek-na-lautsy
Мэгрэ і чалавек на лаўцы
Metadata
Author: Жорж Сімэнон
Title: Мэгрэ і чалавек на лаўцы
Narrator:
Source Group: Аўдыёкнігі
Source:
Notes
The original audio files are preserved as-is:
no conversion;
no re-encoding;
no filename changes inside each split folder, except removing one common top-level archive folder when present.
To avoid Hugging Face Dataset Viewer scan-size errors, the dataset is split into smaller folders.
Target maximum split… See the full description on the dataset page: https://huggingface.co/datasets/archivartaunik/zhorzh-simenon-megre-i-chalavek-na-lautsy.Zhongli
