CoolFace
6 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01erenzhou /AVI-Math Dataset Sources Repository: https://github.com/VisionXLab/avi-math Paper: https://arxiv.org/abs/2509.10059 BibTeX: @ARTICLE{zhou2025avimath, author={Zhou, Yue and Feng, Litong and Lan, Mengcheng and Yang, Xue and Li, Qingyun and Ke, Yiping and Jiang, Xue and Zhang, Wayne}, journal={ISPRS Journal of Photogrammetry and Remote Sensing}, title={Multimodal Mathematical Reasoning Embedded in Aerial Vehicle Imagery: Benchmarking, Analysis, and Exploration}, year={2025}… See the full description on the dataset page: https://huggingface.co/datasets/erenzhou/AVI-Math.imagequestion-answering1K<n<10K1 likes492 downloads3mo agoHugging Face02TUM-AVS /Nurisk-ICRA2026 Nurisk: VQA for Risk Assessment in Autonomous Driving Nurisk is a visual question answering dataset focusing on risk assessment for autonomous driving. Each row contains: image: a BEV image question: a driving-related question answer: the ground truth answer Paper NuRisk: A Visual Question Answering Dataset for Agent-Level Risk Assessment in Autonomous Driving — see the paper on arXiv:2509.25944 . Framework Dataset Structure… See the full description on the dataset page: https://huggingface.co/datasets/TUM-AVS/Nurisk-ICRA2026.imagequestion-answering10K<n<100K1 likes424 downloads3mo agoHugging Face03AV-Odyssey /AV_Odyssey_BenchOfficial dataset for the paper "AV-Odyssey: Can Your Multimodal LLMs Really Understand Audio-Visual Information?". 🌟 For more details, please refer to the project page with data examples: https://av-odyssey.github.io/. [🌐 Webpage] [📖 Paper] [🤗 Huggingface AV-Odyssey Dataset] [🤗 Huggingface Deaftest Dataset] [🏆 Leaderboard] 🔥 News 2024.11.24 🌟 We release AV-Odyssey, the first-ever comprehensive evaluation benchmark to explore whether MLLMs really understand audio-visual… See the full description on the dataset page: https://huggingface.co/datasets/AV-Odyssey/AV_Odyssey_Bench.audioquestion-answeringn<1K5 likes303 downloads2y agoHugging Face04Yuan-avs /Nurisk Nurisk: VQA for Risk Assessment in Autonomous Driving Nurisk is a visual question answering dataset focusing on risk assessment for autonomous driving. Each row contains: image: a BEV image question: a driving-related question answer: the ground truth answer Paper NuRisk: A Visual Question Answering Dataset for Agent-Level Risk Assessment in Autonomous Driving — see the paper on arXiv:2509.25944 . Framework Dataset Structure… See the full description on the dataset page: https://huggingface.co/datasets/Yuan-avs/Nurisk.imagequestion-answering10K<n<100K4 likes221 downloads4mo agoHugging Face05AV-Odyssey /Deaftest_datasetOfficial Deaftest dataset for the paper "AV-Odyssey: Can Your Multimodal LLMs Really Understand Audio-Visual Information?". 🌟 For more details, please refer to the project page with data examples: https://av-odyssey.github.io/. [🌐 Webpage] [📖 Paper] [🤗 Huggingface AV-Odyssey Dataset] [🤗 Huggingface Deaftest Dataset] [🏆 Leaderboard] 🔥 News 2024.11.24 🌟 We release AV-Odyssey, the first-ever comprehensive evaluation benchmark to explore whether MLLMs really understand… See the full description on the dataset page: https://huggingface.co/datasets/AV-Odyssey/Deaftest_dataset.audioquestion-answeringn<1K1 likes42 downloads2y agoHugging Face06SW-Yoon /AVSimagequestion-answering10K<n<100K0 likes17 downloads2y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.