CoolFace
11 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01turing-motors /STRIDE-QA-Dataset STRIDE-QA Dataset 📦 Dataset STRIDE-QA is a large-scale visual question answering (VQA) dataset for physically grounded spatiotemporal reasoning in autonomous driving. Constructed from 100 hours of multi-sensor driving data in Tokyo, it offers 16 M QA pairs over 270 K frames with dense annotations including 3D bounding boxes, segmentation masks, and multi-object tracks. Category Description Object-centric Spatial QA Spatial relations between two… See the full description on the dataset page: https://huggingface.co/datasets/turing-motors/STRIDE-QA-Dataset.imagevisual-question-answering100K<n<1M9 likes1.4k downloads8mo agoHugging Face02mispeech /MECAT-QAMECAT: A Multi-Experts Constructed Benchmark for Fine-Grained Audio Understanding Tasks 📖 Paper | 🛠️ GitHub | 🎧 Demo | 🔊 MECAT-Caption (HF) Dataset Description MECAT (Multi-Expert Chain for Audio Tasks) is a comprehensive benchmark constructed on large-scale data to evaluate machine understanding of audio content through two core tasks: Audio Captioning: Generating textual descriptions for given audio Audio Question Answering: Answering questions about given audio… See the full description on the dataset page: https://huggingface.co/datasets/mispeech/MECAT-QA.audioaudio-classification100K<n<1M4 likes626 downloads5mo agoHugging Face03Holi-Spatial /HoliSpatial-QA-2Mimage100K<n<1M0 likes200 downloads6mo agoHugging Face04mlfoundations /downstream_validation_qatext1K<n<10K0 likes152 downloads2y agoHugging Face05ragavsachdeva /holoassist-qatextn<1K0 likes72 downloads1y agoHugging Face06Chenrongxin /STRIDE-QA-Dataset STRIDE-QA Dataset 📦 Dataset STRIDE-QA is a large-scale visual question answering (VQA) dataset for physically grounded spatiotemporal reasoning in autonomous driving. Constructed from 100 hours of multi-sensor driving data in Tokyo, it offers 16 M QA pairs over 270 K frames with dense annotations including 3D bounding boxes, segmentation masks, and multi-object tracks. Category Description Object-centric Spatial QA Spatial relations between two… See the full description on the dataset page: https://huggingface.co/datasets/Chenrongxin/STRIDE-QA-Dataset.imagevisual-question-answering100K<n<1M0 likes16 downloads5mo agoHugging Face07DevonPeroutky /reddit-roastme-visual-qaimage1K<n<10K0 likes9 downloads2y agoHugging Face080xScratch /qa-rusttextn<1K1 likes6 downloads2y agoHugging Face090xScratch /Ethereum_QAtextn<1K0 likes3 downloads2y agoHugging Face10WHR-14159 /Grover_qasm_241001gatedtext10K<n<100K0 likes1 downloads2y agoHugging Face11MussEsSein /qa_data_0216text100K<n<1M0 likes1 downloads7mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.