datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
csvps
Cityscapes VPS
This dataset is derived from the videos in the validation split of the Cityscapes[^1] dataset.
It aggregates the images and metadata from Cityscapes[^1], Cityscapes-VPS[^2] and Cityscapes-DVPS[^3] into a single structured format.
This comprehensive derivative was created out of the need for a batteries-included variant of the dataset for academic purposes.
Specifically, joining samples from the individual datasets in their original structure (each is organized… See the full description on the dataset page: https://huggingface.co/datasets/khwstolle/csvps.CSVQA
# CSVQA (Chinese Science Visual Question Answering)
| 🏆 Leaderboard | 📄 arXiv | 💻 GitHub | 🌐 Webpage | 📄 Paper |
🔥News
June 2, 2025: Our paper is now available on arXiv and we welcome citations:CSVQA: A Chinese Multimodal Benchmark for Evaluating STEM Reasoning Capabilities of VLMs
May 30, 2025: We developed a complete evaluation pipeline, and the implementation details are available on GitHub
📖 Dataset Introduction
Vision-Language Models (VLMs) have… See the full description on the dataset page: https://huggingface.co/datasets/Skywork/CSVQA.SDXL_Data_CSVformat_txtdiabetes_binary_health_indicators_BRFSS2015.csvLaion_aesthetics_5plus_1024_33M_csvDATA_CSVchatinterface_with_image_csv
Dataset Card for Dataset Name
Dataset Summary
[More Information Needed]
Supported Tasks and Leaderboards
[More Information Needed]
Languages
[More Information Needed]
Dataset Structure
Data Instances
[More Information Needed]
Data Fields
[More Information Needed]
Data Splits
[More Information Needed]
Dataset Creation
Curation Rationale
[More Information Needed]
Source Data… See the full description on the dataset page: https://huggingface.co/datasets/abidlabs/chatinterface_with_image_csv.HCS_Dataset-csvchatinterface_with_image_csv3
Dataset Card for Dataset Name
Dataset Summary
[More Information Needed]
Supported Tasks and Leaderboards
[More Information Needed]
Languages
[More Information Needed]
Dataset Structure
Data Instances
[More Information Needed]
Data Fields
[More Information Needed]
Data Splits
[More Information Needed]
Dataset Creation
Curation Rationale
[More Information Needed]
Source Data… See the full description on the dataset page: https://huggingface.co/datasets/abidlabs/chatinterface_with_image_csv3.quare-primitive-counter-9k-raw-csvtest_cups_with_csvchatinterface_with_image_csv
Dataset Card for Dataset Name
Dataset Summary
[More Information Needed]
Supported Tasks and Leaderboards
[More Information Needed]
Languages
[More Information Needed]
Dataset Structure
Data Instances
[More Information Needed]
Data Fields
[More Information Needed]
Data Splits
[More Information Needed]
Dataset Creation
Curation Rationale
[More Information Needed]
Source Data… See the full description on the dataset page: https://huggingface.co/datasets/freddyaboulton/chatinterface_with_image_csv.philosophy-culture-translations-html-csv
AI-Culture Philosophy and Culture Translations CSV + HTML Corpus
The corpus contains an exceptionally diverse range of cultural, philosophical, and literary texts, available in 12 major languages. Among other topics, there is extensive engagement with the ethics and aesthetics of artificial intelligence and its cultural and philosophical implications, as well as connections between AI and philosophy of language and philosophy of mind.
This project is maintained by a non-profit… See the full description on the dataset page: https://huggingface.co/datasets/AI-Culture-Commons/philosophy-culture-translations-html-csv.CSVault-Items
Dataset Card for Counter-Strike 2 Skins Database
Dataset Summary
This dataset contains a comprehensive collection of all skins from Counter-Strike 2. It includes metadata and 1534 high-quality PNG images for each skin. The dataset is useful for researchers, developers, building applications related to CS2 skins.
Dataset Structure
Data Format
The dataset is provided in JSON format, where each entry represents a skin with associated… See the full description on the dataset page: https://huggingface.co/datasets/Llakadmatatag/CSVault-Items.chatinterface_with_image_csv2
Dataset Card for Dataset Name
Dataset Summary
[More Information Needed]
Supported Tasks and Leaderboards
[More Information Needed]
Languages
[More Information Needed]
Dataset Structure
Data Instances
[More Information Needed]
Data Fields
[More Information Needed]
Data Splits
[More Information Needed]
Dataset Creation
Curation Rationale
[More Information Needed]
Source Data… See the full description on the dataset page: https://huggingface.co/datasets/abidlabs/chatinterface_with_image_csv2.test-imagefolder-metadata-csvCSVQA
# CSVQA (Chinese Science Visual Question Answering)
| 🏆 Leaderboard | 📄 arXiv | 💻 GitHub | 🌐 Webpage | 📄 Paper |
🔥News
June 2, 2025: Our paper is now available on arXiv and we welcome citations:CSVQA: A Chinese Multimodal Benchmark for Evaluating STEM Reasoning Capabilities of VLMs
May 30, 2025: We developed a complete evaluation pipeline, and the implementation details are available on GitHub
📖 Dataset Introduction
Vision-Language Models (VLMs)… See the full description on the dataset page: https://huggingface.co/datasets/jzzzw/CSVQA.csvchatinterface_with_image_csv4
Dataset Card for Dataset Name
Dataset Summary
[More Information Needed]
Supported Tasks and Leaderboards
[More Information Needed]
Languages
[More Information Needed]
Dataset Structure
Data Instances
[More Information Needed]
Data Fields
[More Information Needed]
Data Splits
[More Information Needed]
Dataset Creation
Curation Rationale
[More Information Needed]
Source Data… See the full description on the dataset page: https://huggingface.co/datasets/abidlabs/chatinterface_with_image_csv4.csv_outupDreamLIP_capion_csv_w_keyaftermath_exp2_testtime.csv
Aftermath of DrawEduMath
This contains exp2_testtime.csv, for recreating the results of the paper titled "The Aftermath of DrawEduMath: Vision Language Models Underperform with Struggling Students and Misdiagnose Errors".
This file contains model predictions for DrawEduMath QA from eleven vision-language models. Unlike the original benchmark, we input models' self-generated descriptions of student images as part of the QA prompt, to see how test-time scaling may improve results.… See the full description on the dataset page: https://huggingface.co/datasets/lucy3/aftermath_exp2_testtime.csv.Danbooru2023-CSV-Simpletest-imagefolder-metadata-csv-nov-22test-csv-conversion
license: cc-by-nc-nd-4.0
prompt_handbook_csvcsv_vs_viz
CSVvsVizQA
Dataset to compare question answering ability from CSV Data vs Data Visualization images.
ms2525d-landunit-symbols-csvtest_csv_data3923Combined_HCS_Dataset_csv
