chart
Datasets
All datasets matching “chart”ChartQA
Dataset Card for "ChartQA"
More Information needed
ChartGalaxy
ChartGalaxy: A Dataset for Infographic Chart Understanding and Generation
🤗 Dataset | 🖥️ Code | 📄 Paper | 📄 Arxiv
🔥 News
[2026.09] 🎉🎉 A new high-quality batch of 14,809 synthetic infographic charts has been added.
This update features more complex layouts and richer chart variations.
[2026.02] 🎉🎉 A new batch of data has been added, comprising 108,208 infographic charts.
This update features broader diversity in title designs and more polished layouts… See the full description on the dataset page: https://huggingface.co/datasets/ChartGalaxy/ChartGalaxy.ChartQA
Large-scale Multi-modality Models Evaluation Suite
Accelerating the development of large-scale multi-modality models (LMMs) with lmms-eval
🏠 Homepage | 📚 Documentation | 🤗 Huggingface Datasets
This Dataset
This is a formatted version of ChartQA. It is used in our lmms-eval pipeline to allow for one-click evaluations of large multi-modality models.
@article{masry2022chartqa,
title={ChartQA: A benchmark for question answering about charts with visual and… See the full description on the dataset page: https://huggingface.co/datasets/lmms-lab-encoder/ChartQA.ChartNet
ChartNet: A Million-Scale Multimodal Dataset for Chart Understanding
🌐 Homepage | 📖 arXiv
📝 Changelog
June 3, 2026 — Release of grounded_qa subset and completed reasoning subset (both subject to Notice Regarding Data Availability)
May 15, 2026 — Added link to 30K real-world charts and detailed captions dataset released by our collaborators Abaka AI/2077AI.
April 29, 2026 — Release of an additional 2.5 million row subset core_permissive (subject to… See the full description on the dataset page: https://huggingface.co/datasets/ibm-granite/ChartNet.ChartDiff
ChartDiff: A Large-Scale Benchmark for Comprehending Pairs of Charts
Overview
ChartDiff is a large-scale benchmark for cross-chart comparative summarization, designed to evaluate whether vision-language models can identify differences and generate coherent comparative descriptions across pairs of charts.
Unlike existing chart understanding datasets that emphasize single-chart interpretation, ChartDiff requires models to compare two charts jointly and generate a concise… See the full description on the dataset page: https://huggingface.co/datasets/ckchaos/ChartDiff.Chart2CodeFrom Charts to Code: A Hierarchical Benchmark for Multimodal Models
Welcome to Chart2Code! If you find this repo useful, please give a star ⭐ for encouragement.
Data Overview
Chart2Code is a hierarchical benchmark for evaluating multimodal models on chart understanding and chart-to-code generation. The dataset is organized into five Hugging Face configurations:
level1_direct
level1_customize
level1_figure
level2
level3
In the current Hugging Face… See the full description on the dataset page: https://huggingface.co/datasets/CSU-JPG/Chart2Code.

