datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
TABLET-Large
TABLET-Large
This is the Large sized train set of the TABLET dataset. It contains all train examples for all TABLET tasks, resulting in a total of 3,505,311 training examples across 17 tasks.This dataset is self-contained, each example includes a table image, its HTML representation, and the associated task data.However, if you're interested in downloading just the TABLET tables, check out TABLET-tables.
All TABLET Subsets:
(train) TABLET-Small: The smallest TABLET subset… See the full description on the dataset page: https://huggingface.co/datasets/alonsoapp/TABLET-Large.elements_annotated_tables_4500_docs
Dataset
🚀 Progress
Last update (UTC): 2025-11-11 15:40:21Z
Documents processed: 4500 / 500058
Batches completed: 30
Total pages/rows uploaded: 89882
Latest batch summary
Batch index: 30
Docs in batch: 150
Pages/rows added: 1487
TABLET-Medium
TABLET-Medium
This is the Medium sized train set of the TABLET dataset. It contains the train examples for all TABLET tasks.Each task is capped at 140,000 examples, resulting in a total of 1,117,217 training examples across 17 tasks.This dataset is self-contained, each example includes a table image, its HTML representation, and the associated task data.However, if you're interested in downloading just the TABLET tables, check out TABLET-tables.
All TABLET Subsets:
(train)… See the full description on the dataset page: https://huggingface.co/datasets/alonsoapp/TABLET-Medium.clean-up-tableThis dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v2.1",
"robot_type": "franka",
"total_episodes": 51,
"total_frames": 28867,
"total_tasks": 1,
"total_videos": 0,
"total_chunks": 1,
"chunks_size": 1000,
"fps": 20,
"splits": {
"train": "0:51"
},
"data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet",
"video_path": null,
"features": {… See the full description on the dataset page: https://huggingface.co/datasets/danielsanjosepro/clean-up-table.table-vqa
Dataset description
The table-vqa Dataset integrates images of tables from the dataset AFTdb (Arxiv Figure Table Database) curated by cmarkea.
This dataset consists of pairs of table images and corresponding LaTeX source code, with each image linked to an average of ten questions and answers. Half of the Q&A pairs are in English and the other half in French. These questions and answers were generated using Gemini 1.5 Pro and Claude 3.5 sonnet, making the dataset well-suited for… See the full description on the dataset page: https://huggingface.co/datasets/cmarkea/table-vqa.TableVQA-Bench
Dataset Card for "TableVQA-Bench"
More Information needed
Visual-TableQA
🧠 Visual-TableQA: Open-Domain Benchmark for Reasoning over Table Images
Welcome to Visual-TableQA, a project designed to generate high-quality synthetic question-answer datasets associated to images of tables. This resource is ideal for training and evaluating models on visually-grounded table understanding tasks such as document QA, table parsing, and multimodal reasoning.
🚀 Latest Update
We have refreshed the dataset with newly generated QA pairs created by… See the full description on the dataset page: https://huggingface.co/datasets/AI-4-Everyone/Visual-TableQA.TABLET-Small
TABLET-Small
This is the Small sized train set of the TABLET dataset. It contains the train examples for 14 TABLET tasks.Each task is capped at 140,000 examples, resulting in a total of 776,602 training examples across 14 tasks.This dataset is self-contained, each example includes a table image, its HTML representation, and the associated task data.However, if you're interested in downloading just the TABLET tables, check out TABLET-tables.
All TABLET Subsets:
(train)… See the full description on the dataset page: https://huggingface.co/datasets/alonsoapp/TABLET-Small.table_spill_cleanup_bimanual_rgbd_segmentation_poses
Exylos Bimanual Table Spill Cleanup Rich-Modality Sample
A compact, rich-modality bimanual robot manipulation dataset for tabletop spill cleanup.
Each episode combines synchronized dual-arm Panda state/action trajectories, 7 RGB camera streams, per-frame depth maps, per-frame segmentation masks, object pose streams, phase annotations, and an objective cleanup success metric based on the remaining spill fraction.
This dataset is a rich-modality inspection sample for the Exylos… See the full description on the dataset page: https://huggingface.co/datasets/ExylosAi/table_spill_cleanup_bimanual_rgbd_segmentation_poses.text_table_md_v0table-image-html-pairsnutrition-table-detection
Open Food Facts Nutrition table detection dataset
This dataset was used to train the nutrition table object detection model running in production at Open Food Facts.
Images were collected from the Open Food Facts database and labeled manually.
Just like the original images, the images in this dataset are licensed under the Creative Commons Attribution Share Alike license (CC-BY-SA 3.0).
Fields
image_id: Unique identifier for the image, generated from the barcode and… See the full description on the dataset page: https://huggingface.co/datasets/openfoodfacts/nutrition-table-detection.table-vqa_beirThis is a copy of https://huggingface.co/datasets/jinaai/table-vqa reformatted into the BEIR format. For any further information like license, please refer to the original dataset.
Disclaimer
This dataset may contain publicly available images or text data. All data is provided for research and educational purposes only. If you are the rights holder of any content and have concerns regarding intellectual property or copyright, please contact us at "support-data (at) jina.ai" for… See the full description on the dataset page: https://huggingface.co/datasets/jinaai/table-vqa_beir.09012025_clean_table_closerclosercamThis dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v2.1",
"robot_type": null,
"total_episodes": 50,
"total_frames": 9169,
"total_tasks": 1,
"total_videos": 0,
"total_chunks": 1,
"chunks_size": 1000,
"fps": 30,
"splits": {
"train": "0:50"
},
"data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet",
"video_path": null,
"features": {… See the full description on the dataset page: https://huggingface.co/datasets/rli14/09012025_clean_table_closerclosercam.armanual-dinner-table
armanual — bimanual SO-101 dinner-table demonstrations
Demonstrations for Bimanual VLA Manipulation with Multi-Modal Reasoning (Intel Physical AI
Online Challenge). Two simulated SO-101 arms set a dinner table in MuJoCo: placing plates, cups
and cutlery, opening a drawer to reach the cutlery, handing objects between arms, and pouring.
Code: https://github.com/pythonsniffer/armanual
What is in it
Episodes
114
Frames
22,206 (20 Hz)
Distinct… See the full description on the dataset page: https://huggingface.co/datasets/pythonsniffer/armanual-dinner-table.TABLET-testwipe_tableThis dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v2.1",
"robot_type": "Stretch",
"total_episodes": 40,
"total_frames": 11145,
"total_tasks": 2,
"total_videos": 0,
"total_chunks": 1,
"chunks_size": 1000,
"fps": 10,
"splits": {
"train": "0:40"
},
"data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/AnranZZ/wipe_table.pubmed-tables-latex-768pxfruit-on-table-recognition-for-sorting-training-50500a2b-f18369e7
Red Fruit Detection for Sorting
Computer vision training dataset of 30 renders (640x640, morning lighting, RGB + albedo pass, with bounding-box and label annotations) staging red-colored fruits in restaurant, cafe and lounge environments, built to train a YOLO model to detect red fruits for sorting.
This dataset mirrors public data-pack render outputs from Physicl.
Each row represents one render view. The image column contains a stable URL to the primary render image uploaded… See the full description on the dataset page: https://huggingface.co/datasets/physicl-community/fruit-on-table-recognition-for-sorting-training-50500a2b-f18369e7.fruit-on-table-recognition-for-sorting-next-pack-9a889d4d-aee688d0
Outdoor Fruit Detection & Sorting
Training dataset of 100 outdoor scenes (market squares and an alleyway) framing fruit among stalls and produce, at 640x640 with albedo, material index, world normals and metric depth passes plus per-frame annotations, to train a YOLOv8 model to detect, classify and count fruit types for eventual sorting.
This dataset mirrors public data-pack render outputs from Physicl.
Each row represents one render view. The image column contains a stable URL… See the full description on the dataset page: https://huggingface.co/datasets/physicl-community/fruit-on-table-recognition-for-sorting-next-pack-9a889d4d-aee688d0.08232025_clean_tableThis dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v2.1",
"robot_type": null,
"total_episodes": 0,
"total_frames": 0,
"total_tasks": 0,
"total_videos": 0,
"total_chunks": 0,
"chunks_size": 1000,
"fps": 30,
"splits": {},
"data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/rli14/08232025_clean_table.table-row-yolo-datasetlightonocr-pubtables-table-onlyminivlm-tablevqa-sft
minivlm-tablevqa-sft
Supervised fine-tuning mixture for improving visual table question answering in
small vision-language models — built for the MiniVLMDocEval project to lift
Qwen3.5-0.8B on TableVQABench.
Why this mixture (data-driven targeting)
We measured Qwen3.5-0.8B per TableVQABench sub-domain and found the weakness is
Wikipedia-style visual-table lookup, not financial tables:
sub-domain
Qwen3.5-0.8B
vwtq (Wikipedia lookup)
27.8
weakest, and… See the full description on the dataset page: https://huggingface.co/datasets/savoji/minivlm-tablevqa-sft.arocrbench_tablesPlease see paper & code for more information:
https://github.com/mbzuai-oryx/KITAB-Bench
https://arxiv.org/abs/2502.14949
ViRL39K-Tables-Diagrams-ChartsTableSynth-semTABLET-dev
TABLET-dev
This is the dev set of the TABLET dataset. It contains the development/validation examples for all TABLET tasks.This dataset is self-contained, each example includes a table image, its HTML representation, and the associated task data, you don't need to download anything else to use it.However, if you're interested in downloading just the TABLET tables, check out TABLET-tables.
All TABLET Subsets:
(train) TABLET-Small: The smallest TABLET subset, including 776,602… See the full description on the dataset page: https://huggingface.co/datasets/alonsoapp/TABLET-dev.Viet-Table-MarkdownTableBench-V-test
