datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
cua_debugger_traj
CUA Debugger Trajectories
204 failed computer-use agent (CUA) trajectories on OSWorld, each with a human root-cause annotation.
Three agents were run on OSWorld (Ubuntu desktop, screenshot-only observation, pyautogui execution at 1920×1080). Every trajectory in this dataset is a failure (no task reached evaluator score 1.0). For each trajectory, a human annotator identified the root error step — the earliest step responsible for the failure — and labeled it with an… See the full description on the dataset page: https://huggingface.co/datasets/CyT1ng/cua_debugger_traj.lib90-rotate-debugged-90This dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v2.1",
"robot_type": "panda",
"total_episodes": 90,
"total_frames": 18110,
"total_tasks": 74,
"total_videos": 0,
"total_chunks": 1,
"chunks_size": 1000,
"fps": 10,
"splits": {
"train": "0:90"
},
"data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/windfromthenorth/lib90-rotate-debugged-90.rotate90-debugged-270This dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v2.1",
"robot_type": "panda",
"total_episodes": 90,
"total_frames": 17924,
"total_tasks": 74,
"total_videos": 0,
"total_chunks": 1,
"chunks_size": 1000,
"fps": 10,
"splits": {
"train": "0:90"
},
"data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/windfromthenorth/rotate90-debugged-270.debug_MMMU_mcq_to_remove
Dataset Card for "debug_MMMU_mcq_to_remove"
More Information needed
rotate90-debugged-90This dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v2.1",
"robot_type": "panda",
"total_episodes": 90,
"total_frames": 20338,
"total_tasks": 74,
"total_videos": 0,
"total_chunks": 1,
"chunks_size": 1000,
"fps": 10,
"splits": {
"train": "0:90"
},
"data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/windfromthenorth/rotate90-debugged-90.debug_MathVista_open_ended_to_remove
Dataset Card for "debug_MathVista_open_ended_to_remove"
More Information needed
crag-mm-single-turn-debug-public
CRAG-MM: Comprehensive multi-modal, multi-turn RAG Benchmark
This repository contains the CRAG-MM dataset, a high-quality conversational benchmark for multimodal assistants. The dataset features conversations about images with varied complexity levels, designed to evaluate AI systems' visual understanding and conversational abilities.
CRAG-MM is a visual question-answering benchmark that focuses on factual questions, offering a unique collection of image and question-answering sets… See the full description on the dataset page: https://huggingface.co/datasets/crag-mm-2025/crag-mm-single-turn-debug-public.rotate90-debugged-270-actualThis dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v2.1",
"robot_type": "panda",
"total_episodes": 270,
"total_frames": 58702,
"total_tasks": 74,
"total_videos": 0,
"total_chunks": 1,
"chunks_size": 1000,
"fps": 10,
"splits": {
"train": "0:270"
},
"data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/windfromthenorth/rotate90-debugged-270-actual.debug_MathVista_mcq_to_remove
Dataset Card for "debug_MathVista_mcq_to_remove"
More Information needed
calvin-debug-lerobotdebug_MMMU_open_ended_to_remove
Dataset Card for "debug_MMMU_open_ended_to_remove"
More Information needed
test_debug_editdata_2A part remove dataset of the any edit dataset
copy from any edit dataset to debug quickly:
cite:
https://huggingface.co/datasets/Bin1117/AnyEdit
https://huggingface.co/datasets/Bin1117/anyedit-split
@article{yu2024anyedit,
title={AnyEdit: Mastering Unified High-Quality Image Editing for Any Idea},
author={Yu, Qifan and Chow, Wei and Yue, Zhongqi and Pan, Kaihang and Wu, Yang and Wan, Xiaoyang and Li, Juncheng and Tang, Siliang and Zhang, Hanwang and Zhuang, Yueting},
journal={arXiv… See the full description on the dataset page: https://huggingface.co/datasets/laulampaul/test_debug_editdata_2.test_debug_editdataA part replace dataset of the any edit dataset
copy from any edit dataset to debug quickly https://huggingface.co/datasets/Bin1117/AnyEdit
cc12m_openai-clip-vit-patch32_image_retrieval_top4_start1000000_end3000000_DEBUGcalvin_debug_lerobo_formatThis dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v2.0",
"robot_type": "panda",
"total_episodes": 9,
"total_frames": 503,
"total_tasks": 9,
"total_videos": 0,
"total_chunks": 1,
"chunks_size": 1000,
"fps": 10,
"splits": {
"train": "0:9"
},
"data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/ducido/calvin_debug_lerobo_format.crag-mm-multi-turn-debug-public
CRAG-MM: Comprehensive multi-modal, multi-turn RAG Benchmark
This repository contains the CRAG-MM dataset, a high-quality conversational benchmark for multimodal assistants. The dataset features conversations about images with varied complexity levels, designed to evaluate AI systems' visual understanding and conversational abilities.
CRAG-MM is a visual question-answering benchmark that focuses on factual questions, offering a unique collection of image and question-answering sets… See the full description on the dataset page: https://huggingface.co/datasets/crag-mm-2025/crag-mm-multi-turn-debug-public.flickr-audio-image-debugrs_debugThis dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v3.0",
"robot_type": "so101_follower",
"total_episodes": 10,
"total_frames": 1096,
"total_tasks": 1,
"chunks_size": 1000,
"data_files_size_in_mb": 100,
"video_files_size_in_mb": 500,
"fps": 30,
"splits": {
"train": "0:10"
},
"data_path": "data/chunk-{chunk_index:03d}/file-{file_index:03d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/msmandelbrot/rs_debug.jxu124_refcoco_debugdebug-rotationThis dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v2.1",
"robot_type": "panda",
"total_episodes": 29,
"total_frames": 6742,
"total_tasks": 23,
"total_videos": 0,
"total_chunks": 1,
"chunks_size": 1000,
"fps": 10,
"splits": {
"train": "0:29"
},
"data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/windfromthenorth/debug-rotation.debug
license: apache-2.0
vis-debug_iip_rl-judgeAndroidControl_debug3x-libero-drawer-test-debugThis dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v2.1",
"robot_type": "panda",
"total_episodes": 150,
"total_frames": 300,
"total_tasks": 1,
"total_videos": 0,
"total_chunks": 1,
"chunks_size": 1000,
"fps": 10,
"splits": {
"train": "0:150"
},
"data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/windfromthenorth/3x-libero-drawer-test-debug.rs_debug2This dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v3.0",
"robot_type": "so101_follower",
"total_episodes": 14,
"total_frames": 1999,
"total_tasks": 1,
"chunks_size": 1000,
"data_files_size_in_mb": 100,
"video_files_size_in_mb": 500,
"fps": 30,
"splits": {
"train": "0:14"
},
"data_path": "data/chunk-{chunk_index:03d}/file-{file_index:03d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/msmandelbrot/rs_debug2.vis-debug_iip_sft-judgelaion_4_to_3_debugsam3-segment-test-debug
Image Segmentation: Animal using SAM3
This dataset contains semantic segmentation maps for animal segmented in images from davanstrien/ena24-detection using Meta's SAM3.
Generated using: uv-scripts/sam3 segmentation script
Statistics
Objects Segmented: animal
Total Instances: 8
Images with Detections: 5 / 5 (100.0%)
Average Instances per Image: 1.60
Output Format: semantic-mask
Processing Details
Source Dataset: davanstrien/ena24-detection
Model:… See the full description on the dataset page: https://huggingface.co/datasets/davanstrien/sam3-segment-test-debug.gia-dataset-parquet-debug
Dataset Card for "gia-dataset-parquet-debug"
More Information needed
jxu124_refcocoplus_debug
