datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
data_verticalrandomsorange-pick-testsoft_orangeThis dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v2.1",
"robot_type": "soft",
"total_episodes": 50,
"total_frames": 8002,
"total_tasks":1,
"total_videos": 0,
"total_chunks": 1,
"chunks_size": 1000,
"fps": 5,
"splits": {
"train": "0:50"
},
"data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/HCSuMoss/soft_orange.OR_anonymizationlab_data_orange_cube_single_point_paired_25This dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v2.1",
"robot_type": "panda",
"total_episodes": 50,
"total_frames": 12001,
"total_tasks": 1,
"total_videos": 0,
"total_chunks": 1,
"chunks_size": 1000,
"fps": 15,
"splits": {
"train": "0:50"},
"data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/ceilingfan456/lab_data_orange_cube_single_point_paired_25.tictac-pick-orange-cubeVisualReasoner-30k
Dataset Card for VisualReasoner-30k
Dataset Details
This dataset is an extension of VisualReasoner-1M, containing approximately 30k cases and can be used for training visual reasoning tasks.
Unlike VisualReasoner-1M, this dataset models the reasoning process in an end-to-end format to better accommodate scenarios where explicit tool invocation is not allowed.
Dataset Descriptions
The structure of each case is as follows:
{
"identity": "Case ID"… See the full description on the dataset page: https://huggingface.co/datasets/orange-sk/VisualReasoner-30k.libero_unlearned_orange_juiceThis dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v3.0",
"robot_type": "panda",
"total_episodes": 1693,
"total_frames": 273465,
"total_tasks": 40,
"chunks_size": 1000,
"data_files_size_in_mb": 100,
"video_files_size_in_mb": 200,
"fps": 10.0,
"splits": {
"train": "0:1693"
},
"data_path": "data/chunk-{chunk_index:03d}/file-{file_index:03d}.parquet",
"video_path": null… See the full description on the dataset page: https://huggingface.co/datasets/leonardo-russo/libero_unlearned_orange_juice.droid_pnp_carrot_orangeplateThis dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v2.0",
"robot_type": "panda",
"total_episodes": 49,
"total_frames": 11910,
"total_tasks": 1,
"total_videos": 0,
"total_chunks": 1,
"chunks_size": 1000,
"fps": 15,
"splits": {
"train": "0:49"
},
"data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/jaehyeokdoo2/droid_pnp_carrot_orangeplate.Gen-nuScenesorange_on_greenbackupsoarm101_pickplace_orange_060e_tw_closedThis dataset was created using LeRobot.
Dataset Description
Pick-and-place dataset of 60 episodes. Single orange cube placed at varied positions across the workspace.
The objective in each episode is to approach, grasp, and relocate the cube to a designated placement area.
Pickup strategy: Robot rotates toward the target cube and approaches with the gripper closed. In this dataset, because of the
robot's wrist orientation, the camera is tilted 90deg, and the frontal view can only… See the full description on the dataset page: https://huggingface.co/datasets/edgarcancinoe/soarm101_pickplace_orange_060e_tw_closed.VisualReasoner-1M
Dataset Card for VisualReasoner-1M
Dataset Details
This is a dataset for the paper From the Least to the Most: Building a Plug-and-Play Visual Reasoner via Data Synthesis. The dataset contains approximately 1 million cases and can be used for training visual reasoning tasks. The reasoning process involves breaking down tasks and utilizing tools to solve complex and challenging visual question-answering tasks progressively.
For detailed data synthesis methods, please… See the full description on the dataset page: https://huggingface.co/datasets/orange-sk/VisualReasoner-1M.so101-orange-smokeorange-bowl-in-purple-bowl-rl-npz
orange-bowl-in-purple-bowl RL npz — dataset card
작성: 2026-09-16 (Claude, Opus). Bigenlight/carrot-in-pot-rl-npz의 자매 문서 — 같은 파이프라인(build_dataset.py / build_px_dataset.py)으로 orange-bowl 실물 코퍼스를 chunk-MDP transition으로 변환한 것. raw 영상은 Bigenlight/orange_bowl_in_purple_bowl_lerobot_v3에 있고, 여기엔 파생 npz(W/data/rl/)만 있음. 모델은 Bigenlight/orange-bowl-in-purple-bowl-ifql.
수치는 npz meta + .shape(mmap)와 결과 노트에서 읽음. 출처 file:line은 cards/CARD_SOURCES.md §2026-09-16 "orange".
0. 소스… See the full description on the dataset page: https://huggingface.co/datasets/Bigenlight/orange-bowl-in-purple-bowl-rl-npz.so100_orange_50ep_2cam_2-trajectoryThis dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v3.0",
"robot_type": "so100",
"total_episodes": 50,
"total_frames": 14965,
"total_tasks": 1,
"chunks_size": 1000,
"data_files_size_in_mb": 100,
"video_files_size_in_mb": 200,
"fps": 30,
"splits": {
"train": "0:50"
},
"data_path": "data/chunk-{chunk_index:03d}/file-{file_index:03d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/edmos7/so100_orange_50ep_2cam_2-trajectory.orange-juice-ad
Dataset Card for "orange-juice-ad"
More Information needed
continuous-baseline-review-dataffw_isaacsim_orangeThis dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v2.1",
"robot_type": "ffw",
"total_episodes": 1,
"total_frames": 144,
"total_tasks": 1,
"total_videos": 0,
"total_chunks": 1,
"chunks_size": 1000,
"fps": 10,
"splits": {
"train": "0:1"
},
"data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/PatoTMR/ffw_isaacsim_orange.imnet1k_orangeOrandCar_RS_think
OrandCar — OrandCar_RS_think
Rejection-sampled from the OrandCar train split. This split holds the accepted items, with the model's reasoning trace.
rows
1,576
QA pairs
1,576
shards
1
accepted / rejected (whole family)
1,576 / 445
accept rate
78.0%
verifier
alnum
How the data was produced
A VLM answers every question at temperature 0 with reasoning enabled. Its answer is compared with
the official ground truth by the verifier described… See the full description on the dataset page: https://huggingface.co/datasets/elliot-mllm/OrandCar_RS_think.Data_Orange_2_Leaves_augimnet1k_orangutan_orang_orangutang_Pongo_pygmaeusOrandCar_RS_nothink
OrandCar — OrandCar_RS_nothink
Rejection-sampled from the OrandCar train split. This split holds the accepted items, answer only.
rows
1,576
QA pairs
1,576
shards
1
accepted / rejected (whole family)
1,576 / 445
accept rate
78.0%
verifier
alnum
How the data was produced
A VLM answers every question at temperature 0 with reasoning enabled. Its answer is compared with
the official ground truth by the verifier described below; matches go… See the full description on the dataset page: https://huggingface.co/datasets/elliot-mllm/OrandCar_RS_nothink.OrandCar_rejected
OrandCar — OrandCar_rejected
Rejection-sampled from the OrandCar train split. This split holds the rejected items — the answer field holds the official ground truth.
rows
445
QA pairs
445
shards
1
accepted / rejected (whole family)
1,576 / 445
accept rate
78.0%
verifier
alnum
The rejected split is training data, not just diagnostics: answer is the official ground truth, and wrong_vlm records what the model said instead.
How the data was… See the full description on the dataset page: https://huggingface.co/datasets/elliot-mllm/OrandCar_rejected.Orange-ou-tomatelab_data_orange_cube_single_pointThis dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v2.1",
"robot_type": "panda",
"total_episodes": 25,
"total_frames": 4041,
"total_tasks": 1,
"total_videos": 0,
"total_chunks": 1,
"chunks_size": 1000,
"fps": 15,
"splits": {
"train": "0:25"
},
"data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/ceilingfan456/lab_data_orange_cube_single_point.Vico_orangeopenpi_mo-aloha_sjj_orangeThis dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v2.0",
"robot_type": "aloha",
"total_episodes": 51,
"total_frames": 51000,
"total_tasks": 1,
"total_videos": 0,
"total_chunks": 1,
"chunks_size": 1000,
"fps": 50,
"splits": {
"train": "0:51"
},
"data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/virtualkss/openpi_mo-aloha_sjj_orange.
