datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
EmbodiedGenDatahttps://huggingface.co/spaces/HorizonRobotics/EmbodiedGen-Gallery-Explorer
Embodied-R1.5-SFT-Dataset
Embodied-R1.5-SFT-Dataset
🌐 Project Page |
📄 arXiv |
💻 Code |
🧰 EmbodiedEvalKit |
🤗 Models & Datasets
🗓️ Update — 2026-08-20 (20260820). All 34 Stage 1 SFT JSON annotation files have been uploaded to sft_datasets_json/. The complete JSON ↔ image/video data mapping is documented in the Dataset composition table below.
⚠️ Partial release. This repository currently contains only a subset of the full Stage 1 SFT… See the full description on the dataset page: https://huggingface.co/datasets/IffYuan/Embodied-R1.5-SFT-Dataset.embodied_reasoner
Embodied-Reasoner Dataset
Dataset Overview
Embodied-Reasoner is a multimodal reasoning dataset designed for embodied interactive tasks. It contains 9,390 Observation-Thought-Action trajectories for training and evaluating multimodal models capable of performing complex embodied tasks in indoor environments.
Key Features
📸 Rich Visual Data: Contains 64,000 first-person perspective interaction images🤔 Deep Reasoning Capabilities: 8 million thought… See the full description on the dataset page: https://huggingface.co/datasets/zwq2018/embodied_reasoner.EmbodiedEvalThis repository contains the dataset of the paper EmbodiedEval: Evaluate Multimodal LLMs as Embodied Agents.
Github repository: https://github.com/thunlp/EmbodiedEval
Project Page: https://embodiedeval.github.io/
Embodied-R1.5-RFT-Dataset
Embodied-R1.5-RFT-Dataset
🌐 Project Page |
📄 arXiv |
💻 Code |
🧰 EmbodiedEvalKit |
🤗 Models & Datasets
🗓️ Update — 2026-08-20 (20260820). All 28 Stage 2 RFT JSON annotation files have been uploaded to rft_datasets_json/. The complete JSON ↔ media archive mapping is documented in the Dataset composition table below.
⚠️ Partial release. This repository currently contains only a subset of the full Stage 2 RFT data… See the full description on the dataset page: https://huggingface.co/datasets/IffYuan/Embodied-R1.5-RFT-Dataset.EmbodiedGenRLv2-BGembodied-spatial-reasoning
Embodied Spatial Reasoning Tasks
Dataset Description
This dataset is part of the embodied-spatial-reasoning project, where the agent has to actively explore the environment to determine if certain spatial relationships hold true. The tasks involve spatial reasoning with various objects and scenes. Each task includes a query about the spatial relationships between objects within a scene, which the agent must verify through exploration.
Dataset Structure
The… See the full description on the dataset page: https://huggingface.co/datasets/thanhqt2002/embodied-spatial-reasoning.vsi-benchEmbodiedRestore
EmbodiedRestore
Paired robotic first-frame observations (low-quality / ground-truth) under 25 distortions from the TID2013 / KADID-10k taxonomy, evaluated by three policies (π0.5, π0, OpenVLA). Built for benchmarking image restoration / IQA on robot-observation distributions, with downstream policy success rates(SR) and steps to successas(StS) as secondary signals.
To promote the development of image restoration model for robot vision systems, we will continue to maintain this… See the full description on the dataset page: https://huggingface.co/datasets/qruisjtu/EmbodiedRestore.BasicSpatialAbility
[ACL'25 Main] Defining and Evaluating Visual Language Models’ Basic Spatial Abilities: A Perspective from Psychometrics
[!IMPORTANT]
You can find the sample testing code on GitHub!
This dataset is a benchmark designed for evaluating Multimodal Large Language Models' Basic Spatial Abilities based on authentic Psychometric theories. It is structured specifically to support both Zero-shot and Few-shot evaluation protocols.
Split Name
Role
Description
test
Query Set… See the full description on the dataset page: https://huggingface.co/datasets/EmbodiedCity/BasicSpatialAbility.Jai-World-VRM-3D-Embodied-AI
Jai World - VRM 3D Embodied AI
Interact with an AI powered 3D vrm avatar in a virtual world. A lightweight desktop folder-based Flask app that supports both Ollama and OpenRouter.
This project explores personalized 3D world generation, embodied AI, AI spatial navigation, and virtual character interaction. This code is a working example. It's a starting point that can be tuned and expanded.
Tech stack:
Three.js + HTML + CSS + JS + Flask + Ollama/OpenRouter (qwen3.5:9b/… See the full description on the dataset page: https://huggingface.co/datasets/vbookshelf/Jai-World-VRM-3D-Embodied-AI.EmbodiedGenRLVSiQAPaperBench-X-Embodied-Manipulation-Pi05-LIBERO
PaperBench-X — pi05 / LIBERO
Standalone embodied manipulation public-demo package for the released openpi
pi05_libero checkpoint on all four LIBERO suites. This repository intentionally
contains no navigation tasks and is not the general manipulation bundle.
What is here
path
content
task/
complete Harbor task, evaluator, rubric, scoring regression tests and report builder
raw_task/
canonical source used by convert_raw_to_rp.py
demo/
static… See the full description on the dataset page: https://huggingface.co/datasets/SueMintony/PaperBench-X-Embodied-Manipulation-Pi05-LIBERO.automomaembodied_gaussiansDataset for "Physically Embodied Gaussian Splatting: A Realtime Correctable World Model for
Robotics"
Open3DVQAReinforced_Reasoning_for_Embodied_Planningembodied-spatial-resoningEmbodiedVerse-BenchObjaverseSyntheticEditableEmbodied-CoTibb-embodied-agi-notes
IBB: Intention–Body–Behavior as a Loop for Embodied Intelligence
Most stacks still treat cognition as something completed before the body executes. IBB (Intention–Body–Behavior) models intelligence as a non-divisible loop:
Intention orients — task entry, purpose, what would count as success.
Body bears — limitation, capability, physics, vulnerability.
Behavior exposes the system to the world and returns consequence as write-back.
Write-back reshapes the next cycle of intention.… See the full description on the dataset page: https://huggingface.co/datasets/AIxHAI/ibb-embodied-agi-notes.EmbodiedGenRLv2unieqaembodied_gaussiansDataset for "Physically Embodied Gaussian Splatting: A Realtime Correctable World Model for
Robotics"
Uni-EmbodiedEB-Man_environment_anchored_prior_datasetEmbodiedGenRL-articulateopen-eqa
