datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
VLNVerse_sceneVLNVerse_dataVLN_Dataset_2This repository contains encrypted visual features for an ongoing academic research project. Decryption keys are managed internally for reproducibility.
SAGE-3D_VLN_Data
SAGE-3D VLN Data: Vision-Language Navigation Dataset with Hierarchical Instructions
Paper | Project Page | Code
A comprehensive VLN dataset featuring 2 million trajectory-instruction pairs across 1,000 indoor scenes, with hierarchical instruction design covering high-level semantic goals to low-level control commands.
Overview of SAGE-3D VLN Data. SAGE-3D VLN Data includes a hierarchical instruction, and two major task types (VLN + No-goal).
📢 News… See the full description on the dataset page: https://huggingface.co/datasets/spatialverse/SAGE-3D_VLN_Data.AGC-VLN-Town10HD-100episodes
AGC-VLN — Town10HD 100-Episode Results
The 100 closed-loop evaluation runs of AGC-VLN (Air-Ground Collaborative
Vision-and-Language Navigation via Shared Bird's-Eye Maps) on the CARLA-Air
Town10HD scene.
Overview
100 episodes = 50 scenes × 2 runs each.
Aggregated metrics: joint success 77.0%, UAV success 50.0%, UGV success
75.0%, collaboration gain +27.0% (CG = SR_joint − min(SR_uav, SR_ugv)).
Success threshold: either agent within 5 m of the goal.
Each run… See the full description on the dataset page: https://huggingface.co/datasets/Shuning1997/AGC-VLN-Town10HD-100episodes.Rule-VLN
Rule-VLN Dataset
Rule-VLN is a rule-compliant outdoor vision-and-language navigation benchmark built on the Touchdown / StreetLearn urban navigation environment. It studies whether navigation agents can follow language instructions while also complying with semantic traffic rules, such as regulatory signs that prohibit otherwise reachable movements.
This dataset accompanies the paper:
Rule-VLN: Bridging Perception and Compliance via Semantic Reasoning and Geometric… See the full description on the dataset page: https://huggingface.co/datasets/jeffry77/Rule-VLN.Matterport3D-Scansvln_n1_trainMP3D_marked_obsvlnverse_sceneMP3D_featureR2R_RxR_VLNCE_preprocessedLHPR-VLN
Long-Horizon Planning and Reasoning in VLN (LHPR-VLN) Benchmark
This repository contains the long-horizon planning and reasoning in VLN (LHPR-VLN) benchmark, introduced in Towards Long-Horizon Vision-Language Navigation: Platform, Benchmark and Method.
Dataset is also available in ModelScope.
Files
The trajectory dataset file organization is as follows. The task/ directory contains the complete LH-VLN task trajectories, while the step_task/… See the full description on the dataset page: https://huggingface.co/datasets/Starry123/LHPR-VLN.vln4navidUAV-VLN-FOVVLN-Ego-making
Project Page: VLN-R1: Vision-Language Navigation via Reinforcement Fine-Tuning
This repository demonstrates how to create the VLN-Ego dataset using VLN-CE. The specific code can be found in VLN-Ego-making.
Before proceeding, you need to download the content from this Hugging Face resource to create VLN-Ego. It mainly includes necessary downloads such as checkpoints and configuration files for VLN-CE.
这个repo展示了如何使用VLN-CE制作VLN-Ego数据集,具体代码在VLN-Ego-making。
在这之前,您需要下载这个hugging… See the full description on the dataset page: https://huggingface.co/datasets/alexzyqi/VLN-Ego-making.SID-VLN Datasets of Learning Goal-Oriented Language-Guided Navigation with Self-Improving Demonstrations at Scale.
vln_r2r_rxr_hfov90VLN-Ego
Project Page: VLN-R1: Vision-Language Navigation via Reinforcement Fine-Tuning
This is the training dataset of VLN-R1. It is collected from R2R and RxR, and can be directly used for training with LLaMA-Factory.
We have also released the process of making VLN-Ego using VLN-CE, with the code available at VLN-Ego-making.
这是VLN-R1的训练数据集。它源自R2R和RxR,可直接用于LLaMA-Factory的训练。
我们也公布了使用VLN-CE制作VLN-Ego的过程,相关代码位于VLN-Ego-making。
FloorPlan-VLN-RxRVLN-CEVLN-CE-R2R_easi
VLN-CE R2R Dataset for EASI
Vision-and-Language Navigation in Continuous Environments (VLN-CE) Room-to-Room
(R2R) benchmark, repackaged for the EASI
evaluation framework.
Task
An agent receives a natural language navigation instruction and must navigate
through a Matterport3D indoor environment to reach a goal location. The agent
uses discrete actions: STOP, MOVE_FORWARD (0.25m), TURN_LEFT (15 deg),
TURN_RIGHT (15 deg).
Success is measured when the agent stops within 3.0m… See the full description on the dataset page: https://huggingface.co/datasets/oscarqjh/VLN-CE-R2R_easi.VLN-CE-IsaacVLNCE_datavln_n1_tensorsOutdoor_VLNVLN_annotationsHA-VLN🚀🚀🚀
HAPS Dataset 2.0
In real-world scenarios, human motion typically adapts and interacts with the surrounding region. The proposed Human Activity and Pose Simulation (HAPS) Dataset 2.0 improves upon HAPS 1.0 by making the following enhancements:
Refining and diversifying human motions.
Providing descriptions closely tied to region awareness.
HAPS 2.0 mitigates the limitations of existing human motion datasets by identifying 26 distinct regions across 90 architectural… See the full description on the dataset page: https://huggingface.co/datasets/fly1113/HA-VLN.mas-vln-isaac-rgbd
MAS-VLN Randomized Warehouse RGBD
This dataset contains Isaac Sim randomized warehouse multi-robot rollouts packaged
for Hugging Face release. RGB and metric depth frames are stored inside one plain
tar file per rollout. Metadata tables index scenes, rollouts, and rendered camera
frames.
Release
Dataset version: v0.1.1
Created at: 2026-05-14T07:37:38.436998+00:00
Isaac Sim: 5.1.0
ROS distro: humble
Scenes: 25
Rollouts: 121
Frame rows: 890572
Layout… See the full description on the dataset page: https://huggingface.co/datasets/yang-jiao/mas-vln-isaac-rgbd.VLN-CE_Habitat_Images
