mateoguaman/vamos_50pct_traj_25pct_atraj_25pct_anno
vamos_50pct_traj_25pct_atraj_25pct_anno Description VLN Navigation dataset with 100% of tartandrive data, 50% of scand data, 25% of coda data, 100% of in-domain spot data, and 10% of annotated/augmented data using gpt5-mini. Whenever daatsets aren't 100%, they are ranked by curvature and output of length 5. Processing Parameters {} Dataset Configuration Train dataset: mixer: mateoguaman/annotations_only_10pct_gpt5_mini: 0.25… See the full description on the dataset page: https://huggingface.co/datasets/mateoguaman/vamos_50pct_traj_25pct_atraj_25pct_anno.
vamos50pcttraj25pctatraj25pctanno
Description
VLN Navigation dataset with 100% of tartandrive data, 50% of scand data, 25% of coda data, 100% of in-domain spot data, and 10% of annotated/augmented data using gpt5-mini. Whenever daatsets aren't 100%, they are ranked by curvature and output of length 5.
Processing Parameters
{}
Dataset Configuration
Train dataset:
mixer: mateoguaman/annotations_only_10pct_gpt5_mini: 0.25
mateoguaman/vlmn_tartandrive100_scand50_coda25_spot100_sub5: 0.5
mateoguaman/vlmn_tartandrive100_scand50_coda25_spot100_sub5_filtered_trajectories_training_10_fixed: 0.25
split: train
Validation dataset:
mixer: mateoguaman/annotations_only_10pct_gpt5_mini: 0.25
mateoguaman/vlmn_tartandrive100_scand50_coda25_spot100_sub5: 0.5
mateoguaman/vlmn_tartandrive100_scand50_coda25_spot100_sub5_filtered_trajectories_training_10_fixed: 0.25
split: validationAdditional Information
This dataset was created by consolidating cached processed datasets from the VLM Navigation project.
