mateoguaman/vamos_10pct_gpt5_mini
vamos_10pct_gpt5_mini Description VLN Navigation dataset with 100% of tartandrive data, 50% of scand data, 25% of coda data, 100% of in-domain spot data, and 10% of annotated/augmented data using gpt5-mini. Whenever daatsets aren't 100%, they are ranked by curvature and output of length 5. Processing Parameters {} Dataset Configuration Train dataset: mixer: mateoguaman/vlmn_tartandrive100_scand50_coda25_spot100_sub5: 1.0… See the full description on the dataset page: https://huggingface.co/datasets/mateoguaman/vamos_10pct_gpt5_mini.
vamos10pctgpt5_mini
Description
VLN Navigation dataset with 100% of tartandrive data, 50% of scand data, 25% of coda data, 100% of in-domain spot data, and 10% of annotated/augmented data using gpt5-mini. Whenever daatsets aren't 100%, they are ranked by curvature and output of length 5.
Processing Parameters
{}
Dataset Configuration
Train dataset:
mixer: mateoguaman/vlmn_tartandrive100_scand50_coda25_spot100_sub5: 1.0
mateoguaman/vlmn_tartandrive100_scand50_coda25_spot100_sub5_filtered_trajectories_training_10: 1.0
split: train
Validation dataset:
mixer: mateoguaman/vlmn_tartandrive100_scand50_coda25_spot100_sub5: 1.0
mateoguaman/vlmn_tartandrive100_scand50_coda25_spot100_sub5_filtered_trajectories_training_10: 1.0
split: validationAdditional Information
This dataset was created by consolidating cached processed datasets from the VLM Navigation project.
