mateoguaman/vamos_10pct_gpt5_mini
vamos_10pct_gpt5_mini Description VLN Navigation dataset with 100% of tartandrive data, 50% of scand data, 25% of coda data, 100% of in-domain spot data, and 10% of annotated/augmented data using gpt5-mini. Whenever daatsets aren't 100%, they are ranked by curvature and output of length 5. Processing Parameters {} Dataset Configuration Train dataset: mixer: mateoguaman/vlmn_tartandrive100_scand50_coda25_spot100_sub5: 1.0… See the full description on the dataset page: https://huggingface.co/datasets/mateoguaman/vamos_10pct_gpt5_mini.
Upload README.md with huggingface_hub
Create consolidated dataset from cached datasets (part 00006-of-00007)
Create consolidated dataset from cached datasets (part 00005-of-00007)
Create consolidated dataset from cached datasets (part 00004-of-00007)
Create consolidated dataset from cached datasets (part 00003-of-00007)
Create consolidated dataset from cached datasets (part 00002-of-00007)
Create consolidated dataset from cached datasets (part 00001-of-00007)
Create consolidated dataset from cached datasets (part 00000-of-00007)
initial commit
