VAMOS
Datasets
All datasets matching “VAMOS”vamos_25pct_gpt5_nano
vamos_25pct_gpt5_nano
Description
VLN Navigation dataset with 100% of tartandrive data, 50% of scand data, 25% of coda data, 100% of in-domain spot data, and 25% of annotated/augmented data using gpt5-nano. Whenever daatsets aren't 100%, they are ranked by curvature and output of length 5.
Processing Parameters
{}
Dataset Configuration
Train dataset:
mixer: mateoguaman/vlmn_tartandrive100_scand50_coda25_spot100_sub5: 1.0… See the full description on the dataset page: https://huggingface.co/datasets/mateoguaman/vamos_25pct_gpt5_nano.vamos_dataset
vamos_dataset
Description
VAMOS Navigation Dataset (with visual-language co-training data) combines multiple publicly available datasets and in-domain Spot data collected by the authors.This version includes navigation data as well as co-training data from COCO-QA and Localized Narratives.
The dataset includes:
100% of TartanDrive 2 data
50% of SCAND data
25% of CODa data
100% of in-domain Spot data (collected by the authors)
COCO-QA and Localized Narratives… See the full description on the dataset page: https://huggingface.co/datasets/mateoguaman/vamos_dataset.vamos_10pct_gpt5_mini
vamos_10pct_gpt5_mini
Description
VLN Navigation dataset with 100% of tartandrive data, 50% of scand data, 25% of coda data, 100% of in-domain spot data, and 10% of annotated/augmented data using gpt5-mini. Whenever daatsets aren't 100%, they are ranked by curvature and output of length 5.
Processing Parameters
{}
Dataset Configuration
Train dataset:
mixer: mateoguaman/vlmn_tartandrive100_scand50_coda25_spot100_sub5: 1.0… See the full description on the dataset page: https://huggingface.co/datasets/mateoguaman/vamos_10pct_gpt5_mini.vamos_25pct_traj_25pct_atraj_50pct_anno
vamos_25pct_traj_25pct_atraj_50pct_anno
Description
VLN Navigation dataset with 100% of tartandrive data, 50% of scand data, 25% of coda data, 100% of in-domain spot data, and 10% of annotated/augmented data using gpt5-mini. Whenever daatsets aren't 100%, they are ranked by curvature and output of length 5.
Processing Parameters
{}
Dataset Configuration
Train dataset:
mixer: mateoguaman/annotations_only_10pct_gpt5_mini: 0.5… See the full description on the dataset page: https://huggingface.co/datasets/mateoguaman/vamos_25pct_traj_25pct_atraj_50pct_anno.vamos_navigation_only_dataset
vamos_navigation_only_dataset
Description
VAMOS navigation-only (without natural language preference annotations nor co-training VQA data) dataset with 100% of TartanDrive 2 data, 50% of SCAND data, 25% of CODa data, and 100% of in-domain spot data. Whenever daatsets aren't 100%, they are ranked by curvature and output of length 5.
License
This dataset is released under the Creative Commons Attribution-NonCommercial-ShareAlike 4.0 International License (CC… See the full description on the dataset page: https://huggingface.co/datasets/mateoguaman/vamos_navigation_only_dataset.finance_emotions
Citation
Please cite the following if you use this data:
Vamossy, Domonkos F., and Rolf Skog. "EmTract: Extracting Emotions from Social Media." Available at SSRN 3975884 (2023).
BibTex citation:
@article{vamossy2023emtract,
title={EmTract: Extracting Emotions from Social Media},
author={Vamossy, Domonkos F and Skog, Rolf},
journal={Available at SSRN 3975884},
year={2023}
}
