datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
olmoearth_pretrain_datasetThis is the pre-training dataset for training the OlmoEarth pre-trained remote sensing foundation models.
Documentation is on GitHub at https://github.com/allenai/olmoearth_pretrain/blob/main/docs/Pretraining-Dataset.md
The dataset is released under CC BY 4.0. It includes data from the following sources:
Sentinel-2 L2A imagery from the European Space Agency, available under the Copernicus Sentinel Data and Service Legal Notice
Sentinel-1 GRD IW vv+vh imagery from the European Space Agency… See the full description on the dataset page: https://huggingface.co/datasets/allenai/olmoearth_pretrain_dataset.Molmo2-ER-RoboPoint
Molmo2-ER · wentao-yuan/robopoint-data
1.43M robotics affordance instruction-tuning examples (pointing + detection + VQA).
This is a re-hosted, loader-ready subset of the upstream dataset, used to train allenai/Molmo2-ER-4B. Files mirror the upstream layout; nothing in the data has been modified.
Upstream source
Original dataset: wentao-yuan/robopoint-data
Paper: RoboPoint: A Vision-Language Model for Spatial Affordance Prediction for Robotics (arXiv:2406.10721)
License:… See the full description on the dataset page: https://huggingface.co/datasets/allenai/Molmo2-ER-RoboPoint.Molmo2-ER-RefSpatial
Molmo2-ER · JingkunAn/RefSpatial
2.5M spatial-referring corpus (web + indoor + simulated) covering 31 spatial relations.
This is a re-hosted, loader-ready subset of the upstream dataset, used to train allenai/Molmo2-ER-4B. Files mirror the upstream layout; nothing in the data has been modified.
Upstream source
Original dataset: JingkunAn/RefSpatial
Paper: RoboRefer: Towards Spatial Referring with Reasoning in Vision-Language Models for Robotics (arXiv:2506.04308)… See the full description on the dataset page: https://huggingface.co/datasets/allenai/Molmo2-ER-RefSpatial.wds_vtab-clevr_count_allall2024speedtest_1all2024speedtest_0all2024speedtest_7all2024speedtest_12olmoearth_projects_awfThis model is for fine-tuning OlmoEarth-v1-Base to map land use and land cover in southern Kenya.
The data is annotated by experts at the African Wildlife Foundation.
For more details, please see the documentation on GitHub at https://github.com/allenai/olmoearth_projects/blob/main/docs/awf.md
all2024speedtest_5all2024speedtest_2all2024speedtest_13all2024speedtest_3all2024speedtest_19all2024speedtest_23all2024speedtest_6all2024speedtest_18all2024speedtest_17all2024speedtest_16all2024speedtest_14all2024speedtest_25all2024speedtest_45all2024speedtest_9all2024speedtest_33all2024speedtest_15all2024speedtest_39all2024speedtest_50all2024speedtest_8wds_vtab-clevr_count_all_test
CLEVR Count All Webdataset (Test set only)
Original paper: CLEVR: A Diagnostic Dataset for Compositional Language and Elementary Visual Reasoning
Homepage: https://cs.stanford.edu/people/jcjohns/clevr/
Bibtex:
@article{DBLP:journals/corr/JohnsonHMFZG16,
author = {Justin Johnson and
Bharath Hariharan and
Laurens van der Maaten and
Li Fei{-}Fei and
C. Lawrence Zitnick and
Ross B. Girshick},
title =… See the full description on the dataset page: https://huggingface.co/datasets/djghosh/wds_vtab-clevr_count_all_test.all2024speedtest_30
