datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
olmoearth_pretrain_datasetThis is the pre-training dataset for training the OlmoEarth pre-trained remote sensing foundation models.
Documentation is on GitHub at https://github.com/allenai/olmoearth_pretrain/blob/main/docs/Pretraining-Dataset.md
The dataset is released under CC BY 4.0. It includes data from the following sources:
Sentinel-2 L2A imagery from the European Space Agency, available under the Copernicus Sentinel Data and Service Legal Notice
Sentinel-1 GRD IW vv+vh imagery from the European Space Agency… See the full description on the dataset page: https://huggingface.co/datasets/allenai/olmoearth_pretrain_dataset.Molmo2-ER-RoboPoint
Molmo2-ER · wentao-yuan/robopoint-data
1.43M robotics affordance instruction-tuning examples (pointing + detection + VQA).
This is a re-hosted, loader-ready subset of the upstream dataset, used to train allenai/Molmo2-ER-4B. Files mirror the upstream layout; nothing in the data has been modified.
Upstream source
Original dataset: wentao-yuan/robopoint-data
Paper: RoboPoint: A Vision-Language Model for Spatial Affordance Prediction for Robotics (arXiv:2406.10721)
License:… See the full description on the dataset page: https://huggingface.co/datasets/allenai/Molmo2-ER-RoboPoint.Molmo2-ER-RoboVQA
Molmo2-ER · Google DeepMind RoboVQA
Human-annotated long-horizon robotics video QA across three embodiments.
This is a re-hosted, loader-ready subset of the upstream dataset, used to train allenai/Molmo2-ER-4B. Files mirror the upstream layout; nothing in the data has been modified.
Upstream source
Original dataset: Google DeepMind RoboVQA
Paper: RoboVQA: Multimodal Long-Horizon Reasoning for Robotics (arXiv:2311.00899)
License: cc-by-4.0 (inherits from upstream)
If you… See the full description on the dataset page: https://huggingface.co/datasets/allenai/Molmo2-ER-RoboVQA.AllTheBacteria-FCGR-7merMolmo2-ER-RefSpatial
Molmo2-ER · JingkunAn/RefSpatial
2.5M spatial-referring corpus (web + indoor + simulated) covering 31 spatial relations.
This is a re-hosted, loader-ready subset of the upstream dataset, used to train allenai/Molmo2-ER-4B. Files mirror the upstream layout; nothing in the data has been modified.
Upstream source
Original dataset: JingkunAn/RefSpatial
Paper: RoboRefer: Towards Spatial Referring with Reasoning in Vision-Language Models for Robotics (arXiv:2506.04308)… See the full description on the dataset page: https://huggingface.co/datasets/allenai/Molmo2-ER-RefSpatial.wds_vtab-clevr_count_allWildDet3D-Bench
WildDet3D Benchmark
In-the-wild 3D object detection benchmark (val and test splits) from COCO, LVIS, and Objects365.
Split
Images
Annotations
Categories
Val
2,470
9,256
785
Note: The test set is held out for hidden evaluation and is not publicly available. Please submit predictions to [TODO: evaluation server] for test set evaluation.
Download
pip install huggingface_hub
# Download everything
huggingface-cli download allenai/WildDet3D-Bench… See the full description on the dataset page: https://huggingface.co/datasets/allenai/WildDet3D-Bench.Hive-ALLA Semantically Consistent Dataset for Data-Efficient Query-Based Universal Sound Separation
Kai Li*, Jintao Cheng*, Chang Zeng, Zijun Yan, Helin Wang, Zixiong Su, Bo Zheng, Xiaolin Hu
Tsinghua University, Shanda AI, Johns Hopkins University
*Equal contribution
📜 Arxiv 2026 | 🎶 Demo | 🤗 Metadata | 🤗 Hive-ALL Audio | 🤗 Space
💥 News
[2026-05-21] Hive-ALL is now also available on ModelScope for users in China who prefer faster downloads via the… See the full description on the dataset page: https://huggingface.co/datasets/JusperLee/Hive-ALL.lrs3_all_video_wavall2024speedtest_1voxceleb2-40k-part1-preprocess-all-files-separatehey-ginolmoearth_projects_awfThis model is for fine-tuning OlmoEarth-v1-Base to map land use and land cover in southern Kenya.
The data is annotated by experts at the African Wildlife Foundation.
For more details, please see the documentation on GitHub at https://github.com/allenai/olmoearth_projects/blob/main/docs/awf.md
all2024speedtest_0all2024speedtest_12all2024speedtest_2all2024speedtest_7all2024speedtest_19all2024speedtest_13wds_vtab-clevr_count_all_test
CLEVR Count All Webdataset (Test set only)
Original paper: CLEVR: A Diagnostic Dataset for Compositional Language and Elementary Visual Reasoning
Homepage: https://cs.stanford.edu/people/jcjohns/clevr/
Bibtex:
@article{DBLP:journals/corr/JohnsonHMFZG16,
author = {Justin Johnson and
Bharath Hariharan and
Laurens van der Maaten and
Li Fei{-}Fei and
C. Lawrence Zitnick and
Ross B. Girshick},
title =… See the full description on the dataset page: https://huggingface.co/datasets/djghosh/wds_vtab-clevr_count_all_test.all2024speedtest_5all2024speedtest_14all2024speedtest_3all2024speedtest_16all2024speedtest_18all2024speedtest_6all2024speedtest_30all2024speedtest_25all2024speedtest_10all2024speedtest_17
