datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
OLIVES_Dataset
OLIVES_Dataset
Abstract
Clinical diagnosis of the eye is performed over multifarious data modalities including scalar clinical labels, vectorized biomarkers, two-dimensional fundus images, and three-dimensional Optical Coherence Tomography (OCT) scans. While the clinical labels, fundus images and OCT scans are instrumental measurements, the vectorized biomarkers are interpreted attributes from the other measurements. Clinical practitioners use all these data modalities… See the full description on the dataset page: https://huggingface.co/datasets/gOLIVES/OLIVES_Dataset.OLIVES_Dataset
OLIVES_Dataset
Abstract
Clinical diagnosis of the eye is performed over multifarious data modalities including scalar clinical labels, vectorized biomarkers, two-dimensional fundus images, and three-dimensional Optical Coherence Tomography (OCT) scans. While the clinical labels, fundus images and OCT scans are instrumental measurements, the vectorized biomarkers are interpreted attributes from the other measurements. Clinical practitioners use all these data modalities… See the full description on the dataset page: https://huggingface.co/datasets/SOHAIBSUL123/OLIVES_Dataset.ProGAN-Eval
ProGAN eval set
This repository hosts the test set used in the FPBA on the LSUN Synthesis dataset, consisting of 1,000 test images from the LSUN dataset and 1,000 images synthesized by ProGAN.
shelf_jimUSE FIRST COMMIT FOR ORIGINAL DATASET. LATEST COMMIT IS MAHATHI-SHELF DATASET
mahathi_bottleOLiVES
OLiVES: An Outdoor Low-Light Video Benchmark for Enhancement and Segmentation
OLiVES is a new benchmark for low-light video enhancement and video object segmentation. It contains over 25,000 aligned normal/low-light frames and over 46,000 video object segmentation annotations.
LLVE
The video folder names and frame names for aligned frames are identical in input/ (low-light videos) and gt/ (normal-light videos)
VOS
Please run python VOS_dataset_mapper.py to… See the full description on the dataset page: https://huggingface.co/datasets/Anonymous-OLiVES/OLiVES.diffusion_shelf_deployspot-the-diffOriginal dataset: https://github.com/harsh19/spot-the-diff/
@inproceedings{jhamtani2018learning,
title={Learning to Describe Differences Between Pairs of Similar Images},
author={Jhamtani, Harsh and Berg-Kirkpatrick, Taylor},
booktitle={Proceedings of the 2018 Conference on Empirical Methods in Natural Language Processing (EMNLP)},
year={2018}
}
olive_tree_crown_detection
Olive Tree Crown Detection
A dataset for detection of crowns of olive trees. The dataset contains 5,842 images with 66,950 bounding box annotations. There are scales to the images taken.
This dataset is indexed on https://project-agml.github.io/ as part of the AgML python library.
Citation
@article{hnida2025olivetreecrownsdb,
title={OliveTreeCrownsDb: A high-resolution UAV dataset for detection and segmentation in agricultural computer vision},
author={Hnida… See the full description on the dataset page: https://huggingface.co/datasets/Project-AgML/olive_tree_crown_detection.siemens_difficult_task2sep20_smoketestclock-faces-v1-timesclock-faces-v1-hoursOLIVESintersection-count-reasoning-seedOLIVES_Dataset
OLIVES_Dataset
Abstract
Clinical diagnosis of the eye is performed over multifarious data modalities including scalar clinical labels, vectorized biomarkers, two-dimensional fundus images, and three-dimensional Optical Coherence Tomography (OCT) scans. While the clinical labels, fundus images and OCT scans are instrumental measurements, the vectorized biomarkers are interpreted attributes from the other measurements. Clinical practitioners use all these data… See the full description on the dataset page: https://huggingface.co/datasets/zycl2001/OLIVES_Dataset.stack_lego_simple_pi05_deploy_1openfarm-horse-grimace-region
OpenFARM Horse Grimace Region
Prepared horse facial-region pain-score data from the public Mendeley Data record 10.17632/t8rtzcgwxm.3.
Source
Source dataset: https://data.mendeley.com/datasets/t8rtzcgwxm
Source DOI: 10.17632/t8rtzcgwxm.3
Related paper DOI: 10.1371/journal.pone.0258672
License: CC BY 4.0
Splits
{
"train": 945,
"test": 279,
"train_raw": 3931,
"test_raw": 916
}
train and test are animal-heldout, region-score-balanced views.… See the full description on the dataset page: https://huggingface.co/datasets/oliveirabruno01/openfarm-horse-grimace-region.pi05_shelf_deploystack_lego_difficult_v1ditflow_shelf_deployOLIVES_Dataset
OLIVES_Dataset
Abstract
Clinical diagnosis of the eye is performed over multifarious data modalities including scalar clinical labels, vectorized biomarkers, two-dimensional fundus images, and three-dimensional Optical Coherence Tomography (OCT) scans. While the clinical labels, fundus images and OCT scans are instrumental measurements, the vectorized biomarkers are interpreted attributes from the other measurements. Clinical practitioners use all these data modalities… See the full description on the dataset page: https://huggingface.co/datasets/wangyiran2006/OLIVES_Dataset.litbench-mrl-eyesoundwel-pig-vocalizationssheep-facial-expression-benchmark
Sheep Facial Expression Benchmark
Prepared OpenFARM sheep facial-expression benchmark data from the public Mendeley Data record 10.17632/y5sm4smnfr.5.
Source
Source dataset: https://data.mendeley.com/datasets/y5sm4smnfr
Source DOI: 10.17632/y5sm4smnfr.5
Related paper DOI: 10.1016/j.compag.2020.105528
License: CC BY 4.0
Splits
{
"train": 172,
"test": 74,
"train_raw": 898,
"test_raw": 225
}
train and test are filtered/balanced views for benchmark and… See the full description on the dataset page: https://huggingface.co/datasets/oliveirabruno01/sheep-facial-expression-benchmark.diffusion_siemens_difficult_generalization_deployxchatbenchVanGogh_OliveTrees1889_MoMA_vs_TreeOil_ParallelMotionTorqueStudy
Van Gogh Olive Trees 1889 vs The Tree Oil Painting – Parallel Motion Torque Study
📁 Dataset: VanGogh_OliveTrees1889_MoMA_vs_TreeOil_ParallelMotionTorqueStudy🖼 Paintings:
The Olive Trees, 1889 — Vincent van Gogh (MoMA, New York)
The Tree Oil Painting (Private Collection)📊 Study Type: Physics-Informed Artistic Torque Identity Comparison📅 Last Updated: August 03, 2025
🧠 Dataset Summary
This dataset presents a comparative analysis of artistic brushstroke… See the full description on the dataset page: https://huggingface.co/datasets/HaruthaiAi/VanGogh_OliveTrees1889_MoMA_vs_TreeOil_ParallelMotionTorqueStudy.NP_MMPaper | Github Repository
🖊️ Citation
If you find this work helpful, please consider to star🌟 and cite this repo. Thanks for your support!
@misc{dmpo,
title={Beyond Mode Collapse: Distribution Matching for Diverse Reasoning},
author={Xiaozhe Li and Yang Li and Xinyu Fang and Shengyuan Ding and Peiji Li and Yongkang Chen and Yichuan Ma and Tianyi Lyu and Linyang Li and Dahua Lin and Qipeng Guo and Qingwen Liu and Kai Chen},
year={2026},
eprint={2605.19461}… See the full description on the dataset page: https://huggingface.co/datasets/OliverLee/NP_MM.stack_lego_simple_jim
