datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
coffee-bean-grading-dataset
An Image Dataset of Pre-Roast Robusta Coffee Beans with Polygon Annotations for Automated Grading
Abstract
Automated quality assessment of raw agricultural products is critical for ensuring fair trade and supply chain efficiency. This dataset presents 3,877 high-resolution images of pre-roast Arabica coffee beans, collected from farms in **Coorg, Karnataka (India)**—a major coffee-producing region. Each bean is categorized into one of four quality grades:
Grade… See the full description on the dataset page: https://huggingface.co/datasets/SamruddhK/coffee-bean-grading-dataset.idrid-disease-grading
Indian Diabetic Retinopathy Image Dataset (IDRiD)
This dataset is the disease grading portion of the IDRiD.
The original source of the dataset is here: https://ieee-dataport.org/open-access/indian-diabetic-retinopathy-image-dataset-idrid
paper-gradingdiabetic-retinopathy-grading-africa
DR-Grading — Fundus Diabetic-Retinopathy Grading with Africa-Grounded Synthetic Clinical Context
A dataset for cross-sectional 5-class diabetic-retinopathy grading, pairing
resized colour fundus photographs with minimal, epidemiology-grounded synthetic
point-of-care context (age, sex, diabetes duration).
Version 1.0.0 · core dr_synth 1.0.0 · part of the DR-Africa dataset family
(see also dr-progression and
dr-africa-benchmark).
Abstract
Diabetic retinopathy (DR)… See the full description on the dataset page: https://huggingface.co/datasets/macular/diabetic-retinopathy-grading-africa.coffee-bean-grading-dataset
An Image Dataset of Pre-Roast Robusta Coffee Beans with Polygon Annotations for Automated Grading
Abstract
Automated quality assessment of raw agricultural products is critical for ensuring fair trade and supply chain efficiency. This dataset presents 3,877 high-resolution images of pre-roast Arabica coffee beans, collected from farms in **Coorg, Karnataka (India)**—a major coffee-producing region. Each bean is categorized into one of four quality grades:… See the full description on the dataset page: https://huggingface.co/datasets/kushi-ai-2027/coffee-bean-grading-dataset.grading_group1lithium-ion-cell-capacity-grading-process-curves
Lithium-ion Cell Capacity-Grading (FR) Process Curves
Channel-level process curves from the capacity-grading station of a cylindrical
lithium-ion cell line. Every tester channel is sampled natively every 30 seconds for
the whole process, giving the full voltage / current / capacity trajectory of each cell
from the moment it is clamped.
This is the capacity-grading (FR) dataset. Pre-charge is a completely different
process and is published separately; the two are deliberately… See the full description on the dataset page: https://huggingface.co/datasets/michealsmitch/lithium-ion-cell-capacity-grading-process-curves.DR_Grading
Dataset Card for "DR_Grading"
More Information needed
imo-gradingbench
IMO-GradingBench
Dataset Description
IMO-GradingBench is a benchmark dataset for evaluating the automatic grading capabilities of large language models. It consists of 1,000 human gradings of model-generated solutions to mathematical problems.
This dataset is part of the IMO-Bench suite, released by Google DeepMind in conjunction with their 2025 IMO gold medal achievement.
Supported Tasks and Leaderboards
The primary task for this dataset is automatic grading… See the full description on the dataset page: https://huggingface.co/datasets/Hwilner/imo-gradingbench.id_short_answer_gradingIndonesian short answers for Biology and Geography subjects from 534 respondents where the answer grading was done by 7 experts.\ap-rowa-thesis-grading
AP World History Row A Thesis-Grading Dataset (synthetic)
Supervised fine-tuning + DPO data for a single-criterion grader: the AP World History
LEQ/DBQ Row A (thesis/claim) point. Each example is a chat-format row whose assistant
turn is a JSON object {"point": 0|1, "reason": "..."} produced under a compact Row A
grader prompt.
This dataset trains a small open model (Qwen3-0.6B + LoRA) to make the Row A earn/deny
decision without importing higher-row criteria ("evaluate the… See the full description on the dataset page: https://huggingface.co/datasets/anshulmago1/ap-rowa-thesis-grading.grading-question-triage-datasetdr-grading-pipelineAI-Grading-SystemDisease_Grading_for_DR_and_Mucula
Dataset Card for "Disease_Grading_for_DR_and_Mucula"
More Information needed
english-gradinghttps://www.kaggle.com/competitions/feedback-prize-english-language-learning
grading-templatesafrican_plum_grading_classification
African Plum Grading Classification
A dataset for grade classification of plums. The dataset contains 4,507 images across 6 classes: bruised, cracked, rotten, spotted, unaffected, unripe.
Images per class:
bruised: 319
cracked: 162
rotten: 720
spotted: 759
unaffected: 1,721
unripe: 826
This dataset is indexed on https://project-agml.github.io/ as part of the AgML python library.
Citation
@article{fadja2025dataset,
title={A dataset of annotated African plum… See the full description on the dataset page: https://huggingface.co/datasets/Project-AgML/african_plum_grading_classification.Melodic_pattern_reproduction_performances_gradingEssay-quetions-auto-grading-arabicDataset Overview
The Open Orca Enhanced Dataset is meticulously designed to improve the performance of automated essay grading models using deep learning techniques. This dataset integrates robust data instances from the FLAN collection, augmented with responses generated by GPT-3.5 or GPT-4, creating a diverse and context-rich resource for training models.
Dataset Structure
The dataset is structured in a tabular format, with the following key fields:
id: A unique identifier for each data… See the full description on the dataset page: https://huggingface.co/datasets/mohamedemam/Essay-quetions-auto-grading-arabic.Essay-quetions-auto-gradingDataset Overview
The Open Orca Enhanced Dataset is meticulously designed to improve the performance of automated essay grading models using deep learning techniques. This dataset integrates robust data instances from the FLAN collection, augmented with responses generated by GPT-3.5 or GPT-4, creating a diverse and context-rich resource for training models.
Dataset Structure
The dataset is structured in a tabular format, with the following key fields:
id: A unique identifier for each data… See the full description on the dataset page: https://huggingface.co/datasets/mohamedemam/Essay-quetions-auto-grading.DR_Grading_413_103
Dataset Card for "DR_Grading_413_103"
More Information needed
gradingGradingBench
task_categories:
- image-to-text
- visual-question-answering
language:
- zh
- en
tags:
- exam-grading
- ocr
- vision-language
- education
pretty_name: GradingBench
size_categories:
- 1K<n<10K
GradingBench
Comprehensive complex-instruction benchmark for exam paper grading (L1/L2/L3).
Images and annotations are separated:
data/
├── images/
│ ├── L1/
│ │ ├── Mathematics/ *.jpg
│ │ ├── Chinese/
│ │ ├── English/
│ │ ├── Science/
│ │ └──… See the full description on the dataset page: https://huggingface.co/datasets/ERRORSEMI/GradingBench.matharena-gradingbenchPumpkin-Maturity-Grading-Dataset
Pumpkin Maturity Grading Dataset
The current agricultural industry faces challenges in quality control of crops, especially in judging the maturity of pumpkins. Traditional methods often rely on manual identification, which is inefficient and prone to errors. Although existing image recognition technology has made some progress, there is a lack of high-quality datasets specifically targeted at pumpkin maturity. This dataset aims to improve the accuracy of maturity assessment for… See the full description on the dataset page: https://huggingface.co/datasets/Mobiusi/Pumpkin-Maturity-Grading-Dataset.LLM-GradingEval-1000essay-grading-criteriaessay_grading_for_instruction_tuninggarment-grading-specs
