datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
FiVE-Fine-Grained-Video-Editing-Benchmark
FiVE-Bench
FiVE-Bench: A Fine-Grained Video Editing Benchmark for Evaluating Diffusion and Rectified Flow Models
Minghan Li1*, Chenxi Xie2*, Yichen Wu13, Lei Zhang2, Mengyu Wang1†
1Harvard University 2The Hong Kong Polytechnic University 3City University of Hong Kong
*Equal contribution †Corresponding Author
💜 Leaderboard (coming soon) |
💻 GitHub |
🤗 Hugging Face
📝 Project Page |
📰 Paper |
🎥 Video Demo
FiVE is a benchmark comprising 100 videos for… See the full description on the dataset page: https://huggingface.co/datasets/LIMinghan/FiVE-Fine-Grained-Video-Editing-Benchmark.fine-grained-medical-reasoning
Dataset Card for Fine-Grained Medical Reasoning
Fine-grained medical reasoning QA dataset introduced in "Can LLMs Reason Like Doctors? Exploring the Limits of Large Language Models in Complex Medical Reasoning"
(Findings of EACL 2026). Manually annotated from the MedAgentsBench test_hard set,
it evaluates LLMs’ abduction, deduction, and induction capabilities, offering detailed insights into physician-like reasoning.
Dataset Details
Dataset… See the full description on the dataset page: https://huggingface.co/datasets/expertailab/fine-grained-medical-reasoning.Amazon_Reviews_for_Sentiment_Analysis_fine_grained_5_classes
Dataset Card for Dataset Name
The Amazon reviews full score dataset is constructed by randomly taking 600,000 training samples and 130,000 testing samples for each review score from 1 to 5. In total there are 3,000,000 trainig samples and 650,000 testing samples.
Dataset Details
Dataset Description
The files train.csv and test.csv contain all the training samples as comma-sparated values. There are 3 columns in them, corresponding to class index (1 to 5)… See the full description on the dataset page: https://huggingface.co/datasets/yassiracharki/Amazon_Reviews_for_Sentiment_Analysis_fine_grained_5_classes.test-fine-grained-challenges
Fine-Grained Challenges
Groups of morphologically similar species for evaluating and tuning fine-grained classifiers on top of BioCLIP ecosystem model embeddings. Each group gathers species that are easily confused with one another, and the groups span several clades so a classifier can be probed on the distinctions that actually matter rather than on coarse taxonomy.
The corpus lives in a Lance dataset and can be acted on as a whole, on any single group independently, or on any… See the full description on the dataset page: https://huggingface.co/datasets/thompsonmj/test-fine-grained-challenges.fine-grained-challenges
Dataset Card for Fine-Grained Challenges
Fine-Grained Challenges collects focused groups of visually similar animals for testing biological image classifiers. It combines images, taxonomy, provenance, and frozen embeddings from three BioCLIP-family models in one Lance dataset.
Dataset Details
Release v0.1.0 contains three challenge groups:
Challenge group
Focus
Rows
Species-labeled rows
Genus or higher rows
Labeled species
Peromyscus
Deermice and close… See the full description on the dataset page: https://huggingface.co/datasets/imageomics/fine-grained-challenges.Fine_Grained_Fandom_Benchmark_Action_Sequences
Codified Decision Tree (CDT) Action Sequences
This dataset contains scene-action pairs derived from storylines, used to train and evaluate role-playing (RP) agents using the Codified Decision Trees (CDT) framework.
Paper: Deriving Character Logic from Storyline as Codified Decision Trees
Repository: https://github.com/KomeijiForce/Codified_Decision_Tree
Introduction
Role-playing (RP) agents rely on behavioral profiles to act consistently across diverse narrative… See the full description on the dataset page: https://huggingface.co/datasets/KomeijiForce/Fine_Grained_Fandom_Benchmark_Action_Sequences.Yelp_Reviews_for_Sentiment_Analysis_fine_grained_5_classes
Dataset Card for Dataset Name
The Yelp reviews full star dataset is constructed by randomly taking 130,000 training samples and 10,000 testing samples for each review star from 1 to 5. In total there are 650,000 trainig samples and 50,000 testing samples.
Dataset Description
The files train.csv and test.csv contain all the training samples as comma-sparated values. There are 2 columns in them, corresponding to class index (1 to 5) and review text. The review texts are… See the full description on the dataset page: https://huggingface.co/datasets/yassiracharki/Yelp_Reviews_for_Sentiment_Analysis_fine_grained_5_classes.OpenCaption-FineGrained
OpenCaption-FineGrained
OpenCaption-FineGrained is a high-quality dense image captioning dataset containing fine-grained, long-form image descriptions synthesized using the Qwen3.6 Multimodal model. Rather than generating a single caption per image, the dataset follows a multi-sample caption selection pipeline, where multiple candidate captions are generated for every image and the highest-quality caption is selected using an automated quality filtering strategy.
The resulting… See the full description on the dataset page: https://huggingface.co/datasets/prithivMLmods/OpenCaption-FineGrained.formosa-vision-finegrained
Formosa Vision Fine-grained (Expanded)
Dataset Summary
此資料集以台灣在地文化與地景為核心,提供具細節的中文描述,並保留原始圖像。
擴充版本針對每張圖像生成更長、更密集的語義描述,以強化模型在細節理解上的表現。
Motivation
『資料合成』FLAIR 的核心在於訓練模型「聽得懂細節」。這意味著「長文本」越具體、包含越多方位詞 (左上角、紅色物體旁...),模型學到的局部特徵就越好。因為在此階段會透過大型多模態模型生成豐富且長的中文描述夠「碎唸」(包含大量方位、顏色、材質等細節)。相較於網路爬蟲數據,此資料庫具備高品質的本土文化實體 (Entity) 標註,是訓練台灣在地化 AI 的最佳基石。
Source Data
原始資料集:twinkle-ai/Formosa-Vision(Hugging Face Datasets)
擴充流程:以本地 VLM 產生更細緻的中文長描述
Dataset Structure… See the full description on the dataset page: https://huggingface.co/datasets/renhehuang/formosa-vision-finegrained.fine_grained_unlearning
Fine-Grained Knowledge Unlearning — Namesake Benchmark
A benchmark for fine-grained knowledge unlearning: can a method remove a fact
about entity X without damaging the same fact on entity Y, when X and Y
have (near-)identical names and share exactly that one attribute?
Each sample is a pair of real people who
share an identical or near-identical name,
share one career element (e.g. both are basketball players) — the fact to
unlearn on X and retain on Y,
differ on everything… See the full description on the dataset page: https://huggingface.co/datasets/ernlavr/fine_grained_unlearning.FiVE-Fine-Grained-Video-Editing-Benchmark
FiVE-Bench
FiVE-Bench: A Fine-Grained Video Editing Benchmark for Evaluating Diffusion and Rectified Flow Models
Minghan Li1*, Chenxi Xie2*, Yichen Wu13, Lei Zhang2, Mengyu Wang1†
1Harvard University 2The Hong Kong Polytechnic University 3City University of Hong Kong
*Equal contribution †Corresponding Author
💜 Leaderboard (coming soon) |
💻 GitHub |
🤗 Hugging Face
📝 Project Page |
📰 Paper |
🎥 Video Demo
FiVE is a benchmark comprising 100 videos for… See the full description on the dataset page: https://huggingface.co/datasets/CiaranCw/FiVE-Fine-Grained-Video-Editing-Benchmark.finegrained_modality_conflicts_trainreplanner_fine_grained_full_3_all_with_episode_rewardshateful_memes_fine_grained
Hateful Memes Fine-Grained Dataset
This dataset is a fine-grained extension of the widely used Hateful Memes dataset, designed to enable more nuanced analysis of harmful multimodal content. While the original dataset focuses on binary hatefulness classification, this extension introduces additional annotation dimensions capturing incivility and intolerance at a more granular level.
The dataset consists of a subset of 2,030 memes, each annotated independently by three annotators.… See the full description on the dataset page: https://huggingface.co/datasets/nils-herrmann/hateful_memes_fine_grained.gretel-pii-fine-grainedreplanner_fine_grainedreplanner_fine_grained2replanner_fine_grained_full_3_allAmazon_Reviews_for_Sentiment_Analysis_fine_grained_5_classes
Dataset Card for Dataset Name
The Amazon reviews full score dataset is constructed by randomly taking 600,000 training samples and 130,000 testing samples for each review score from 1 to 5. In total there are 3,000,000 trainig samples and 650,000 testing samples.
Dataset Details
Dataset Description
The files train.csv and test.csv contain all the training samples as comma-sparated values. There are 3 columns in them, corresponding to class index (1 to 5)… See the full description on the dataset page: https://huggingface.co/datasets/ydjgfdj/Amazon_Reviews_for_Sentiment_Analysis_fine_grained_5_classes.finegrained_vehicle_labelskodcode-complete_1000_qwen7b_sol_iter0_att10_sol5_lr5e5_3ep_finegrained_dpo_10000kodcode-complete_1000_qwen7b_att_iter1_att20_sol5_finegrained_at1.2_st0.7car-bdd-fine-grained
Fine-Grained Vehicle Detection Dataset (Corolla × BMW 3-Series, BDD100K-derived)
Real-world dashcam frames with fine-grained make annotations on top of
BDD100K's existing car bounding boxes. Identifies which BDD-labeled cars
are specifically Toyota Corolla sedans or BMW 3-Series sedans.
Images: 1410 unique frames · Annotations: 1524
(558 Corolla + 966 BMW 3-Series)
Source imagery: BDD100K dashcam corpus (Berkeley DeepDrive)
Why this dataset
BDD100K labels every car as… See the full description on the dataset page: https://huggingface.co/datasets/arrmlet/car-bdd-fine-grained.fine_grained_media_sentiments_annotationsTask: Aspect based sentiment recognition. Domain: Political news coverage. Named entities have been masked with [NEG], [NEU], or [POS] tokens. Dataset size: 1400 clippits.
An ABSA-BERT model successfully leveraged this dataset, with semi-supervised learning, to get very good results on fine grained sentiment recognition.
Creation: Five of my friends voluntarily used my browser plugin annotation tool I made to send me marked sentences of the news they were reading. How well they understood the… See the full description on the dataset page: https://huggingface.co/datasets/bitsinthesky/fine_grained_media_sentiments_annotations.kodcode-complete_1000_qwen7b_sol_best_of_25_finegrainedkodcode-complete_1000_qwen7b_att_iter0_att20_sol5_finegrained_at1.2_st0.7kodcode-complete_1000_qwen7b_att_iter1_att20_sol5_finegrained_at1.2_st0.7_zeroed_shapedkodcode-complete_1000_qwen7b_att_iter0_att20_sol5_finegrained_at1.2_st0.7_filtered_shapingkodcode-complete_1000_qwen7b_att_iter0_att20_sol5_finegrained_at1.2_st0.7_filtered_shapedkodcode-complete_1000_qwen7b_att_iter0_att20_sol5_finegrained_at1.2_st0.7_shaped
