datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
jizzakh-regionToshkent_viloyatitashkentaidetector-data
Companion dataset for the aidetector
project (live demo: https://humanorai.online). Code, trained model, and the
reproducibility record live in that GitHub repo.
Dataset Card — AI Image Detector
This card documents the data the v2 models were trained and evaluated on. All
counts are taken from the reproducibility record
experiment_v1.json (dataset content hash
72b88efc0497...). The image files themselves are not redistributed in this
repository (see Access below).
The intent… See the full description on the dataset page: https://huggingface.co/datasets/aman213/aidetector-data.her-trees-simple-we
Her Trees Simple World-Engine Dataset
Synthetic 832x480, 81-frame solution videos for the Her Trees letter-assembly task.
Train: 3,000 cases; easy, medium, and hard have 1,000 each.
IID validation: validation/; 25 cases per difficulty and 75 total.
Same-letter fragments share one color and move independently.
See SPLITS.md and the split manifests for seeds and validation details.
MuSLR
🧩 MuSLR: Multimodal Symbolic Logical Reasoning Benchmark
Project page: "Multimodal Symbolic Logical Reasoning".
Paper Link: https://arxiv.org/abs/2509.25851
Multimodal symbolic logical reasoning, which aims to deduce new facts from multimodal input via formal logic, is critical in high-stakes applications such as autonomous driving and medical diagnosis, where rigorous, deterministic reasoning helps prevent serious consequences.
To evaluate such capabilities of current… See the full description on the dataset page: https://huggingface.co/datasets/Aiden0526/MuSLR.ai-det-test-human-refined-stat-no-labels-test
Auto-Generated FastDetector Dataset
Dataset: G-reen/ai-det-test-human-refined-stat-no-labels-test
Globals Config: config/globals_aidet_gemma_no_labels.toml
Analysis Config: config/analysis_nofilter.toml
Rows: 1,677
Evaluation Results
Prompt Subsets: 4 (direct_reference, indirect_reference, revise, rewrite)
Generator Configs: 1 (gemma-4-31B-it-AWQ-4bit (Temp: 1.0))
Classifiers: 13 (EditLens Roberta-Large Score, EditLens Roberta-Large Bucket, Perplexity… See the full description on the dataset page: https://huggingface.co/datasets/G-reen/ai-det-test-human-refined-stat-no-labels-test.ai-det-test-human-refined-stat-labels-test
Auto-Generated FastDetector Dataset
Dataset: G-reen/ai-det-test-human-refined-stat-labels-test
Globals Config: config/globals_aidet_gemma_labels.toml
Analysis Config: config/analysis_nofilter.toml
Rows: 1,679
Evaluation Results
Prompt Subsets: 4 (direct_reference, indirect_reference, revise, rewrite)
Generator Configs: 1 (gemma-4-31B-it-AWQ-4bit (Temp: 1.0))
Classifiers: 13 (EditLens Roberta-Large Score, EditLens Roberta-Large Bucket, Perplexity… See the full description on the dataset page: https://huggingface.co/datasets/G-reen/ai-det-test-human-refined-stat-labels-test.Karakalpakstancontinuous-rush-hour
Continuous Rush Hour
Deterministic continuous-space Rush Hour video-generation and reasoning data.
Train: 25,000 generated cases, 5,000 for each level 1--5.
Validation: 25 generated ID-held-out cases plus 4 true-example OOD cases.
Seed and environment provenance are recorded under manifests/.
Prompt templates are versioned under prompts/.
FerganaSirdaryoBukharasamarkand-regionaide_elderly
Aide Elderly Dataset
Description
Small dataset for experiments related to project Aide.
Authors
Fabien Allemand (fabien.allemand@cea.fr)
ai-detector-benchmark-test-data
🎯 AI Detector Benchmark Test Dataset
A comprehensive benchmark dataset for testing AI image detection models.
📊 Dataset Summary
Total Images: 700
AI-Generated: 250 images (from 5 different generators)
Real Images: 450 images (from 9 diverse datasets)
Perfect for:
✅ Testing AI detection models
✅ Creating leaderboards
✅ Comparing model performance
✅ Benchmarking new approaches
🤖 AI Generators Included
Generator
Images
Accuracy Baseline
FLUX… See the full description on the dataset page: https://huggingface.co/datasets/Robo531/ai-detector-benchmark-test-data.ai-detector-benchmark-test-data
🎯 AI Detector Benchmark Test Dataset
A comprehensive benchmark dataset for testing AI image detection models.
📊 Dataset Summary
Total Images: 700
AI-Generated: 250 images (from 5 different generators)
Real Images: 450 images (from 9 diverse datasets)
Perfect for:
✅ Testing AI detection models
✅ Creating leaderboards
✅ Comparing model performance
✅ Benchmarking new approaches
🤖 AI Generators Included
Generator
Images
Accuracy Baseline
FLUX… See the full description on the dataset page: https://huggingface.co/datasets/ash12321/ai-detector-benchmark-test-data.Historical_hero_AlpomishHistorical_hero_Shiroqai-design-benchmark
AI Design Benchmark
English | 中文
English
A reproducible benchmark framework for evaluating AI design tools across 7 design scenarios
News
2026-04: v1.1 released — added reference images (53 files), bilingual tasks, CI pipeline, and 9 scoring tests; paper PDF published as Release asset
2026-04: Round 1 results published — Lovart 56% (95% CI: 47%–64%), Jimeng 27%, Roboneo 19% across 7 scenes
Overview
This project provides a reproducible… See the full description on the dataset page: https://huggingface.co/datasets/Sanjue-Logic/ai-design-benchmark.Historical_hero_AIdetection-dataHistorical_hero_TomarisHistorical_hero_ManguberdiAiderAIDE_image_detector_codeISIC_BCC_1CFD
Dataset Card for Dataset Name
This dataset card aims to be a base template for new datasets. It has been generated using this raw template.
Dataset Details
Dataset Description
Curated by: [More Information Needed]
Funded by [optional]: [More Information Needed]
Shared by [optional]: [More Information Needed]
Language(s) (NLP): [More Information Needed]
License: [More Information Needed]
Dataset Sources [optional]
Repository: [More… See the full description on the dataset page: https://huggingface.co/datasets/AIdeveric/CFD.imagesISIC_MEL_1
