datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
cardio-mark
CardioMark Review Subset
This repository contains an anonymized review subset of the CardioMark benchmark introduced for automated vertebral heart score (VHS) estimation in canine thoracic radiographs.
The subset is provided to support reproducibility and data-quality inspection during peer review.
Dataset Overview
CardioMark is a large-scale benchmark for evaluating the complete VHS measurement pipeline, including:
cardiac landmark localization
geometric VHS estimation… See the full description on the dataset page: https://huggingface.co/datasets/gen-ai-researcher/cardio-mark.GenAI-RealEstate-TestSet
🏙️ GenAI Real Estate Test Set (Track B)
Dataset for the MenaML Winter School 2026 Challenge.
📊 Dataset Structure
This dataset contains 1,000 images split evenly between:
Authentic: Real estate photography from the Places365 dataset.
Manipulated: Synthetically generated deepfake artifacts (Inpainting, Diffusion Noise, GAN Grids).
🕵️ How to Use
This dataset is designed for testing forensic detection models.
genai-manipulation-detection-interior
GenAI Manipulation Detection Dataset - Interior Design Images
📋 Dataset Description
This dataset contains 1000 paired images (real + manipulated) for training and evaluating GenAI manipulation detection models. Created for the MenaML Winter School 2026 Hackathon.
Dataset Summary
Total Images: 1000 pairs (2000 total images)
Image Size: 512x512
Format: JPEG
Source: Pinterest Interior Design Images (Kaggle)
License: MIT
🎯 Challenge Context
This… See the full description on the dataset page: https://huggingface.co/datasets/FatimahEmadEldin/genai-manipulation-detection-interior.
