datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
ids-project-artifactssr-artifact-prominence
SR Artifact Prominence
Annotated super-resolution artifact regions across four image subsets, with
crowdsourced per-region prominence scores, artifact type labels, and
natural-language descriptions.
Prominence is the fraction of valid crowd workers who answered that the
highlighted region contains a noticeable super-resolution artifact.
Subsets
Subset
Source dataset
Source images
Masks
Notes
open_images
Open Images
547
1,523
GT + LR-bicubic + multiple SR… See the full description on the dataset page: https://huggingface.co/datasets/imolodetskikh/sr-artifact-prominence.annotated-3DGS-artifacts
Puzzle Similarity
Project page | Paper | Code
This repository contains the dataset presented in the ICCV 2025 paper "Puzzle Similarity: A Perceptually-guided Cross-Reference Metric for Artifact Detection in 3D Scene Reconstructions"Authors: Nicolai Hermann, Jorge Condor, and Piotr Didyk
Dataset Description
The Dataset consists of 36 hand-selected 3D Gaussian Splatting renderings containing common reconstruction artefacts, (aligned) ground truths, human-annotated… See the full description on the dataset page: https://huggingface.co/datasets/nihermann/annotated-3DGS-artifacts.deepseek-ocr-artifacts-test-XXszl-artifacts
Part of the SZL Holdings governed estate — claims are designed to carry checkable receipts. Verification proves integrity & origin, never accuracy or performance.
SZL Artifacts — Build Artifact Registry
Artifact boundary - audited 2026-07-15: this Hugging Face dataset repository
is a mixed 106-file, 36,100,909-byte publication/build mirror at revision
91bcb443857f2884ef2bfabaaa6bfdc606c7134a. It is not a uniform table of
DSSE envelopes, and… See the full description on the dataset page: https://huggingface.co/datasets/SZLHOLDINGS/szl-artifacts.Snowball-67B-A2B-Mixed-RLVR-Experiment-Artifacts
Snowball 67B-A2B RL artifact release
2026 mixed-domain RLVR campaign
This release also contains the complete releasable record of the September 2026 Snowball mixed-domain RLVR campaign.
It covers the September 11 synchronous and bounded-staleness asynchronous RLVR1→RLVR2 lineages and the 5.7T
Agentic-start RLVR1 lineage. All training arms are terminal. The final campaign figure,
trace audit, canonical configs, timing reports, retained
traces, and operational… See the full description on the dataset page: https://huggingface.co/datasets/open-athena/Snowball-67B-A2B-Mixed-RLVR-Experiment-Artifacts.mod-arith-probe-artifactsgroundjudge-artifacts
GroundJudge: constructed artifacts and judge verdicts
Supporting data for "Do Vision-Language Judges Use the Image? A Causal
Audit of Visual Grounding in Multimodal Evaluation" (ICLR 2027
submission). Code: see the paper's GitHub repository.
This repo contains only this project's own constructed/generated
artifacts — every image variant, injected trace, and per-instance judge
verdict behind the paper's tables. It deliberately does not include:
Raw source datasets (GQA, ChartQA… See the full description on the dataset page: https://huggingface.co/datasets/rasulkhanbayov/groundjudge-artifacts.artifacts_evaluation
GRADE artifact evaluation (MobiCom 2026)
This anonymous artifact reproduces the quantitative evaluation of the GRADE paper. The release provides 19 model checkpoints plus one auxiliary TAESD weight file (20 .safetensors files in total), code for local inference and metric computation, and the saved camera-ready quantitative results.
Two evaluation paths are supported:
E1 — saved-result reproduction: regenerate the quantitative tables and figures from the supplied camera-ready… See the full description on the dataset page: https://huggingface.co/datasets/mypersonalsharingspot11/artifacts_evaluation.05-Artifactsdeepseek-ocr-artifacts-test-Zimproving-ca-lora-artifacts
Improving CA-LoRA — CA measurements and generated images
The two large evidence sets behind the reproduction-and-extension study of CA-LoRA
(Concept-Aware LoRA) on SDXL / Cityscapes:
measurements/ — the raw head-granularity concept-attribution tensors for a 13-timestep
sweep (backs Table 2 and Figure 1 of the report).
generated/ — every image that was scored for the main results (backs Table 3 of the
report).
The adapters that produced the images live in… See the full description on the dataset page: https://huggingface.co/datasets/chs35/improving-ca-lora-artifacts.deepseek-ocr-artifacts-test-01rescore-eval-artifacts
ReScore Eval Artifacts
Evaluation artifacts for ReScore OMR experiments.
Layout
benchmarks/teacher-forced/: teacher-forced benchmark reports.
benchmarks/free-generation/: free-generation predictions, renders, and summaries.
benchmarks/model-comparisons/: cross-model comparison reports.
FaceLinkGen-artifactslocus-bench
LOCUS-Bench (anonymized release for peer review)
A benchmark for embodied multi-robot task planning that grades two
difficulties separately. Axis S (state judgment): S0 no judgment; S1 whether
a single named target is already in place; S2 which of 2 to 3 conditional
candidates are absent; S3 which of 3 to 6 quantified instances are unsatisfied;
S4 whether invisible implies absent under occlusion, with single-frame fallback
planning. Axis M (mechanical structure): M0 none; M1 an… See the full description on the dataset page: https://huggingface.co/datasets/review-artifacts/locus-bench.sushi_atelier_artifacts
sushi_atelier_artifacts
Images, videos and other artefacts needed to run the website
sr-artifact-detection-demo-datasetdeepseek-ocr-artifactsdapi-artifact-imagesArtiFact
ArtiFact
ArtiFact is a large-scale multimodal benchmark of museum artwork records with aligned images and structured metadata. It is designed for evaluating metadata extraction, error detection, semantic querying, and multimodal reasoning over cultural-heritage collections.
The dataset combines records from the Rijksmuseum, the Metropolitan Museum of Art (Met), and the Art Institute of Chicago (AIC), with normalized fields for artists, dates, materials, techniques, dimensions… See the full description on the dataset page: https://huggingface.co/datasets/deem-data/ArtiFact.mva-repro-artifactsdeepseek-ocr-artifacts-test-XZtime-lapse-artifacts-derived
Time-Lapse Artifacts: Derived Access Layer
This repository contains reproducible access derivatives for
maxwellinked/time-lapse-artifacts.
The archival masters remain in the source dataset and remain authoritative.
Related interfaces
Public browser (pinned snapshot)
Browser source and revision history
The browser is a static presentation layer and may lag the current Hugging Face
dataset. Hugging Face remains authoritative for media, record identities, and… See the full description on the dataset page: https://huggingface.co/datasets/maxwellinked/time-lapse-artifacts-derived.AI_Image_Artifacts_Presentjupiter-tasktrove-dapo-artifacts
Jupiter TaskTrove DAPO — campaign artifacts
Standalone artifact archive for the jupiter-tasktrove-dapo campaign: a six-arm
objective ablation (GRPO control vs DAPO variants vs GSPO) training
Qwen/Qwen3-Coder-30B-A3B-Instruct with agentic RL (MarinSkyRL / SkyRL fully-async GRPO
trainers, Harbor + Daytona sandboxed terminus-2 rollouts) on competitive-programming
tasks, run on JSC Jupiter (GH200) 2026-08-20 → 2026-08-31.
The campaign closed inconclusive (platform degradation, a… See the full description on the dataset page: https://huggingface.co/datasets/penfever/jupiter-tasktrove-dapo-artifacts.review-dataset-index
MANUS-Bench — anonymous release bundle (WACV 2027 submission #1065)
MANUS-Bench (Multimodal Annotated Naturalistic Hand Understanding) is a
benchmark for geometry-conditioned hand-gesture synthesis in natural scenes:
aligned full-scene and localized hand conditions over an official
condition-complete protocol of 88,294 samples (64,931 train / 23,363 test)
drawn from an 88,312-record archive, with fixed in-domain, near-domain, and
hand-object OOD split roles.
This repository is… See the full description on the dataset page: https://huggingface.co/datasets/anon-review-artifact-7k3p9/review-dataset-index.deepseek-ocr-artifacts-testsplat-growth-artifactssphere-encoder-fid-artifacts
Sphere Encoder FID Evaluation Artifacts
This repository contains the evaluation artifacts for the paper Image Generation with a Sphere Encoder.
Project Page | GitHub Repository
These artifacts include data statistic files (fid_stats) and reference images (fid_refs) used to calculate Fréchet Inception Distance (FID) for generative models across several datasets, including CIFAR-10, ImageNet, Animal Faces, and Oxford Flowers.
Workspace Setup
Download the evaluation… See the full description on the dataset page: https://huggingface.co/datasets/kaiyuyue/sphere-encoder-fid-artifacts.
