datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
visa-anomaly-detection
VisA — Visual Anomaly Dataset
Mirror of the VisA (Visual Anomaly) dataset for research use. Staged as a proxy/pretraining
dataset for CoRe's Situational Control paint-inspection work (core-lab/situational-control).
Source
Original repo: https://github.com/amazon-science/spot-diff
Original download: https://amazon-visual-anomaly.s3.us-west-2.amazonaws.com/VisA_20220922.tar
License: CC BY 4.0 (confirmed in the source repo README)
Contents
10,821… See the full description on the dataset page: https://huggingface.co/datasets/imaadd05/visa-anomaly-detection.vehicle-mixed-traffic-detection
Visaitech Mixed-Traffic Vehicle Detection Dataset (v0.1)
Dashcam frames annotated for pedestrian / 2-wheeler / 3-wheeler / 4-wheeler
detection in South Asian mixed traffic, a class taxonomy general-purpose
COCO-trained detectors don't cover (COCO has no concept of an auto-rickshaw
or motorcycle-vs-bicycle-as-one-class "2-wheeler" grouping tuned for how
this traffic actually mixes on the road).
This is an early v0.1 release: 293 annotated frames from 6 source videos,
published… See the full description on the dataset page: https://huggingface.co/datasets/visaitech/vehicle-mixed-traffic-detection.mirage_mvtec_visavisaOriginal dataset:
@inproceedings{zou2022spot,
title={Spot-the-difference self-supervised pre-training for anomaly detection and segmentation},
author={Zou, Yang and Jeong, Jongheon and Pemula, Latha and Zhang, Dongqing and Dabeer, Onkar},
booktitle={European Conference on Computer Vision},
pages={392--408},
year={2022},
organization={Springer}
}
VisA-2KHigh-resolution industrial image anomaly detection dataset VisA-2K.For more information, see HiAD.
Download
huggingface-cli download --repo-type dataset XimiaoZhang/VisA-2K --local-dir VisA-2K --resume-download
paper-visaCapacitacao_Visao_Computacional
Visão Geral
Este repositório contém as atividades práticas e teóricas do curso de Capacitação em Visão Computacional. O curso aborda fundamentos de processamento digital de imagens, técnicas de filtragem, segmentação, extração de características e aplicações em aprendizado de máquina.
Estrutura do Repositório
O repositório está organizado em pastas por atividade, cada uma contendo:
Enunciado da atividade em PDF
Notebook Jupyter (quando aplicável)
README com… See the full description on the dataset page: https://huggingface.co/datasets/arvoredossaberes/Capacitacao_Visao_Computacional.visa-datasetwiki-visafineweb-visaVisAlign
VisAlign: Dataset for Measuring the Alignment between AI and Humans in Visual Perception
This is the test set of VisAlign (NeurIPS 2023 Datasets and Benchmarks Track), a dataset for measuring the degree of alignment between AI models and humans in visual perception. It contains 900 images across 8 categories.
Ground-truth labels and per-image categories are withheld, and filenames are anonymized IDs — to evaluate your model, submit your predictions to the VisAlign Leaderboard.… See the full description on the dataset page: https://huggingface.co/datasets/jiyounglee0523/VisAlign.defectforge-visa-synthetic
DefectForge VisA Synthetic Defects
Synthetic defect images and generation-time masks for the VisA pcb1 and capsules
objects. Each object is generated from only 10 real anomalous training images, while the
frozen high-shot test partition is never visible to generation, filtering, or quality
reference sets.
繁中摘要:這是 VisA pcb1/capsules 的少樣本工業瑕疵合成資料。每個物件只用
10 張真實瑕疵訓練圖;mask 是生成時使用的標註,不是模型事後預測。資料同時提供
filtered 與 unfiltered 版本,並公開 provenance 與 test SHA-256 blocklist。
What is… See the full description on the dataset page: https://huggingface.co/datasets/steven0226/defectforge-visa-synthetic.VisAnalog
VisAnalog: A Diagnostic Suite for Visual Concept Transfer on Natural Images
VisAnalog is a diagnostic benchmark for visual concept transfer on natural images.
Each example follows an analogy pattern: infer the transformation from pair1_source
to pair1_target, transfer that concept to pair2_source, and answer a multiple-choice
question about the expected pair2_target.
Dataset Structure
The uploaded split is test with 617 examples.
Main columns:
pair1_source: first source… See the full description on the dataset page: https://huggingface.co/datasets/zli99/VisAnalog.VIS-APP-Bench
anchor_tasks_web — Dataset
This dataset is part of the benchmark presented in the paper VISTA: An End-to-End Benchmark for Visual Spec-to-Web-App Coding Agents.
Project Page | GitHub Repository | Paper
A web-app generation benchmark. Each task is a multi-page UI taken from a public Figma community file. For every task we ship the textual page descriptions, the rendered mockup PNGs, the Figma node structure, the per-page click-annotations, and the distilled data-testid "anchors"… See the full description on the dataset page: https://huggingface.co/datasets/JunJiaGuo/VIS-APP-Bench.visargs
Dataset Card for VisArgs Benchmark
Dataset Summary
Data from: Selective Vision is the Challenge for Visual Reasoning: A Benchmark for Visual Argument Understanding
@article{chung2024selective,
title={Selective Vision is the Challenge for Visual Reasoning: A Benchmark for Visual Argument Understanding},
author={Chung, Jiwan and Lee, Sungjae and Kim, Minseo and Han, Seungju and Yousefpour, Ashkan and Hessel, Jack and Yu, Youngjae},
journal={arXiv preprint… See the full description on the dataset page: https://huggingface.co/datasets/jiwan-chung/visargs.VisA_Extended
Introduction
This is the extended VisA dataset with the ground truth segmentations for each defect types for each image. It represents the dataset along with the paper "MultiADS: Defect-aware Supervision for Multi-type Anomaly Detection and Segmentation in Zero-Shot Learning" which is published on ICCV 2025 conference.
Dataset Structure
VisA_Extended/
|-- candle/
|-----|--- Data/
|-----|-----|---- Anomaly_grouped/
|-----|-----|------------|----- 000/… See the full description on the dataset page: https://huggingface.co/datasets/zhKingg/VisA_Extended.ViSAEvis_ann_rep_benchmark_v3mirage_mvtec_visavisa-splitVisA-DISthreshvisa_split_oneshotVisAtomVisA
