datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
GlitchBench
GlitchBench
This repository contains the dataset for the paper GlitchBench: Can large multimodal models detect video game glitches?
by
Mohammad Reza Taesiri,
Tianjun Feng,
Anh Nguyen, and
Cor-Paul Bezemer
(CVPR 2024)
Abstract
Large multimodal models (LMMs) have evolved from large language models (LLMs) to integrate multiple input modalities, such as visual inputs. This integration augments the… See the full description on the dataset page: https://huggingface.co/datasets/glitchbench/GlitchBench.ROCOv2-radiology-minicityscapes-pseudo-labels
Cityscapes Unsupervised Panoptic Pseudo-Labels
Pseudo-labels for unsupervised panoptic segmentation on Cityscapes, generated using overclustered k-means semantics + depth-guided instance splitting.
Contents
Pseudo-Labels
Directory
Description
Files
Format
pseudo_semantic_raw_k80/
Overclustered k=80 semantic labels
~3.5K PNGs + centroids.npz
PNG (values 0-79), train/val split
cups_pseudo_labels_depthpro_tau020/
CUPS-format combined labels (DepthPro… See the full description on the dataset page: https://huggingface.co/datasets/qbit-glitch/cityscapes-pseudo-labels.sample_glitch_data2GlitchBenchv2-GlitchDetectioncityscapes_pseudo_labelsGlitchBenchPrivateGlitchBenchv2GlitchBenchv2-ParametricGlitchBenchv2-BugReportBugsBunny-GlitchNoGlitchBugsBunny-GlitchNoGlitch-Unbalancedsteamimages_glitchesGlitchBenchv2-WorkingSetGlitchBenchv2-VisualRegression
