rasulkhanbayov/groundjudge-artifacts
GroundJudge: constructed artifacts and judge verdicts Supporting data for "Do Vision-Language Judges Use the Image? A Causal Audit of Visual Grounding in Multimodal Evaluation" (ICLR 2027 submission). Code: see the paper's GitHub repository. This repo contains only this project's own constructed/generated artifacts — every image variant, injected trace, and per-instance judge verdict behind the paper's tables. It deliberately does not include: Raw source datasets (GQA, ChartQA… See the full description on the dataset page: https://huggingface.co/datasets/rasulkhanbayov/groundjudge-artifacts.
Conversations for this repository live on Hugging Face.
CoolFace shows imported repositories read-only. Posting into someone else’s repository from here would need an authorised integration and the account holder’s consent, so the link goes to the source instead.
Open discussions on Hugging Face