datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
drawvla-prompt-validation-clean
DrawVLA — Sketch-Prompt Validation
Circle (which) + arrow (where) + caption (what) visual instructions overlaid on
LIBERO observations, each labelled with a binary
verdict for training a prompt validator or a self-checking VLA:
right — every channel is correct and exactly one reading survives; execute.
wrong — a channel is incorrect or the deictic prompt remains under-determined;
reject. Formerly ambiguous prompts are retained in this class.
All captions are name-free L2/L3… See the full description on the dataset page: https://huggingface.co/datasets/shibuina/drawvla-prompt-validation-clean.drawvla-prompt-validation
DrawVLA — Sketch-Prompt Validation
Circle (which) + arrow (where) + caption (what) visual instructions overlaid on
LIBERO observations, each labelled with a binary
verdict for training a prompt validator or a self-checking VLA:
right — every channel is correct and exactly one reading survives; execute.
wrong — a channel is incorrect or the deictic prompt remains under-determined;
reject. Formerly ambiguous prompts are retained in this class.
All captions are name-free L2/L3… See the full description on the dataset page: https://huggingface.co/datasets/shibuina/drawvla-prompt-validation.drawvla-prompt-validation-v3
DrawVLA — Sketch-Prompt Validation
Circle (which) + arrow (where) + caption (what) visual instructions overlaid on
LIBERO observations, each labelled with a binary
verdict for training a prompt validator or a self-checking VLA:
right — every channel is correct and exactly one reading survives; execute.
wrong — a channel is incorrect or the deictic prompt remains under-determined;
reject. Formerly ambiguous prompts are retained in this class.
All captions are name-free L2/L3… See the full description on the dataset page: https://huggingface.co/datasets/shibuina/drawvla-prompt-validation-v3.image-prompt-injection
Image-Based Prompt Injection Dataset
Synthetic dataset for prompt injection attacks on Large Vision-Language Models (LVLMs).
Team
Name
GitHub
Ritik Sinha
@Ritik1207-ind
Siddhant Kumar
@siddhantkumar101
Udit Dadhich
@UditDadhich
GitHub Repository: prompt-injection-attacks-on-LVLMS
Dataset Details
4,859 labeled samples
4 attack types: typographic, structural, adversarial, metadata
4 injection goals: jailbreak, exfiltration… See the full description on the dataset page: https://huggingface.co/datasets/Reet1207/image-prompt-injection.PromptedArtistIdentificationDataset-ViewSamples
Sample Dataset Viewer for Prompted Artist Identification Dataset
Website | Paper | GitHub
Identifying Prompted Artist Names from Generated Images
Grace Su, Sheng-Yu Wang, Aaron Hertzmann, Eli Shechtman, Jun-Yan Zhu, Richard Zhang
arXiv, 2025
Description
This page serves as a viewer for sample images from the Prompted Artist Identification Dataset.
Please visit the main dataset page for a description of the full dataset.
The entire benchmark dataset consists of 1.95… See the full description on the dataset page: https://huggingface.co/datasets/cmu-gil/PromptedArtistIdentificationDataset-ViewSamples.
