CoolFace
Datasetpublic

keypa/vision-adapter-images

Vision Adapter Image Corpus Processed image corpus used to train the Vision-Adapter: 79,659 agentic UI images ("screenshots" + "multistep" subsets derived from wave-ui-25k, ShowUI-desktop, and aguvis-stage2) plus full embeddable subsets of HuggingFaceM4/the_cauldron used for general/reasoning/conversational fine-tuning. Each image has been resized to ≤300k pixels and padded to 28-pixel multiples to match MoonViT-V2 preprocessing. Fields image: raw PNG bytes… See the full description on the dataset page: https://huggingface.co/datasets/keypa/vision-adapter-images.

sourceHugging Faceapache-2.0updated 1mo agoView on Hugging Face
0likes141downloads
settings

This repository belongs to keypa on Hugging Face.

CoolFace never edits a repository it does not host. Visibility, licence, collaborators and gating are all managed at the source.

namevision-adapter-images
visibilitypublic
licenceapache-2.0
gatedno
ownerkeypa
Account settings
keypa/vision-adapter-images · CoolFace