darknoon/simple-shapes-svg
The goal of this dataset is to measure and improve the ability of VLMs to see accurately in spatial dimensions. I've tried to ensure that all of the examples are not too hard have sufficient contrast between foreground and background shapes are not clipped or ambiguous solid background canvas is square 512x512 Initially, I've kept the "canvas" that they're working with 512x512 points, but you can learn more by experimenting with the dimensions as well.
156
Conversations for this repository live on Hugging Face.
CoolFace shows imported repositories read-only. Posting into someone else’s repository from here would need an authorised integration and the account holder’s consent, so the link goes to the source instead.
Open discussions on Hugging Face