CoolFace
Datasetpublic

monica-sekoyan/TickTockVQA-segmented

TickTockVQA Segmented (SAM3 crops) Real world analog clock images from jaeha-choi/TickTockVQA, each cropped to the clock face by a SAM3 segmentation pass, packaged for grounded visual reasoning experiments with vision language models. Every image is paired with a single fixed instruction and a ground truth time label. All answers are in H:MM format: this corpus contains no second hand, so it exercises real world hour and minute reading only and carries no H:MM:SS signal.… See the full description on the dataset page: https://huggingface.co/datasets/monica-sekoyan/TickTockVQA-segmented.

sourceHugging Faceotherupdated 2mo agoView on Hugging Face
0likes197downloads
7 commits on main
17a05512mo ago

Fix image decoding for image and image_full

monica-sekoyan
ce733542mo ago

Add full images and pixel-identity columns

monica-sekoyan
f1219db3mo ago

Update README.md

monica-sekoyan
1a868ed3mo ago

Upload resolution_distribution.png

monica-sekoyan
2ea971c3mo ago

Add dataset card

monica-sekoyan
f755cac3mo ago

Upload dataset

monica-sekoyan
56135803mo ago

initial commit

monica-sekoyan