datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
image-pointing-1M-sft-swiftimage-pointing-sft-swiftpointing_demo_normalizedimage-pointing-w-explanations-sft-swiftpixmo-pointingpointing_demo_diversepixmo_pixmo_pointing_cleaned
pixmo_pixmo_pointing_cleaned
The pixmo_pixmo_pointing__x family of the ElliotVL supervised-fine-tuning pool, after VLM cleaning.
images
9,797
QA turns
148,602
answers rewritten by the cleaning pass
2,943
QA created by the cleaning pass (new_qa)
38,700 (26.0%)
shards
5
How this was cleaned
A vision-language model read each image together with its QA and judged the item. The pass is
not a filter that only removes rows — it rewrites answers… See the full description on the dataset page: https://huggingface.co/datasets/elliot-mllm/pixmo_pixmo_pointing_cleaned.arrow_pointing_extrapolation
Arrow Pointing Extrapolation
This dataset contains the exact images used for the extrapolation experiments in pLSTM.
It is a synthetic dataset of arrows pointing to circles and should measure how well an image model can learn the classification
'if the arrow points to the circle' at small (192x192) scales and extrapolate/generalize (without previous resizing of the image input)
to larger scales (384x384).
Note that for the correct validation and test extrapolation subsets, you… See the full description on the dataset page: https://huggingface.co/datasets/ml-jku/arrow_pointing_extrapolation.image-pointing-w-explanations-prompt-single-color-sft-swiftpointing_cv_datasetMMEB-eval-Visual7W-Pointing-beirMMEB-eval-Visual7W-Pointing-beir-v3image-pointing-w-explanations-agree-disagree-sft-swiftMMEB-eval-Visual7W-Pointing-beir-v2lab-pointing-sft-swift
