datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
product-photography-v1-tiny-prompts-tasks-collage-filteredVisCoT_VStar_CollageThis is part of the training data for vSearcher introduced in "InSight-o3: Empowering Multimodal Foundation Models with Generalized Visual Search".
The data comprise collages made from a subset of images from VisualCoT and the training data of V*.
Each entry of this dataset contains a collage (with a randomly placed "core" image within it) and a QA for the core image.
The other images are filler images sampled from the same image pool as the core images.
Every image (both core and filler) is… See the full description on the dataset page: https://huggingface.co/datasets/m-Just/VisCoT_VStar_Collage.collage-layout-dataset
Collage Layout Synthetic Dataset
Synthetic photo-collage layouts for layout-quality analysis & correction,
built on a six-ingredient framework (Format, Photos, Visual Weight, Hierarchy,
Readability, Harmony). Corrector-not-generator: every collage carries a
naive (v1_center) and a corrected (fit) placement, so a model can learn the
correction. Faces are synthetically replaced (privacy-safe).
How to load
from datasets import load_dataset
ds =… See the full description on the dataset page: https://huggingface.co/datasets/jjobear/collage-layout-dataset.giantess-collageproduct-photography-v1-tiny-prompts-tasks-collagecollagecollageTestCollage_test3product-photography-v1-tiny-prompts-tasks-collage-filtered-annotatedcollage
