datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
Visual-SearchVisual_Search
🌟 This repo contains part of the training dataset for model ThinkMorph-7B.
Dataset Description
We create an enriched interleaved dataset centered on four representative tasks requiring varying degrees of visual engagement and cross-modal interactions, including Jigsaw Assembly, Spatial Navigation, Visual Search and Chart Refocus.
Statistics
Dataset Usage
Data Downloading… See the full description on the dataset page: https://huggingface.co/datasets/ThinkMorph/Visual_Search.visual-search-dataset
Visual Search Image Dataset
Overview
This dataset contains 10,343 real-world images organized into 139 semantic categories.
Images represent animals, vehicles, people, food, nature scenes, technology, architecture, and everyday environments.
The dataset was created for experiments in semantic image retrieval and multimodal embedding search using models such as CLIP and vector search libraries like FAISS.
Images were automatically collected using the Python icrawler… See the full description on the dataset page: https://huggingface.co/datasets/darshvit20/visual-search-dataset.visual_search_klab
Visual Search Asymmetry: Deep Nets and Humans Share Similar Inherent Biases
This dataset contains the required data for running the Visual Search Experiment for our NeurIPS paper: [Visual Search Asymmetry: Deep Nets and Humans Share Similar Inherent Biases]
Usage instructions can be found on this GitHub Repo
Citations
Shashi Kant Gupta, Mengmi Zhang, Chia-Chien Wu, Jeremy M. Wolfe, & Gabriel Kreiman (2021). Visual Search Asymmetry: Deep Nets and Humans Share Similar… See the full description on the dataset page: https://huggingface.co/datasets/shashikg/visual_search_klab.zb_visual_searchvisual_search_datavg_visual_searchde_visual_searchvisual-search-catalogde_visual_search_filteredvg_visual_search_filteredmo_visual_search_filtered
