datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
SpaRRTa
SpaRRTa: A Synthetic Benchmark for Evaluating Spatial Intelligence in Visual Foundation Models
SpaRRTa is a synthetic benchmark that probes whether Visual Foundation Models (VFMs) —
such as DINO, DINOv2/v3, MAE, CroCo, VGGT, SPA and CLIP — encode the spatial relations
between objects in a scene, rather than only their semantic identity.
📄 Paper: arXiv:2601.11729
💻 Code: github.com/gmum/SpaRRTa
🧱 Real-world (lego) split: turhancan97/SpaRRTa-Lego
🔬 Attention-analysis split (images +… See the full description on the dataset page: https://huggingface.co/datasets/turhancan97/SpaRRTa.wroclaw-dwarves
Wrocław Dwarves — Fine-Grained Instance Retrieval
Wrocław is scattered with several hundred small bronze dwarf statues — krasnale — installed
across the city since 2001. They are individually sculpted, but they share a visual vocabulary:
the same scale, the same material, the same crouching poses and hand-held props. Telling one
from another is therefore a fine-grained instance recognition problem rather than a
classification one. Every statue here belongs to the same semantic… See the full description on the dataset page: https://huggingface.co/datasets/turhancan97/wroclaw-dwarves.Turkish-VLM-Mix-BenchmarkThis is a Turkish multimodal (image-text-text triplets) dataset consisting of Turkish translated samples from the datasets google/docci, tomg-group-umd/pixelprose, detection-datasets/coco, rafaelpadilla/coco2017, liuhaotian/LLaVA-Instruct-150K, liuhaotian/LLaVA-CC3M-Pretrain-595K, and HuggingFaceM4/FairFace.
The labels are in Turkish and the dataset is in an instruction-tuning format with separate columns for prompts and completion labels.
The original labels (except… See the full description on the dataset page: https://huggingface.co/datasets/ucsahin/Turkish-VLM-Mix-Benchmark.Turkish-Visual-Reasoning-Dataset
Turkish Visual Reasoning Dataset
The Turkish Visual Reasoning Dataset is a Turkish multimodal reasoning dataset designed to evaluate and improve the abstract reasoning capabilities of Vision-Language Models (VLMs).
It was created by adapting established visual reasoning benchmarks into Turkish and combining them with original Turkish BİLSEM preparation questions. The dataset targets challenging reasoning tasks such as logical pattern discovery, spatial reasoning, analogical… See the full description on the dataset page: https://huggingface.co/datasets/Berkesule/Turkish-Visual-Reasoning-Dataset.Farfetch.Product.prices.Turkey
Farfetch web scraped data
About the website
The Ecommerce industry in the EMEA region, specifically in Turkey, has shown significant growth in recent years. Companies like Farfetch have established their presence in the competitive Turkish market. Turkeys rapid digital transformation, favourable demographics, and high levels of internet penetration have resulted in the boom of online retail. This proves profitable for ecommerce platforms specialising in luxury fashion… See the full description on the dataset page: https://huggingface.co/datasets/DBQ/Farfetch.Product.prices.Turkey.Net.a.Porter.Product.prices.Turkey
Net-a-Porter web scraped data
About the website
The Net-a-Porter company operates within the Ecommerce industry in the EMEA region, particularly in Turkey. This expansive sector primarily focuses on the buying and selling of goods and services through the internet, showcasing an array of products from various providers on a global scale. Over the past few years, the Ecommerce sector in Turkey has experienced rapid growth, leading to a highly competitive market. Companies… See the full description on the dataset page: https://huggingface.co/datasets/DBQ/Net.a.Porter.Product.prices.Turkey.
