zoom
Datasets
All datasets matching “zoom”ZoomBench
ZoomBench: A Fine-Grained Multimodal Perception Benchmark
📃 Paper | 🏠 Project | 🤗 Models
Overview
ZoomBench is a challenging benchmark designed to evaluate the fine-grained multimodal perception capabilities of Multimodal Large Language Models (MLLMs). It specifically targets scenarios where decisive visual evidence is small, subtle, or easily overwhelmed by global context — situations that demand "zooming-level" perception from a single full image.
It is… See the full description on the dataset page: https://huggingface.co/datasets/inclusionAI/ZoomBench.corruption-zoom_blur
Corruption Dataset: Zoom_Blur
Dataset Description
This dataset contains corrupted versions of ImageNet-1K images using zoom_blur corruption. It is part of the ImageNet-C benchmark for evaluating model robustness to common image corruptions.
Dataset Structure
Train: 1,281,167 corrupted images
Validation: 50,000 corrupted images
Classes: 1000 ImageNet-1K classes
Format: Arrow (Hugging Face Datasets)
Corruption Type: Zoom_Blur
Applies zoom blur… See the full description on the dataset page: https://huggingface.co/datasets/MarMaster/corruption-zoom_blur.ImageNet-C-zoom_blur-severity_5zoom_in-reorderzoom_in-swapQ-Zoom-Training
Q-Zoom Training Data
Curated training data for the Q-Zoom gated Region-of-Interest mechanism for Vision-Language Models. Companion dataset to the Q-Zoom release repository.
What this repo contains
This repo holds the question JSONLs and the ROI training pickles used to train the three components of Q-Zoom (SD-RPN, Post-SFT, Dynamic Gate). It is the companion to:
YuhengSSS/RoITraining — image archives (*.tar / *.zip) for COCO / GQA / OCR-VQA / DocVQA / TextVQA / ChartQA /… See the full description on the dataset page: https://huggingface.co/datasets/YuhengSSS/Q-Zoom-Training.
