datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
conceptual-12m-mbart-50-multilingualseeingculture-benchmarkPaper | Project Page | Leaderboard | Explorer | Code | CMB, the video successor
Seeing Culture Benchmark (SCB)
Evaluating Visual Reasoning and Grounding in Cultural Context
The Seeing Culture Benchmark (SCB) evaluates cultural reasoning in vision-language models in two stages: i) selecting the correct visual option with multiple-choice visual question answering (VQA), and ii) segmenting the relevant cultural artifact as evidence of reasoning. Visual options in… See the full description on the dataset page: https://huggingface.co/datasets/Multimedia-SMU/seeingculture-benchmark.conceptual-12m-multilingual-marian-esmultimodal-time-series-forecastingmultimodal_query_rewrites
ReVision: Visual Instruction Rewriting Dataset
Dataset Summary
The ReVision dataset is a large-scale collection of task-oriented multimodal instructions, designed to enable on-device, privacy-preserving Visual Instruction Rewriting (VIR). The dataset consists of 39,000+ examples across 14 intent domains, where each example comprises:
Image: A visual scene containing relevant information.
Original instruction: A multimodal command (e.g., a spoken query referencing visual… See the full description on the dataset page: https://huggingface.co/datasets/hsiangfu/multimodal_query_rewrites.multimodal_query_rewrites
ReVision: Visual Instruction Rewriting Dataset
Dataset Summary
The ReVision dataset is a large-scale collection of task-oriented multimodal instructions, designed to enable on-device, privacy-preserving Visual Instruction Rewriting (VIR). The dataset consists of 39,000+ examples across 14 intent domains, where each example comprises:
Image: A visual scene containing relevant information.
Original instruction: A multimodal command (e.g., a spoken query referencing visual… See the full description on the dataset page: https://huggingface.co/datasets/anonymoususerrevision/multimodal_query_rewrites.JourneyBench_Multi_Image_VQA
Dataset Card for Dataset Name
This dataset card aims to be a base template for new datasets. It has been generated using this raw template.
Dataset Details
Dataset Description
Curated by: [More Information Needed]
Funded by [optional]: [More Information Needed]
Shared by [optional]: [More Information Needed]
Language(s) (NLP): [More Information Needed]
License: [More Information Needed]
Dataset Sources [optional]
Repository: [More… See the full description on the dataset page: https://huggingface.co/datasets/JourneyBench/JourneyBench_Multi_Image_VQA.hotel-multimodalMultiFakeRomimage_caption_pairs_for_multimodalrg-7wildlife-multilabel-v1
