datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
citationmapper-atom01-gemini
CitationMapper – Mapping AI Citation Visibility
Entity: CitationMapper (AI visibility tool)
Figure 1. CitationMapper logo – official brand mark.
Explore CitationMapper™ (Gemini version)
Watch the explainer (YouTube): bit.ly/cm-atom1-video-geminiDownload the video (MP4): citationmapper-explainer-ai-visibility-video-gemini.mp4
🧩 What is CitationMapper?
CitationMapper is the first prompt competition analyzer built for AI visibility.It helps SEO agencies, marketing… See the full description on the dataset page: https://huggingface.co/datasets/AIVOMeshLab/citationmapper-atom01-gemini.gemini-2.5-vs-pro-realism-ai-perception
Gemini 2.5 Flash vs Gemini 3 Pro: Photorealism Comparaison
The dataset quantifies how much better Google's newest image model, Gemini 3 Pro Image, is at generating photorealistic images compared to Gemini 2.5 Flash Image.
This text-to-image benchmark dataset contains 6428 human judgments from annotators across 50+ countries, collected in under 20 minutes using the Rapidata Python API, accessible to anyone and ideal for small to large scale evaluation.
Overview
30… See the full description on the dataset page: https://huggingface.co/datasets/indomar/gemini-2.5-vs-pro-realism-ai-perception.Viet-Handwriting-gemini-VQA
Dataset Overview
This dataset is was created from 1252 Vietnamese 🇻🇳 Handwriting images in train split of dataset Cinnamon AI Challenge - Handwriting adddress and UIT-HWDB[1] . Each handwriting image has been analyzed and annotated using advanced Visual Question Answering (VQA) techniques to produce a comprehensive dataset.
There is a set of over 8,700 labels, detailed descriptions and query-based questions and answers generated by the Gemini 1.5 Flash model, currently… See the full description on the dataset page: https://huggingface.co/datasets/5CD-AI/Viet-Handwriting-gemini-VQA.Viet-LAION-Gemini-VQA
Dataset Overview
This dataset is was created from 843,529 Vietnamese 🇻🇳 images from truongpdd/laion-2b-vietnamese-subset.
The dataset contains images from diverse domains, including journalism on social life, sports, tourism, e-commerce product images, clothing, vehicles, sketches, charts, games, technology, and more.
There is a set of 5,061,174 detailed descriptions, and questions-answers generated by the latest Gemini-1.5-Flash-002 model. This results in a richly annotated… See the full description on the dataset page: https://huggingface.co/datasets/5CD-AI/Viet-LAION-Gemini-VQA.Viet-ViTextVQA-gemini-VQA
Dataset Overview
This dataset is was created from 9594 Vietnamese 🇻🇳 images in train split of dataset ViTextVQA [1]. Each image has been analyzed and annotated using advanced Visual Question Answering (VQA) techniques to produce a comprehensive dataset.
There is a set of over 50,000 detailed descriptions and query-based questions and answers generated by the Gemini 1.5 Flash model, currently Google's leading model on the WildVision Arena Leaderboard. This results in a richly… See the full description on the dataset page: https://huggingface.co/datasets/5CD-AI/Viet-ViTextVQA-gemini-VQA.Viet-Menu-gemini-VQA
Dataset Overview
This dataset is was created from 840 Vietnamese 🇻🇳 menu images in train split of dataset in Quy Nhon AI Hackathon 2022 - Smart menu. Each menu image has been analyzed and annotated using advanced Visual Question Answering (VQA) techniques to produce a comprehensive dataset.
There is a set of 5800 detailed descriptions, extractions, and query-based questions and answers generated by the Gemini 1.5 Flash model, currently Google's leading model on the WildVision… See the full description on the dataset page: https://huggingface.co/datasets/5CD-AI/Viet-Menu-gemini-VQA.Viet-Vintext-gemini-VQA
Dataset Overview
This dataset is was created from 1056 Vietnamese 🇻🇳 images in train split of dataset VinText [1]. Each image has been analyzed and annotated using advanced Visual Question Answering (VQA) techniques to produce a comprehensive dataset.
There is a set of over 6,000 detailed descriptions and query-based questions and answers generated by the Gemini 1.5 Flash model, currently Google's leading model on the WildVision Arena Leaderboard. This results in a richly… See the full description on the dataset page: https://huggingface.co/datasets/5CD-AI/Viet-Vintext-gemini-VQA.Viet-OpenViVQA-gemini-VQA
Dataset Overview
This dataset is was created from 8024 Vietnamese 🇻🇳 images in train split of dataset OpenViVQA [1]. Each image has been analyzed and annotated using advanced Visual Question Answering (VQA) techniques to produce a comprehensive dataset.
There is a set of over 63,798 detailed descriptions and query-based questions and answers generated by the Gemini 1.5 Flash model, currently Google's leading model on the WildVision Arena Leaderboard. This results in a richly… See the full description on the dataset page: https://huggingface.co/datasets/5CD-AI/Viet-OpenViVQA-gemini-VQA.Viet-colpali_train_set-Gemini
Dataset Description
This dataset is translated from vidore/colpali_train_set, The content within the images is translated into Vietnamese, and the translated text is replaced directly into the corresponding images. Additionally, the 'query,' 'answer,' 'options,' and 'prompt' fields are translated using Gemini 1.5 Flase 002.
Example 1:
Original image => Vietnamse translated image
Query:
Những đề xuất nào Tập đoàn Y tế Liberty nên xem xét để cải thiện tỷ lệ luân… See the full description on the dataset page: https://huggingface.co/datasets/5CD-AI/Viet-colpali_train_set-Gemini.
