5CD-AI/Viet-Menu-gemini-VQA
Dataset Overview This dataset is was created from 840 Vietnamese 🇻🇳 menu images in train split of dataset in Quy Nhon AI Hackathon 2022 - Smart menu. Each menu image has been analyzed and annotated using advanced Visual Question Answering (VQA) techniques to produce a comprehensive dataset. There is a set of 5800 detailed descriptions, extractions, and query-based questions and answers generated by the Gemini 1.5 Flash model, currently Google's leading model on the WildVision… See the full description on the dataset page: https://huggingface.co/datasets/5CD-AI/Viet-Menu-gemini-VQA.
18
No card is published for this repository, or it could not be fetched from Hugging Face right now.
