5CD-AI/Viet-OpenViVQA-gemini-VQA
Dataset Overview This dataset is was created from 8024 Vietnamese 🇻🇳 images in train split of dataset OpenViVQA [1]. Each image has been analyzed and annotated using advanced Visual Question Answering (VQA) techniques to produce a comprehensive dataset. There is a set of over 63,798 detailed descriptions and query-based questions and answers generated by the Gemini 1.5 Flash model, currently Google's leading model on the WildVision Arena Leaderboard. This results in a richly… See the full description on the dataset page: https://huggingface.co/datasets/5CD-AI/Viet-OpenViVQA-gemini-VQA.
Conversations for this repository live on Hugging Face.
CoolFace shows imported repositories read-only. Posting into someone else’s repository from here would need an authorised integration and the account holder’s consent, so the link goes to the source instead.
Open discussions on Hugging Face