VLAI-AIVN/vietnamtourism-instruction-dataset
VietnamTourism LLaVA Instruction VietnamTourism LLaVA Instruction is a Vietnamese multimodal instruction-tuning dataset built from public tourism article images and metadata, then converted into LLaVA-style multi-turn conversations. The dataset is intended for research and internal experimentation on Vietnamese visual question answering, image-grounded dialogue, and tourism-domain multimodal assistants. Dataset Summary split samples train 5,978… See the full description on the dataset page: https://huggingface.co/datasets/VLAI-AIVN/vietnamtourism-instruction-dataset.
Conversations for this repository live on Hugging Face.
CoolFace shows imported repositories read-only. Posting into someone else’s repository from here would need an authorised integration and the account holder’s consent, so the link goes to the source instead.
Open discussions on Hugging Face