VLAI-AIVN/vietnamtourism-instruction-dataset
VietnamTourism LLaVA Instruction VietnamTourism LLaVA Instruction is a Vietnamese multimodal instruction-tuning dataset built from public tourism article images and metadata, then converted into LLaVA-style multi-turn conversations. The dataset is intended for research and internal experimentation on Vietnamese visual question answering, image-grounded dialogue, and tourism-domain multimodal assistants. Dataset Summary split samples train 5,978… See the full description on the dataset page: https://huggingface.co/datasets/VLAI-AIVN/vietnamtourism-instruction-dataset.
013
No card is published for this repository, or it could not be fetched from Hugging Face right now.
