CoolFace
Datasetpublic

trannhiem/TranNhiem-Vietnamese-ImageText-Reasoning

TranNhiem Vietnamese Image-Text Reasoning (V-LAION) Large-scale Vietnamese multimodal reasoning: multi-turn visual question–answering grounded on natural images, where every answer ships with an explicit chain-of-thought. Reasoning traces and Answer were synthesized by Qwen3.5 over images from the LAION-derived Vi-Laion-gemini-VQA set. Curated by: Trần Nhiệm Languages: Vietnamese (vi) answers · English (en) reasoning Modality: image + text → text Records: 544,795… See the full description on the dataset page: https://huggingface.co/datasets/trannhiem/TranNhiem-Vietnamese-ImageText-Reasoning.

sourceHugging Faceotherupdated 2mo agoView on Hugging Face
7likes313downloads

trannhiem/TranNhiem-Vietnamese-ImageText-Reasoning · main · files are served by the source, never re-hosted here