CoolFace
Datasetpublic

savoji/minivlm-tablevqa-sft

minivlm-tablevqa-sft Supervised fine-tuning mixture for improving visual table question answering in small vision-language models — built for the MiniVLMDocEval project to lift Qwen3.5-0.8B on TableVQABench. Why this mixture (data-driven targeting) We measured Qwen3.5-0.8B per TableVQABench sub-domain and found the weakness is Wikipedia-style visual-table lookup, not financial tables: sub-domain Qwen3.5-0.8B vwtq (Wikipedia lookup) 27.8 weakest, and… See the full description on the dataset page: https://huggingface.co/datasets/savoji/minivlm-tablevqa-sft.

sourceHugging Faceupdated 3mo agoView on Hugging Face
0likes86downloads

savoji/minivlm-tablevqa-sft · main · files are served by the source, never re-hosted here