savoji/minivlm-tablevqa-sft
minivlm-tablevqa-sft Supervised fine-tuning mixture for improving visual table question answering in small vision-language models — built for the MiniVLMDocEval project to lift Qwen3.5-0.8B on TableVQABench. Why this mixture (data-driven targeting) We measured Qwen3.5-0.8B per TableVQABench sub-domain and found the weakness is Wikipedia-style visual-table lookup, not financial tables: sub-domain Qwen3.5-0.8B vwtq (Wikipedia lookup) 27.8 weakest, and… See the full description on the dataset page: https://huggingface.co/datasets/savoji/minivlm-tablevqa-sft.
Conversations for this repository live on Hugging Face.
CoolFace shows imported repositories read-only. Posting into someone else’s repository from here would need an authorised integration and the account holder’s consent, so the link goes to the source instead.
Open discussions on Hugging Face