Yong-Hoon/vlm-sft-mix-en-ko-26-06
VLM SFT Mix — English + Korean (26-06) A unified multimodal supervised fine-tuning (SFT) mix for vision-language models (Gemma / LLaVA family). It combines vision (image–conversation) and text (instruction / reasoning) data in parallel English and Korean. Images are embedded as bytes inside the parquet files, so the dataset loads directly with 🤗 datasets — no separate image files to download. Summary Configs (sub-datasets) 175 Examples (rows) 42… See the full description on the dataset page: https://huggingface.co/datasets/Yong-Hoon/vlm-sft-mix-en-ko-26-06.
032
No card is published for this repository, or it could not be fetched from Hugging Face right now.
