CoolFace
Datasetpublicgated

Yong-Hoon/vlm-sft-mix-en-ko-26-06

VLM SFT Mix — English + Korean (26-06) A unified multimodal supervised fine-tuning (SFT) mix for vision-language models (Gemma / LLaVA family). It combines vision (image–conversation) and text (instruction / reasoning) data in parallel English and Korean. Images are embedded as bytes inside the parquet files, so the dataset loads directly with 🤗 datasets — no separate image files to download. Summary Configs (sub-datasets) 175 Examples (rows) 42… See the full description on the dataset page: https://huggingface.co/datasets/Yong-Hoon/vlm-sft-mix-en-ko-26-06.

sourceHugging Faceupdated 25d agoView on Hugging Face
0likes32downloads

Yong-Hoon/vlm-sft-mix-en-ko-26-06 · main · files are served by the source, never re-hosted here

This repository is gated. The listing is public, but downloading a file means accepting the publisher’s terms at Hugging Face first — the links above take you there rather than around it.