CoolFace
Datasetpublicgated

Yong-Hoon/vlm-sft-mix-en-ko-26-06

VLM SFT Mix — English + Korean (26-06) A unified multimodal supervised fine-tuning (SFT) mix for vision-language models (Gemma / LLaVA family). It combines vision (image–conversation) and text (instruction / reasoning) data in parallel English and Korean. Images are embedded as bytes inside the parquet files, so the dataset loads directly with 🤗 datasets — no separate image files to download. Summary Configs (sub-datasets) 175 Examples (rows) 42… See the full description on the dataset page: https://huggingface.co/datasets/Yong-Hoon/vlm-sft-mix-en-ko-26-06.

sourceHugging Faceupdated 24d agoView on Hugging Face
0likes32downloads
Dataset Card

No card is published for this repository, or it could not be fetched from Hugging Face right now.

Yong-Hoon/vlm-sft-mix-en-ko-26-06 · CoolFace