datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
LaTeX_OCR1% sampled from https://huggingface.co/datasets/linxy/LaTeX_OCR
Radiology_mini0.33% sampled from https://huggingface.co/datasets/eltorio/ROCOv2-radiology
llava-instruct-mix-vsft-miniOriginally from https://huggingface.co/datasets/HuggingFaceH4/llava-instruct-mix-vsft but 0.33% randomnly sampled
Agri-CM3-Vision-Unsloth
Agri-CM3-Vision-Unsloth
An English vision-only dataset prepared for fine-tuning Vision Language Models (VLMs) with Unsloth.
This is a reformatted subset of the original HIT-Kwoo/Agri-CM3 benchmark — a large-scale Chinese agricultural pest and disease dataset. We extracted only the English vision splits, keeping all image-based tasks and formatting them in the ShareGPT conversation format compatible with Unsloth fine-tuning.
Purpose
This dataset was specifically… See the full description on the dataset page: https://huggingface.co/datasets/farukalamai/Agri-CM3-Vision-Unsloth.pcbslm-static-v2-unsloth-vlm
PCBSLM static-v2 Unsloth VLM
Portable multimodal Unsloth dataset for PCB layout/document-grounded training.
The JSONL splits use Unsloth/Gemma-style chat messages:
{
"messages": [
{"role": "user", "content": [
{"type": "image", "image": "assets/raw_docs/.../images/page.png"},
{"type": "text", "text": "instruction..."}
]},
{"role": "assistant", "content": [
{"type": "text", "text": "{...json answer...}"}
]}
]
}
Files… See the full description on the dataset page: https://huggingface.co/datasets/henry1477/pcbslm-static-v2-unsloth-vlm.ogiri-bokete-unsloth-vlm
Japanese Bokete Ogiri — Unsloth VLM format
YANS-official/ogiri-bokete を、UnslothのVision SFTで扱える会話形式に変換した非公開用データセットです。
各JSONLレコードは「1画像 + 1回答」です。
{
"messages": [
{"role": "user", "content": [
{"type": "image", "image": "images/124469.jpg"},
{"type": "text", "text": "この画像のお題に対して、面白い一言を1つ返してください。"}
]},
{"role": "assistant", "content": [
{"type": "text", "text": "..."}
]}
]
}
Files
train.jsonl: 1,678 records / 630 prompts… See the full description on the dataset page: https://huggingface.co/datasets/beezza/ogiri-bokete-unsloth-vlm.RLAIF-V-Dataset
Dataset Card for RLAIF-V-Dataset
GitHub | Paper
News:
[2024.05.28] 📃 Our paper is accesible at arxiv now!
[2024.05.20] 🔥 Our data is used in MiniCPM-Llama3-V 2.5, which represents the first end-side MLLM achieving GPT-4V level performance!
Dataset Summary
RLAIF-V-Dataset is a large-scale multimodal feedback dataset. The dataset provides high-quality feedback with a total number of 83,132 preference pairs, where the instructions are collected from a diverse… See the full description on the dataset page: https://huggingface.co/datasets/unsloth/RLAIF-V-Dataset.unsloth_Radiology_miniunsloth-LaTeX_OCR-for-mlx-vlmReformated clone for MLX-VLM Dev
unslothImages0.33% sampled from https://huggingface.co/datasets/eltorio/ROCOv2-radiology
Qwen2-VL-2B-Instruct-unsloth-bnb-4bit_Qari-0.2-eval-tashkilunsloth-cua-demonstrationsHigh-quality trajectories for UI agent training.
mascot-unsloth-vision-dataset
