CoolFace
6 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01MohamedRashad /Arabic-VLM-Full-Pearl 💎 The Arabic VLM Dataset (Full Pearl Edition) This repository contains the full, unreviewed dataset comprising 309K multimodal examples. This data was generated automatically using the agentic pipeline developed for the Pearl project, as described in our paper. Disclaimer: This is the raw, synthetic data that has not been subject to human review. It was generated as part of the data creation process and is released for research purposes. It may contain noise, errors, or… See the full description on the dataset page: https://huggingface.co/datasets/MohamedRashad/Arabic-VLM-Full-Pearl.imagequestion-answering100K<n<1M10 likes452 downloads10mo agoHugging Face02Kirito-Lab /VLM-ExecRouterBench VLM-ExecRouterBench An execution-oriented benchmark for cost-aware open-set VLM routing. Cost-aware routing | Open-set model onboarding | Multimodal, code, and search tasks Overview VLM-ExecRouterBench is an execution-oriented benchmark for routing vision-language model queries to a pool of candidate VLMs. Each sample is executed by multiple candidate models, producing correctness labels, inference costs, metadata… See the full description on the dataset page: https://huggingface.co/datasets/Kirito-Lab/VLM-ExecRouterBench.imagevisual-question-answering10K<n<100K0 likes226 downloads1mo agoHugging Face03YangyiYY /VLM-SFTimagetext-generation1M<n<10M2 likes69 downloads2y agoHugging Face04True2456 /iq-terrain-vlm-dataset IQ Terrain VLM Dataset A high-fidelity, mathematically pristine Vision-Language Model (VLM) dataset designed specifically to teach models the procedural graphics and raymarching techniques of Inigo Quilez. Dataset Summary Most coding datasets rely on broadly scraped, often buggy code from GitHub or StackOverflow. This dataset takes a highly targeted approach: Mathematical Ground Truth: All GLSL code and mathematical concepts are sourced directly from Inigo… See the full description on the dataset page: https://huggingface.co/datasets/True2456/iq-terrain-vlm-dataset.imagetext-generation1K<n<10K0 likes51 downloads2mo agoHugging Face05beezza /ogiri-bokete-unsloth-vlm Japanese Bokete Ogiri — Unsloth VLM format YANS-official/ogiri-bokete を、UnslothのVision SFTで扱える会話形式に変換した非公開用データセットです。 各JSONLレコードは「1画像 + 1回答」です。 { "messages": [ {"role": "user", "content": [ {"type": "image", "image": "images/124469.jpg"}, {"type": "text", "text": "この画像のお題に対して、面白い一言を1つ返してください。"} ]}, {"role": "assistant", "content": [ {"type": "text", "text": "..."} ]} ] } Files train.jsonl: 1,678 records / 630 prompts… See the full description on the dataset page: https://huggingface.co/datasets/beezza/ogiri-bokete-unsloth-vlm.imageimage-to-text1K<n<10K0 likes26 downloads2mo agoHugging Face06SOGANG-ISDS /VLM_CCAgated [!NOTE] Planned improvements: Human verification (image - keyword alignment; Q&A / translation) Report VLM performance on this dataset Include image license details in metadata We welcome your feedback! Please contact us: Lab: isds.sogang@gmail.com Maintainer: bizli0618@sogang.ac.kr VLM-CCA Korean Culture VQA Dataset Dataset Summary The Korean Culture VQA Dataset for Visual Language Model's Cultural Context Awareness (VLM-CCA) is a multimodal benchmark… See the full description on the dataset page: https://huggingface.co/datasets/SOGANG-ISDS/VLM_CCA.imagevisual-question-answering1K<n<10K3 likes7 downloads9mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.