CoolFace
24 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01MAmmoTH-VL /MAmmoTH-VL-Instruct-12M MAmmoTH-VL-Instruct-12M 🏠 Homepage | 🤖 MAmmoTH-VL-8B | 💻 Code | 📄 Arxiv | 📕 PDF | 🖥️ Demo Introduction Our simple yet scalable visual instruction data rewriting pipeline consists of three steps: manual data source collection, rewriting using MLLMs/LLMs, and filtering via the same MLLM as a judge. Examples below illustrate transformations in math and science categories, showcasing detailed, step-by-step responses. The data distribution of… See the full description on the dataset page: https://huggingface.co/datasets/MAmmoTH-VL/MAmmoTH-VL-Instruct-12M.imagevisual-question-answering10M<n<100M67 likes3.7k downloads2y agoHugging Face02PolyU-ChenLab /ET-Instruct-164K E.T. Instruct 164K arXiv | Project Page | GitHub E.T. Instruct 164K is a large-scale instruction-tuning dataset tailored for fine-grained event-level and time-sensitive video understanding. It contains 101K meticulously collected videos under diverse domains and 9 event-level understanding tasks with well-designed instruction-response pairs. The average video length is around 146 seconds. 📦 Download Dataset You may download the dataset using the following command.… See the full description on the dataset page: https://huggingface.co/datasets/PolyU-ChenLab/ET-Instruct-164K.text100K<n<1M5 likes2.1k downloads2y agoHugging Face03kaiyuyue /llava-1.5-665k-instructionsThis dataset repository, LLaVA-1.5-665K-Instructions, is notably utilized in the paper Zero-Shot Vision Encoder Grafting via LLM Surrogates. The official code repository for the paper can be found here: https://github.com/kaiyuyue/zero LLaVA-1.5-665K-Instructions This dataset repo contains the entire LLaVA-1.5-665K-Instructions in one place, including images and text sequences. The images are in train_split/*.tars and the text sequences are in jsons: llava_v1_5_mix665k.json is the… See the full description on the dataset page: https://huggingface.co/datasets/kaiyuyue/llava-1.5-665k-instructions.imagevisual-question-answering100K<n<1M10 likes639 downloads1y agoHugging Face04Mizukiluke /ureader-instruction-1.0image10K<n<100K15 likes284 downloads3y agoHugging Face05anas-awadalla /open-lm-instruction-datatext1K<n<10K0 likes177 downloads3y agoHugging Face06InnovatorLab /Innovator-VL-Instruct-Scienceimage100K<n<1M1 likes139 downloads7mo agoHugging Face07Intel /fivl-instruct FiVL-Instruct Dataset FiVL: A Frameword for Improved Vision-Language Alignment introduces grounded datasets for both training and evaluation, building upon existing vision-question-answer and instruction datasets Each sample in the original datasets was augmented with key expressions, along with their corresponding bounding box indices and segmentation masks within the images. Dataset Details Creators: Intel Labs Version: 1.0 (Updated: 2024-12-18) License: CC BY 4.0… See the full description on the dataset page: https://huggingface.co/datasets/Intel/fivl-instruct.texttext-generation1M<n<10M0 likes75 downloads2y agoHugging Face08sc-genrm-scaling /GPQA_verifications_GenRM-Base_Llama-3.3-70B-Instructtext1K<n<10K0 likes39 downloads1y agoHugging Face09sc-genrm-scaling /MATH128_verifications_GenRM-FT_Llama-3.1-8B-Instructtext10K<n<100K1 likes27 downloads1y agoHugging Face10sc-genrm-scaling /MATH128_verifications_GenRM-FT_Qwen-2.5-7B-Instructtext10K<n<100K1 likes19 downloads1y agoHugging Face11CJY /Chinese-Dialogue-180k-Instruct-Audioaudio100K<n<1M4 likes14 downloads1y agoHugging Face12sc-genrm-scaling /MATH128_Solutions_Llama-3.1-8B-Instructtextn<1K0 likes13 downloads1y agoHugging Face13sc-genrm-scaling /MATH128_Solutions_Qwen-2.5-7B-Instructtextn<1K0 likes13 downloads1y agoHugging Face14sc-genrm-scaling /MATH128_Solutions_Llama-3.3-70B-Instructtextn<1K0 likes13 downloads1y agoHugging Face15ymhao /X2I-instructgated X2I-instruct (WebP-compressed) This dataset is a WebP q=80 re-encoded version of yzwang/X2I-subject-driven, packed into plain .tar shards (each <= 30 GiB). All images (PNG / JPEG / WebP) were re-encoded as WebP at quality 80. Total size shrunk from ~1.79 TB -> ~140 GB (~13x compression). JSONL metadata files are rewritten to point at the new .webp paths (see *.webp.jsonl). All instructions, sample structure, and the per-subdir directory layout are preserved. Layout… See the full description on the dataset page: https://huggingface.co/datasets/ymhao/X2I-instruct.imagetext-to-image1M<n<10M0 likes13 downloads4mo agoHugging Face16lucasjin /M4-Instruct-Multiimage100K<n<1M0 likes12 downloads1y agoHugging Face17alignmentforever /0504_combination_instruction_wikihowimage10K<n<100K0 likes9 downloads1y agoHugging Face18sc-genrm-scaling /MATH128_verifications_Llama-3.3-70B-Instruct_GenRM-Basetext10K<n<100K1 likes7 downloads1y agoHugging Face19marcosremar2 /instructs2s-webdatasetaudio100K<n<1M0 likes6 downloads8mo agoHugging Face20arianhosseini /lcb128_llama3-8B-instruct_256samples_ver32_temp0-7text10K<n<100K0 likes5 downloads2y agoHugging Face21nishadsinghi /aime24_qwen2.5-7b_ver_Llama-3.1-8B-Instruct_data-qwen_25_7b_gpt_4o_verify_train_e3_LR-5e-7_7Klentext1K<n<10K0 likes4 downloads2y agoHugging Face22aklein4 /mmlu-chat-Llama-3.2-Instructtext10K<n<100K0 likes3 downloads1y agoHugging Face23aklein4 /benchmark-chat-Llama-3.2-Instructtext10K<n<100K0 likes3 downloads1y agoHugging Face24gujintao /InstructIRimage100K<n<1M0 likes2 downloads8mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.