CoolFace
16 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01AdaptLLM /food-visual-instructions Adapting Multimodal Large Language Models to Domains via Post-Training (EMNLP 2025) This repos contains the food visual instructions for post-training MLLMs in our paper: On Domain-Specific Post-Training for Multimodal Large Language Models. The main project page is: Adapt-MLLM-to-Domains Data Information Using our visual instruction synthesizer, we generate visual instruction tasks based on the image-caption pairs from extended Recipe1M+ dataset. These synthetic… See the full description on the dataset page: https://huggingface.co/datasets/AdaptLLM/food-visual-instructions.imagevisual-question-answering100K<n<1M3 likes1.2k downloads1y agoHugging Face02openllmplayground /pandagpt_visual_instruction_dataset[Dataset Details] This dataset is constructed by combining LLaVA Visual Instruct 150K and the dataset released by MiniGPT-4. [License] Attribution-NonCommercial 4.0 International It should abide by the policy of OpenAI: https://openai.com/policies/terms-of-use Intended use Primary intended uses: The primary use of this dataset is research on large multimodal models and chatbots. Primary intended users: The primary intended users of the model are researchers and hobbyists in… See the full description on the dataset page: https://huggingface.co/datasets/openllmplayground/pandagpt_visual_instruction_dataset.image14 likes209 downloads3y agoHugging Face03AdaptLLM /remote-sensing-visual-instructions Adapting Multimodal Large Language Models to Domains via Post-Training (EMNLP 2025) This repos contains the remote-sensing visual instructions for post-training MLLMs in our paper: On Domain-Specific Post-Training for Multimodal Large Language Models. The main project page is: Adapt-MLLM-to-Domains Data Information Using our visual instruction synthesizer, we generate visual instruction tasks based on the image-caption pairs from NWPU-Captions, RSICD, RSITMD… See the full description on the dataset page: https://huggingface.co/datasets/AdaptLLM/remote-sensing-visual-instructions.imagevisual-question-answering10K<n<100K8 likes103 downloads1y agoHugging Face04AdaptLLM /biomed-visual-instructions Adapting Multimodal Large Language Models to Domains via Post-Training (EMNLP 2025) This repos contains the biomedicine visual instructions for post-training MLLMs in our paper: On Domain-Specific Post-Training for Multimodal Large Language Models. The main project page is: Adapt-MLLM-to-Domains Data Information Using our visual instruction synthesizer, we generate visual instruction tasks based on the image-caption pairs from PubMedVision (referred to as PMC_refined… See the full description on the dataset page: https://huggingface.co/datasets/AdaptLLM/biomed-visual-instructions.textvisual-question-answering1M<n<10M5 likes96 downloads1y agoHugging Face05leeroy-jankins /DOD-Instruction-5040-02-Visual-Information DoD Visual Information Question-Answer Dataset Maintainer: Terry Eppler Owner: US Federal Government Dataset Summary This dataset contains 250 document-grounded question-and-answer records based on DoD Instruction 5040.02, “Visual Information (VI),” dated October 27, 2011, and incorporating Change 2 effective April 20, 2018. The source establishes Department of Defense policy, responsibilities, and procedures for managing visual-information records… See the full description on the dataset page: https://huggingface.co/datasets/leeroy-jankins/DOD-Instruction-5040-02-Visual-Information.documentquestion-answering0 likes45 downloads2mo agoHugging Face06jiyatai /visual_instruction_tuning_ID_referenceimage1 likes24 downloads2y agoHugging Face07DebasishDhal99 /bengali-visual-genome-instruction-settext10K<n<100K0 likes24 downloads11mo agoHugging Face08DebasishDhal99 /malayalam-visual-genome-instruction-settext10K<n<100K0 likes20 downloads11mo agoHugging Face09kfkas /hansung_visual_instruction_datasetimagen<1K0 likes19 downloads2y agoHugging Face10DebasishDhal99 /odia-visual-genome-instruction-settext10K<n<100K1 likes19 downloads11mo agoHugging Face11kfkas /hausung_visual_instruction_datasettextn<1K0 likes18 downloads2y agoHugging Face12kfkas /visual_instruction_datasetimagen<1K0 likes16 downloads2y agoHugging Face13akashuee /hindi_visual_genome_instruction_settext10K<n<100K0 likes16 downloads1y agoHugging Face14DebasishDhal99 /hindi-visual-genome-instruction-settext10K<n<100K0 likes12 downloads1y agoHugging Face15iocuydi /amharic-visual-instruction-tuningDataset used for finetuning step of Amharic llava. More details: https://arxiv.org/abs/2403.06354 0 likes9 downloads2y agoHugging Face165CD-AI /Viet-Visual-Instructionsgatedimage10K<n<100K0 likes1 downloads2y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.