CoolFace
8 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01appledora /DORI-instruction-tuning-dataset DORI Spatial Reasoning Instruction Dataset Dataset Description This dataset contains instruction tuning data for spatial reasoning tasks across multiple question types and visual datasets. Dataset Structure Dataset Splits train: 26,626 samples test: 6,672 samples Total: 33,298 samples Question Types q1 q2 q3 q4 q5 q6 q7 Source Datasets 3d_future cityscapes coco coco_space_sea get_3d jta kitti nocs_real objectron… See the full description on the dataset page: https://huggingface.co/datasets/appledora/DORI-instruction-tuning-dataset.imagevisual-question-answering10K<n<100K1 likes451 downloads1y agoHugging Face02openllmplayground /pandagpt_visual_instruction_dataset[Dataset Details] This dataset is constructed by combining LLaVA Visual Instruct 150K and the dataset released by MiniGPT-4. [License] Attribution-NonCommercial 4.0 International It should abide by the policy of OpenAI: https://openai.com/policies/terms-of-use Intended use Primary intended uses: The primary use of this dataset is research on large multimodal models and chatbots. Primary intended users: The primary intended users of the model are researchers and hobbyists in… See the full description on the dataset page: https://huggingface.co/datasets/openllmplayground/pandagpt_visual_instruction_dataset.image14 likes209 downloads3y agoHugging Face03kfkas /hansung_visual_instruction_datasetimagen<1K0 likes19 downloads2y agoHugging Face04kfkas /visual_instruction_datasetimagen<1K0 likes16 downloads2y agoHugging Face05VLAI-AIVN /vietnamtourism-instruction-datasetgated VietnamTourism LLaVA Instruction VietnamTourism LLaVA Instruction is a Vietnamese multimodal instruction-tuning dataset built from public tourism article images and metadata, then converted into LLaVA-style multi-turn conversations. The dataset is intended for research and internal experimentation on Vietnamese visual question answering, image-grounded dialogue, and tourism-domain multimodal assistants. Dataset Summary split samples train 5,978… See the full description on the dataset page: https://huggingface.co/datasets/VLAI-AIVN/vietnamtourism-instruction-dataset.imagevisual-question-answering1K<n<10K0 likes12 downloads5mo agoHugging Face06RepalleHanumansai /car-image-instruction-datasetimage1K<n<10K0 likes6 downloads3mo agoHugging Face07lambdaeranga /h-reflex-instruction-datasetimagen<1K0 likes5 downloads1y agoHugging Face08Anshar0 /simple-instruction-datasetimagen<1K0 likes2 downloads9mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.