CoolFace
16 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01kaiyuyue /llava-1.5-665k-instructionsThis dataset repository, LLaVA-1.5-665K-Instructions, is notably utilized in the paper Zero-Shot Vision Encoder Grafting via LLM Surrogates. The official code repository for the paper can be found here: https://github.com/kaiyuyue/zero LLaVA-1.5-665K-Instructions This dataset repo contains the entire LLaVA-1.5-665K-Instructions in one place, including images and text sequences. The images are in train_split/*.tars and the text sequences are in jsons: llava_v1_5_mix665k.json is the… See the full description on the dataset page: https://huggingface.co/datasets/kaiyuyue/llava-1.5-665k-instructions.imagevisual-question-answering100K<n<1M10 likes639 downloads1y agoHugging Face02pbwpbw /tiny_llavavideoTinyLLaVA-Video This dataset combines data from multiple sources for pre-training and fine-tuning. Pretrain Data: Four subsets of LLaVA-Video-178K (0_30_s_academic_v0_1, 30_60_s_academic_v0_1, 0_30_s_youtube_v0_1, 30_60_s_youtube_v0_1), supplemented with filtered Video-LLaVA data (https://huggingface.co/datasets/LanguageBind/Video-LLaVA) and data from Valley (https://github.com/RupertLuo/Valley). The video data can be downloaded from the linked datasets, and cleaned annotations are provided… See the full description on the dataset page: https://huggingface.co/datasets/pbwpbw/tiny_llavavideo.textvideo-text-to-text100K<n<1M0 likes320 downloads6mo agoHugging Face03lmms-lab /LLaVA-OneVision-Mid-Data Dataset Card for LLaVA-OneVision Due to unknow reasons, we are unable to process dataset with large amount into required HF format. So we directly upload the json files and image folders (compressed into tar.gz files). You can use the following link to directly download and decompress them. https://huggingface.co/datasets/lmms-lab/LLaVA-OneVision-Mid-Data/tree/main/evol_instruct We provide the whole details of LLaVA-OneVision Dataset. In this dataset, we include the data splits… See the full description on the dataset page: https://huggingface.co/datasets/lmms-lab/LLaVA-OneVision-Mid-Data.imagetext-generation100K<n<1M21 likes141 downloads2y agoHugging Face04PixArt-alpha /SAM-LLaVA-Captions10Mtext10M<n<100M64 likes128 downloads3y agoHugging Face05wusize /LLaVA-OneVision-Mid-Data-512image100K<n<1M0 likes59 downloads2y agoHugging Face06wusize /LLaVA-OneVision-Data-512image1M<n<10M0 likes45 downloads2y agoHugging Face07JiashuYang /SynTab-LLaVA-Datasetimage100K<n<1M0 likes31 downloads7mo agoHugging Face08dunghuynh /llava-odimage100K<n<1M0 likes25 downloads1y agoHugging Face09wei682 /LLaVA_data/LLaVA_data/ │ finetune/ │ ├── llava_v1_5_mix665k.json │ └── data/ │ ├── coco/ │ ├── gqa/ │ ├── nohup.out │ ├── prompts/ │ ├── textvqa/ │ ├── coco2014_val_gpt4_qa_30x3.json │ ├── coco2014_val_qa_eval/ │ ├── eval/ │ ├── LLaVA-Pretrain/ │ ├── occ_vqa/ │ ├── temp_eval/ │ └── vg/ └── pretrain/ └── blip_laion_cc_sbu_558k.json └── images/ textn<1K0 likes24 downloads2y agoHugging Face10ApolloVideo /llava_video_subsettext100K<n<1M0 likes16 downloads7mo agoHugging Face11ej2 /llava_mix665image10K<n<100K1 likes13 downloads2y agoHugging Face12vid-modeling /llava_video_max_256_frame_fps1image100K<n<1M0 likes8 downloads2y agoHugging Face13boyuzhuGPT /llavaguard-qwen3image1K<n<10K0 likes7 downloads1y agoHugging Face14abby101 /depth-maps-llavagatedimage100K<n<1M0 likes5 downloads1y agoHugging Face15the-Lin /llava_med_for_cv805image10K<n<100K0 likes3 downloads1y agoHugging Face16vectoryyyy /llava_videotext1M<n<10M0 likes3 downloads7mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.