CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01lmms-lab /LLaVA-Video-178K Dataset Card for LLaVA-Video-178K Uses This dataset is used for the training of the LLaVA-Video model. We only allow the use of this dataset for academic research and education purpose. For OpenAI GPT-4 generated data, we recommend the users to check the OpenAI Usage Policy. Data Sources For the training of LLaVA-Video, we utilized video-language data from five primary sources: LLaVA-Video-178K: This dataset includes 178,510 caption entries, 960,792 open-ended… See the full description on the dataset page: https://huggingface.co/datasets/lmms-lab/LLaVA-Video-178K.textvisual-question-answering1M<n<10M202 likes38k downloads2y agoHugging Face02Ahmed-Nasri /llava-video-178k-siglip-tokens-ftov-new LLaVA-Video-178K SigLIP Token Cache (LLaVA-OV fine-tuned vision tower) Derived data (vision-encoder features of video frames), not a redistribution of the source videos. Source: lmms-lab/LLaVA-Video-178K -- its card restricts use to academic research and education, and its annotations come from GPT-4-class models (see the OpenAI usage policy). Complete: 85000 clips. Subset Folders: 0_30_s_academic_v0_1, 0_30_s_youtube_v0_1, 30_60_s_academic_v0_1… See the full description on the dataset page: https://huggingface.co/datasets/Ahmed-Nasri/llava-video-178k-siglip-tokens-ftov-new.video5 likes7k downloads5d agoHugging Face03LanguageBind /Video-LLaVA19 likes441 downloads3y agoHugging Face045CD-AI /Vietnamese-lmms-lab-LLaVA-Video-178K-gg-translated Dataset Card for 5CD-AI/Vietnamese-lmms-lab-LLaVA-Video-178K-gg-translated This translated dataset includes: LLaVA-Video-178K: 178,509 caption entries, 960,791 open-ended QA (question and answer) items, and 196,198 multiple-choice QA items. The video source of the original dataset is in this repo: lmms-lab/LLaVA-Video-178K textvisual-question-answering1M<n<10M1 likes264 downloads2y agoHugging Face05malterei /LLaVA-Video-large-swift Dataset Card LLaVA-Video-medium-swift A subset of LLaVA-Video-178K for educational purposes to learn how to fine-tune video models. videovisual-question-answeringn<1K1 likes94 downloads2y agoHugging Face06malterei /LLaVA-Video-small-swift Dataset Card LLaVA-Video-small-swift Small subset of LLaVA-Video-178K for educational purposes to learn how to fine-tune video models. textvisual-question-answeringn<1K2 likes85 downloads2y agoHugging Face07divyjx /VideoLLaVA_dataset0 likes67 downloads3y agoHugging Face08aowen14 /tl-llava-video-scratch TL LLAVA VIDEO visual-question-answering0 likes58 downloads2y agoHugging Face0934data /llava-video-178kvideo1K<n<10K0 likes51 downloads3mo agoHugging Face10malterei /LLaVA-Video-medium-swiftvideon<1K0 likes43 downloads2y agoHugging Face11Mitzi4132 /VideoLLava MVBench We introduce a novel static-to-dynamic method for defining temporal-related tasks. By converting static tasks into dynamic ones, we facilitate systematic generation of video tasks necessitating a wide range of temporal abilities, from perception to cognition. Guided by task definitions, we then automatically transform public video annotations into multiple-choice QA for task evaluation. This unique paradigm enables efficient creation of MVBench with minimal manual… See the full description on the dataset page: https://huggingface.co/datasets/Mitzi4132/VideoLLava.1K<n<10K0 likes41 downloads2y agoHugging Face12farewellthree /llava-video-jsontext1M<n<10M1 likes22 downloads2y agoHugging Face13eagle0504 /llava-video-text-dataset eagle0504/llava-video-text-dataset This is a tiny LLaVA dataset with exactly four video samples for training. Field video_url: Video URLs (MP4/GIF format) Field conversation: LLaVA conversation format with user/assistant roles Field num_frames: Number of frames per video (5) Dataset Structure Each sample contains a conversation in LLaVA format: { "video_url": "https://example.com/video.mp4", "conversation": [ { "role": "user", "content": [… See the full description on the dataset page: https://huggingface.co/datasets/eagle0504/llava-video-text-dataset.imagen<1K0 likes20 downloads11mo agoHugging Face14xxtars /Video-R1-LLaVA-Video-83k-woopvideo10K<n<100K0 likes18 downloads10mo agoHugging Face15namzakku /Video-LLaVA-json0 likes17 downloads2y agoHugging Face16lhbit20010120 /VideoLLaVA_Train0 likes17 downloads2y agoHugging Face17ApolloVideo /llava_video_subsettext100K<n<1M0 likes16 downloads7mo agoHugging Face18jayzhu486 /LLaVA-Video-178K-subset0 likes15 downloads5mo agoHugging Face19victor-wang902 /video-llava-processed0 likes14 downloads2y agoHugging Face20gpupk2020 /videollava-7b_tm05_eval_clip0 likes10 downloads2y agoHugging Face21vid-modeling /llava_video_max_256_frame_fps1image100K<n<1M0 likes8 downloads2y agoHugging Face22Xiaodong /LLaVA-Video-2_3_m_youtube_mc-qwen_filter_1text1K<n<10K0 likes8 downloads2y agoHugging Face23ngqtrung /video-r1-llava-mc-v1 --- language: - en license: apache-2.0 size_categories: - 10K<n<100K task_categories: - video-text-to-text tags: - multimodal-rl - qwen3-vl - gspo - grpo --- # ngqtrung/video-r1-llava-mc-v1 Curated v1 dataset for multimodal RL fine-tuning of Qwen3-VL-4B-Instruct. | Property | Value | |---|---| | Rows | 72421 | | Modality | video | | Split | train | | Schema | verl-ready (prompt + images + videos +… See the full description on the dataset page: https://huggingface.co/datasets/ngqtrung/video-r1-llava-mc-v1.text10K<n<100K0 likes6 downloads5mo agoHugging Face24ngqtrung /video-r1-llava-free-v1 --- language: - en license: apache-2.0 size_categories: - 1K<n<10K task_categories: - video-text-to-text tags: - multimodal-rl - qwen3-vl - gspo - grpo --- # ngqtrung/video-r1-llava-free-v1 Curated v1 dataset for multimodal RL fine-tuning of Qwen3-VL-4B-Instruct. | Property | Value | |---|---| | Rows | 9860 | | Modality | video | | Split | train | | Schema | verl-ready (prompt + images + videos +… See the full description on the dataset page: https://huggingface.co/datasets/ngqtrung/video-r1-llava-free-v1.text1K<n<10K0 likes6 downloads5mo agoHugging Face25gpupk2020 /videollava-7b_clip_evaltext10K<n<100K0 likes3 downloads2y agoHugging Face26vectoryyyy /llava_videotext1M<n<10M0 likes3 downloads7mo agoHugging Face27Pavan1996 /Video_Llava0 likes2 downloads2y agoHugging Face28weili-0234 /llava-video-178k-frames0 likes2 downloads1y agoHugging Face29ningchu /videollava_10p_subsetimage0 likes2 downloads5mo agoHugging Face30FrankvLee /videollava_fx0 likes1 downloads2y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.