datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
Vietnamese-lmms-lab-LLaVA-Video-178K-gg-translated
Dataset Card for 5CD-AI/Vietnamese-lmms-lab-LLaVA-Video-178K-gg-translated
This translated dataset includes:
LLaVA-Video-178K: 178,509 caption entries, 960,791 open-ended QA (question and answer) items, and 196,198 multiple-choice QA items.
The video source of the original dataset is in this repo: lmms-lab/LLaVA-Video-178K
LLaVA-Video-small-swift
Dataset Card LLaVA-Video-small-swift
Small subset of LLaVA-Video-178K for educational purposes to learn how to fine-tune video models.
llava-video-jsonLLaVA-Video-2_3_m_youtube_mc-qwen_filter_1videollava-7b_clip_eval
