datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
tiny_llavavideoTinyLLaVA-Video
This dataset combines data from multiple sources for pre-training and fine-tuning.
Pretrain Data: Four subsets of LLaVA-Video-178K (0_30_s_academic_v0_1, 30_60_s_academic_v0_1, 0_30_s_youtube_v0_1, 30_60_s_youtube_v0_1), supplemented with filtered Video-LLaVA data (https://huggingface.co/datasets/LanguageBind/Video-LLaVA) and data from Valley (https://github.com/RupertLuo/Valley). The video data can be downloaded from the linked datasets, and cleaned annotations are provided… See the full description on the dataset page: https://huggingface.co/datasets/pbwpbw/tiny_llavavideo.TinyLLAVA-Training-Datatiny_llava_3tiny_llava_1tiny_llava_20240227184554tiny_llava_20240227183919tiny_llava_2tiny_llava_20240227183328
