tinyllava
tiny_llavavideoTinyLLaVA-Video
This dataset combines data from multiple sources for pre-training and fine-tuning.
Pretrain Data: Four subsets of LLaVA-Video-178K (0_30_s_academic_v0_1, 30_60_s_academic_v0_1, 0_30_s_youtube_v0_1, 30_60_s_youtube_v0_1), supplemented with filtered Video-LLaVA data (https://huggingface.co/datasets/LanguageBind/Video-LLaVA) and data from Valley (https://github.com/RupertLuo/Valley). The video data can be downloaded from the linked datasets, and cleaned annotations are provided… See the full description on the dataset page: https://huggingface.co/datasets/pbwpbw/tiny_llavavideo.TinyLLaVA-Video-R1-training-dataTinyLLaVA-Video-R1
We select multiple choice questions from the NextQA subset of LLaVA-Video-178K as training data. To maintain manageable training time with limited computational resources, we only
choose the subset of data with a duration of 0 to 30 seconds, which contains 5,496 samples.
In addition, we manually annotate 16 samples for cold-starting and provide the annotations.
Organize Data
Organize the files and annotation files as follows in path/to/your/dataset:
dataset
├──… See the full description on the dataset page: https://huggingface.co/datasets/Zhang199/TinyLLaVA-Video-R1-training-data.TinyLLaVA-Video-v1-training-dataTinyLLaVA-Video
This dataset combines data from multiple sources for pre-training and fine-tuning.
Pretrain Data: Four subsets of LLaVA-Video-178K (0_30_s_academic_v0_1, 30_60_s_academic_v0_1, 0_30_s_youtube_v0_1, 30_60_s_youtube_v0_1), supplemented with filtered Video-LLaVA data (https://huggingface.co/datasets/LanguageBind/Video-LLaVA) and data from Valley (https://github.com/RupertLuo/Valley). The video data can be downloaded from the linked datasets, and cleaned annotations are provided… See the full description on the dataset page: https://huggingface.co/datasets/Zhang199/TinyLLaVA-Video-v1-training-data.tinyllavaCheck out the configuration reference at https://huggingface.co/docs/hub/spaces-config-reference
TinyLLAVA-Training-Datatinyllava_train_eval_data
