Mitzi4132/VideoLLava
MVBench We introduce a novel static-to-dynamic method for defining temporal-related tasks. By converting static tasks into dynamic ones, we facilitate systematic generation of video tasks necessitating a wide range of temporal abilities, from perception to cognition. Guided by task definitions, we then automatically transform public video annotations into multiple-choice QA for task evaluation. This unique paradigm enables efficient creation of MVBench with minimal manual… See the full description on the dataset page: https://huggingface.co/datasets/Mitzi4132/VideoLLava.
040
Update README.md
Upload light_yesyes.xlsx
Delete VideoLLava.py
Create VideoLLava.py
Create README.md
Add light and upload light_yes.xlsx
initial commit
