CoolFace
Datasetpublic

Mitzi4132/VideoLLava

MVBench We introduce a novel static-to-dynamic method for defining temporal-related tasks. By converting static tasks into dynamic ones, we facilitate systematic generation of video tasks necessitating a wide range of temporal abilities, from perception to cognition. Guided by task definitions, we then automatically transform public video annotations into multiple-choice QA for task evaluation. This unique paradigm enables efficient creation of MVBench with minimal manual… See the full description on the dataset page: https://huggingface.co/datasets/Mitzi4132/VideoLLava.

sourceHugging Faceupdated 2y agoView on Hugging Face
0likes40downloads
7 commits on main
1b6a2262y ago

Update README.md

Mitzi4132
0f88a542y ago

Upload light_yesyes.xlsx

Mitzi4132
c3153cf2y ago

Delete VideoLLava.py

Mitzi4132
cba55612y ago

Create VideoLLava.py

Mitzi4132
925f5da2y ago

Create README.md

Mitzi4132
b2edc442y ago

Add light and upload light_yes.xlsx

root
7e5cef62y ago

initial commit

Mitzi4132