CoolFace
Datasetpublic

moondream/ssv2-3x3

SSV2 3x3 A multi-frame action dataset for teaching a model to look across several frames and describe the action it sees. Each image is a 3x3 collage built from frames sampled from the original Something-Something V2 videos. Splits Split Samples train 150000 test 1000 Columns image: 3x3 frame collage label_text: action label text annotation_text: action template text video_id: original video identifier

sourceHugging Faceupdated 6mo agoView on Hugging Face
0likes46downloads
discussions and pull requests

Conversations for this repository live on Hugging Face.

CoolFace shows imported repositories read-only. Posting into someone else’s repository from here would need an authorised integration and the account holder’s consent, so the link goes to the source instead.

Open discussions on Hugging Face
moondream/ssv2-3x3 · CoolFace