Qualcomm-AI-Research/QIVD
QIVD: Qualcomm Interactive Video Dataset A collection of 2,900 video clips paired with visual question-answer annotations. Each clip is associated with exactly one question drawn from one of 13 fine-grained QA categories, a full-sentence answer, a concise short answer, and a timestamp pinpointing the relevant moment in the video. Overview QIVD is a dataset and benchmark for online, situated audio-visual question answering. Unlike existing video QA benchmarks… See the full description on the dataset page: https://huggingface.co/datasets/Qualcomm-AI-Research/QIVD.
Conversations for this repository live on Hugging Face.
CoolFace shows imported repositories read-only. Posting into someone else’s repository from here would need an authorised integration and the account holder’s consent, so the link goes to the source instead.
Open discussions on Hugging Face