CoolFace
Datasetpublic

Alibaba-NLP/xvbench

XVBench XVBench is a benchmark for evaluating multimodal retrieval-augmented generation systems on cross-video understanding. It is introduced alongside the paper VimRAG: Navigating Massive Visual Context in Retrieval-Augmented Generation via Multimodal Memory Graph. The questions in XVBench are created based on videos from HowTo100M, a large-scale corpus of narrated instructional videos. The benchmark focuses on questions that require models or agents to retrieve and reason… See the full description on the dataset page: https://huggingface.co/datasets/Alibaba-NLP/xvbench.

sourceHugging Facecc-by-4.0updated 5mo agoView on Hugging Face
0likes37downloads
4 commits on main
35928b75mo ago

update

qiuchenwang
097c08b5mo ago

update

qiuchenwang
c5a3a126mo ago

dataset initial

Qiuchen-Wang
a55f8606mo ago

initial commit

Qiuchen-Wang