roverx12345/jointavbench
JointAVBench: A Benchmark for Joint Audio-Visual Reasoning Evaluation Overview JointAVBench is a benchmark for evaluating omni-modal large language models on joint audio-visual reasoning tasks. Each multiple-choice question is designed to require both visual and auditory information. This repository contains the audited release of JointAVBench under the roverx12345 namespace. The benchmark keeps the original 2,853-question split while refining answer… See the full description on the dataset page: https://huggingface.co/datasets/roverx12345/jointavbench.
01k
No card is published for this repository, or it could not be fetched from Hugging Face right now.
