srqa
Datasets
All datasets matching “srqa”SRQA_Audio
SRQA Audio
SRQA Audio is the public audio asset bundle for the synthetic Spoken Reasoning Question Answering (SRQA) benchmark used in the paper Learning When to Think While Listening in Large Audio-Language Models. It provides audio files for evaluating audio-input models on spoken versions of established reasoning tasks.
Contents
Rewritten and TTS-rendered benchmark audio
The following benchmark tracks were rewritten into spoken queries and rendered with the… See the full description on the dataset page: https://huggingface.co/datasets/Oulasong/SRQA_Audio.SRQA_DatabaseFolder Description:
GT -- original Full HD uncompressed videos. In the folder corresponding to each video there is "GT.yuv" file (the video itself) and "frames" folder containing all frames of the video
GTCompressed -- videos from "GT" folder with reduced resolution and compressed with different video codecs. The folder corresponding to each of the video sequences contains folders with frames of the compressed videos. Folders are named in the following format: "<scale factor>_<codec… See the full description on the dataset page: https://huggingface.co/datasets/Divotion/SRQA_Database.
