CoolFace
8 results

AVQA

gijs /avqa-processedaudio10K<n<100K0 likes5.3k downloads1y agoHugging FaceUnFaZeD07 /Music-AVQAtabular10K<n<100K0 likes3.1k downloads7mo agoHugging Facejuyil /AVQA-videos AVQA — Audio-Visual Question Answering (videos + annotations) A drop-in package of the AVQA dataset (Yang et al., ACM MM 2022): real-life audio-visual question answering over short in-the-wild clips. The original release ships only the QA annotations and expects users to collect the source videos from VGGSound themselves. This repository bundles the source video clips together with the official train/val annotations, so the dataset is usable without any YouTube scraping.… See the full description on the dataset page: https://huggingface.co/datasets/juyil/AVQA-videos.tabularvisual-question-answering10K<n<100K1 likes2.9k downloads4mo agoHugging FaceJoysw909 /AVQA Summary | 摘要 This dataset is collected from the AVQA training subset (train_qa.json). We converted the data to the R1-AQA format, where each line in the text file represents a JSON object with specific keys. The AVQA training set originally consists of approximately 40k samples. However, we use only about 38k samples because some data sources have become invalid (e.g. link failure, or less than 10 seconds). Given that there is no quick link to the audio mentioned in the above two… See the full description on the dataset page: https://huggingface.co/datasets/Joysw909/AVQA.audioquestion-answering10K<n<100K2 likes1.8k downloads10mo agoHugging FaceNight-Quiet /AVQA-Video0 likes452 downloads1y agoHugging Facemteb /MUSIC-AVQA_cls-preprocessedaudio1K<n<10K0 likes239 downloads7mo agoHugging Face