CoolFace
Datasetpublic

mjuicem/StreamingBench

StreamingBench: Assessing the Gap for MLLMs to Achieve Streaming Video Understanding 🏠 Project Page | 📄 arXiv Paper | 📦 Dataset | 🏅Leaderboard StreamingBench evaluates Multimodal Large Language Models (MLLMs) in real-time, streaming video understanding tasks. 🌟 [NEW! 2025.05.15] 🔥: Seed1.5-VL achieved ALL model SOTA with a score of 82.80 on the Proactive Output. [NEW! 2025.03.17] ⭐: ViSpeeker achieved Open-Source SOTA with a score of 61.60 on… See the full description on the dataset page: https://huggingface.co/datasets/mjuicem/StreamingBench.

sourceHugging Faceupdated 1y agoView on Hugging Face
13likes12kdownloads
fileAnomaly Context Understanding.zip8.55 GBdownload
fileEmotion Recognition.zip5.29 GBdownload
fileMisleading Context Understanding.zip5.50 GBdownload
fileMultimodal Alignment.zip8.92 GBdownload
fileProactive Output_1-25.zip4.17 GBdownload
fileProactive Output_26-50.zip4.02 GBdownload
fileReal-Time Visual Understanding_1-50.zip8.46 GBdownload
fileReal-Time Visual Understanding_101-150.zip7.65 GBdownload
fileReal-Time Visual Understanding_151-200.zip16.75 GBdownload
fileReal-Time Visual Understanding_201-250.zip2.08 GBdownload
fileReal-Time Visual Understanding_251-300.zip17.54 GBdownload
fileReal-Time Visual Understanding_301-350.zip15.75 GBdownload
fileReal-Time Visual Understanding_351-400.zip16.67 GBdownload
fileReal-Time Visual Understanding_401-450.zip16.97 GBdownload
fileReal-Time Visual Understanding_451-500.zip12.54 GBdownload
fileReal-Time Visual Understanding_51-100.zip8.11 GBdownload
fileScene Understanding_1-25.zip5.27 GBdownload
fileScene Understanding_26-50.zip7.61 GBdownload
fileSequential Question Answering_1-25.zip5.56 GBdownload
fileSequential Question Answering_26-50.zip6.85 GBdownload
fileSource Discrimination.zip4.37 GBdownload

mjuicem/StreamingBench · main · files are served by the source, never re-hosted here