mhjiang0408/MAC_Bench
MAC: A Live Benchmark for Multimodal Large Language Models in Scientific Understanding 📋 Dataset Description MAC is a comprehensive live benchmark designed to evaluate multimodal large language models (MLLMs) on scientific understanding tasks. The dataset focuses on scientific journal cover understanding, providing challenging testbeds for assessing visual-textual comprehension capabilities of MLLMs in academic domains. 🎯 Tasks 1. Image-to-Text… See the full description on the dataset page: https://huggingface.co/datasets/mhjiang0408/MAC_Bench.
1190
