CoolFace
Datasetpublic

FreedomIntelligence/MileBench

MileBench Introduction We introduce MileBench, a pioneering benchmark designed to test the MultImodal Long-contExt capabilities of MLLMs. This benchmark comprises not only multimodal long contexts, but also multiple tasks requiring both comprehension and generation. We establish two distinct evaluation sets, diagnostic and realistic, to systematically assess MLLMs’ long-context adaptation capacity and their ability to completetasks in long-context scenarios… See the full description on the dataset page: https://huggingface.co/datasets/FreedomIntelligence/MileBench.

sourceHugging Facecc-by-2.0updated 2y agoView on Hugging Face
9likes1.9kdownloads
fileMileBench_part0.tar.gz3.70 GBdownload
fileMileBench_part1.tar.gz2.44 GBdownload
fileMileBench_part2.tar.gz1.88 GBdownload
fileMileBench_part3.tar.gz2.29 GBdownload
fileMileBench_part4.tar.gz2.34 GBdownload
fileMileBench_part5.tar.gz3.14 GBdownload

FreedomIntelligence/MileBench · main · files are served by the source, never re-hosted here