FreedomIntelligence/MileBench
MileBench Introduction We introduce MileBench, a pioneering benchmark designed to test the MultImodal Long-contExt capabilities of MLLMs. This benchmark comprises not only multimodal long contexts, but also multiple tasks requiring both comprehension and generation. We establish two distinct evaluation sets, diagnostic and realistic, to systematically assess MLLMs’ long-context adaptation capacity and their ability to completetasks in long-context scenarios… See the full description on the dataset page: https://huggingface.co/datasets/FreedomIntelligence/MileBench.
Update README.md
Delete MileBench.zip
Update README.md
MileBench
Update README.md
Update README.md
Update README.md
Update README.md
Update README.md
Update README.md
Update README.md
Update README.md
Upload 6 files
Update README.md
Delete data
Upload dataset (part 00001-of-00002)
Upload dataset (part 00000-of-00002)
Rename MLBench.zip to MileBench.zip
upload MLBench.zip
Upload dataset (part 00001-of-00002)
Upload dataset (part 00000-of-00002)
initial commit
