CoolFace
Datasetpublic

vaishnavikedar4/MCIF

Dataset Description, Collection, and Source MCIF (Multimodal Crosslingual Instruction Following) is a multilingual human-annotated benchmark based on scientific talks that is designed to evaluate instruction-following in crosslingual, multimodal settings over both short- and long-form inputs. MCIF spans three core modalities -- speech, vision, and text -- and four diverse languages (English, German, Italian, and Chinese), enabling a comprehensive evaluation of MLLMs'… See the full description on the dataset page: https://huggingface.co/datasets/vaishnavikedar4/MCIF.

sourceHugging Facecc-by-4.0updated 9mo agoView on Hugging Face
0likes39downloads
1 commits on main
f3fd74f9mo ago

Duplicate from FBK-MT/MCIF

vaishnavikedar4, danniliu