CoolFace
Datasetpublic

facebook/Common-O

Common-O measuring multimodal reasoning across scenes Common-O, inspired by cognitive tests for humans, probes multimodal LLMs' ability to reason across scenes by asking "what’s in common?" Common-O is comprised of household objects: We have two subsets: Common-O (3 - 8 objects) and Common-O Complex (8 - 16 objects). Multimodal LLMs excel at single image perception, but struggle with multi-scene reasoning Evaluating a Multimodal LLM on Common-O… See the full description on the dataset page: https://huggingface.co/datasets/facebook/Common-O.

sourceHugging Facemitupdated 7mo agoView on Hugging Face
12likes710downloads
settings

This repository belongs to facebook on Hugging Face.

CoolFace never edits a repository it does not host. Visibility, licence, collaborators and gating are all managed at the source.

nameCommon-O
visibilitypublic
licencemit
gatedno
ownerfacebook
Account settings
facebook/Common-O · CoolFace