CoolFace
Datasetpublic

phiyodr/InpaintCOCO

InpaintCOCO - Fine-grained multimodal concept understanding (for color, size, and COCO objects) Dataset Summary A data sample contains 2 images and 2 corresponding captions that differ only in one object, the color of an object, or the size of an object. Many multimodal tasks, such as Vision-Language Retrieval and Visual Question Answering, present results in terms of overall performance. Unfortunately, this approach overlooks more nuanced concepts, leaving us… See the full description on the dataset page: https://huggingface.co/datasets/phiyodr/InpaintCOCO.

sourceHugging Faceupdated 2y agoView on Hugging Face
5likes5.1kdownloads
6 commits on main
1ffac842y ago

Update README.md

phiyodr
4fd4a062y ago

Update README.md

phiyodr
3927dec3y ago

Update README.md

phiyodr
f906b6c3y ago

Init

phiyodr
c56e3193y ago

Upload dataset

phiyodr
e2882143y ago

initial commit

phiyodr