CoolFace
Datasetpublic

sci-m-wang/C4-Eval

C4-Eval C4-Eval is the evaluation set for C4 Bench, a Chengyu-based benchmark for measuring whether multimodal language models can understand cross-concept creativity. The release contains the original images, the corresponding idiom answers, and the complete task-specific questions used for evaluation. 221 base items: 37 human-designed seed figures and 184 bridge-controlled synthetic figures. 1,105 evaluation instances: five task formulations for every base item. Language:… See the full description on the dataset page: https://huggingface.co/datasets/sci-m-wang/C4-Eval.

sourceHugging Faceupdated 2mo agoView on Hugging Face
0likes529downloads
9 commits on main
169c3362mo ago

Add arXiv citation

sci-m-wang
d395c462mo ago

Restore visual question answering category

sci-m-wang
b3251fa2mo ago

Fix task category to image-text-to-text

sci-m-wang
b72751a2mo ago

Publish complete C4-Eval benchmark

sci-m-wang
3a934962mo ago

Add files using upload-large-folder tool

sci-m-wang
425f7972mo ago

Add files using upload-large-folder tool

sci-m-wang
052cfc42mo ago

Add files using upload-large-folder tool

sci-m-wang
9ce94352mo ago

Add files using upload-large-folder tool

sci-m-wang
38cc2d02mo ago

initial commit

sci-m-wang