CoolFace
Datasetpublic

Animas1024/MarineEval

Dataset Card for MarineEval The work introduces MarineEval, the first large-scale benchmark specifically designed to evaluate the marine understanding capabilities of Vision-Language Models (VLMs). MarineEval contains 2,000 expert-verified image-based question–answer pairs across 7 task dimensions and 20 domain-specific capacity dimensions, emphasizing specialized marine knowledge, visual reasoning, and real-world complexity. Through comprehensive benchmarking of 17… See the full description on the dataset page: https://huggingface.co/datasets/Animas1024/MarineEval.

sourceHugging Facecc-by-4.0updated 6mo agoView on Hugging Face
0likes43downloads
1 commits on main
3d633ca6mo ago

Duplicate from WongYukKwan/MarineEval

Animas1024, WongYukKwan