CoolFace
Datasetpublic

mapspatial/map-spatial-benchmark

Map-based Spatial Reasoning Benchmark A multi-view map-based spatial reasoning benchmark. Each row is one multiple-choice question instance over a registered map image; models must answer with a single option letter. Four tasks (T1–T4), four base-map views, and controlled evidence conditions (direct / query / oracle) and world perturbations (transform / world layers) allow fine-grained analysis of spatial reasoning robustness. Task overview Task Question… See the full description on the dataset page: https://huggingface.co/datasets/mapspatial/map-spatial-benchmark.

sourceHugging Facecc-by-4.0updated 10d agoView on Hugging Face
0likes651downloads
t4_oracle_supervision.jsonl4 linesDownload Raw Back to data
1version https://git-lfs.github.com/spec/v12oid sha256:05a10b15d92ecbbece67e31aa425b266cc04d141221ae8a93e176f0fbed9d4273size 165777164 
mapspatial/map-spatial-benchmark · CoolFace