CoolFace
Datasetpublic

review-artifacts/locus-bench

LOCUS-Bench (anonymized release for peer review) A benchmark for embodied multi-robot task planning that grades two difficulties separately. Axis S (state judgment): S0 no judgment; S1 whether a single named target is already in place; S2 which of 2 to 3 conditional candidates are absent; S3 which of 3 to 6 quantified instances are unsatisfied; S4 whether invisible implies absent under occlusion, with single-frame fallback planning. Axis M (mechanical structure): M0 none; M1 an… See the full description on the dataset page: https://huggingface.co/datasets/review-artifacts/locus-bench.

sourceHugging Faceotherupdated 5d agoView on Hugging Face
0likes261downloads
9 commits on main
4ff324e5d ago

add train

review-artifacts
c3bbc6d6d ago

add ood_x

review-artifacts
75bb0216d ago

add ood_m

review-artifacts
3f8f1586d ago

add ood_s

review-artifacts
908c51e6d ago

add test_id

review-artifacts
9e6e59e6d ago

add MANIFEST.json

review-artifacts
457eead6d ago

add LICENSE.md

review-artifacts
9df18a06d ago

add README.md

review-artifacts
b8d9b3e6d ago

initial commit

review-artifacts