review-artifacts/locus-bench
LOCUS-Bench (anonymized release for peer review) A benchmark for embodied multi-robot task planning that grades two difficulties separately. Axis S (state judgment): S0 no judgment; S1 whether a single named target is already in place; S2 which of 2 to 3 conditional candidates are absent; S3 which of 3 to 6 quantified instances are unsatisfied; S4 whether invisible implies absent under occlusion, with single-frame fallback planning. Axis M (mechanical structure): M0 none; M1 an… See the full description on the dataset page: https://huggingface.co/datasets/review-artifacts/locus-bench.
0261
add train
add ood_x
add ood_m
add ood_s
add test_id
add MANIFEST.json
add LICENSE.md
add README.md
initial commit
