CoolFace
Datasetpublic

dongxiaoyu/MUSE

MUSE: A CAD Design Benchmark with Multi-modal Ground Truth and Rubric-based Evaluation MUSE is a benchmark of 106 CAD design cases for evaluating language and multi-modal models on engineering-grade 3D design tasks. Each case pairs a natural-language design specification with multi-view ground-truth artefacts (2D engineering drawings + 3D rendered images) and a hand-crafted, rubric-style evaluation guide. Why this benchmark Most CAD/3D benchmarks evaluate either… See the full description on the dataset page: https://huggingface.co/datasets/dongxiaoyu/MUSE.

sourceHugging Facecc-by-4.0updated 4mo agoView on Hugging Face
3likes554downloads
discussions and pull requests

Conversations for this repository live on Hugging Face.

CoolFace shows imported repositories read-only. Posting into someone else’s repository from here would need an authorised integration and the account holder’s consent, so the link goes to the source instead.

Open discussions on Hugging Face
dongxiaoyu/MUSE · CoolFace