cmudrc/Material_Selection_Eval
A benchmark designed to facilitate evaluation and modify the behavior of a foundation model through different existing techniques in the context of material selection for conceptual design. The data is collected by conducting a survey of experts in the field of material selection. The same questions mentioned in keyquestions.csv are asked to experts. This can be used to evaluate a Language model performance and its spread compared to a human evaluation. To get into a more detailed explanation… See the full description on the dataset page: https://huggingface.co/datasets/cmudrc/Material_Selection_Eval.
149
