datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
sysmlv2
Dataset Name: SysMLv2_QA_Dataset
Description:
This dataset contains over 3,000 question-answer pairs generated from a PDF introduction to SysML v2. It is designed for training and evaluating natural language understanding models, especially for tasks involving technical document comprehension, question answering, and knowledge extraction. Each entry consists of a question and a corresponding answer extracted or paraphrased from the SysML v2 introduction material.… See the full description on the dataset page: https://huggingface.co/datasets/Xizhidian/sysmlv2.sysml-v2-reasoning-benchmark
sysml-bench: SysML v2 Reasoning Benchmark
Dataset Summary
sysml-bench is a benchmark for evaluating how CLI tool configurations affect
LLM accuracy on structured systems engineering tasks. 132 tasks across 8
categories test model comprehension of SysML v2 models with varying tool
augmentation strategies.
The primary question: does giving an LLM more tools improve its ability to
answer questions about a SysML v2 model? The answer is nuanced — it depends
on the task type… See the full description on the dataset page: https://huggingface.co/datasets/nomograph/sysml-v2-reasoning-benchmark.
