datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
cve_train
CVE-Factory Agent Traces
This dataset contains 4,078 distilled agent traces from 887 CVE reproduction tasks (the trainset/ split), generated using Claude Opus 4.5 with a Mini SWE-Agent harness. See CVE-Factory for the full pipeline.
Note: CVE-Factory also provides a trainset-2/ split with additional simpler tasks, which is not included in this training dataset.
🚀 Training Results
Fine-tuning on this dataset yields dramatic improvements across security benchmarks:… See the full description on the dataset page: https://huggingface.co/datasets/Luoberta/cve_train.cve_train_v1.1
CVE-Factory Agent Traces v1.1
This dataset is an expanded version of cve_train, containing 18,783 distilled agent traces for CVE reproduction tasks. The traces were generated using Claude Opus 4.5 with a Mini SWE-Agent harness through the CVE-Factory pipeline.
What's New in v1.1
Compared to cve_train (v1.0):
18.8k total samples (up from ~4k in v1.0)
+3k agentic tasks from cve_tasks_3k_compressed
Additional traces from expanded CVE task coverage
Training… See the full description on the dataset page: https://huggingface.co/datasets/Luoberta/cve_train_v1.1.SWE-kokkos-bench
SWE-kokkos-bench
SWE-kokkos-bench is a public, verifier-backed benchmark of 100 repository-level
software-engineering tasks mined from merged pull requests in the Kokkos ecosystem. Each task
starts from the parent revision of a real pull request and asks an agent to implement the
corresponding change. Correctness is checked by task-specific build and regression commands
against a held-out test patch.
The dataset supports two complementary interfaces:
this Hugging Face… See the full description on the dataset page: https://huggingface.co/datasets/luosuu/SWE-kokkos-bench.ruozhiba受COIG-CQIA启发,构建类似数据集,但答案风格相对更简洁。
弱智吧精选问题数据来自github提供的疑问句,调用GPT-4获取答案,并过滤掉明显拒答的回复。
