SWE-bench/SWE-smith-trajectories
SWE-smith Trajectories Code • Paper • Site This dataset contains the 5017 trajectories we fine-tuned Qwen 2.5 Coder Instruct on, leading to SWE-agent-LM-32B, a coding LM agent that achieve 40.2% on SWE-bench Verified (no verifiers or multiple rollouts, just 1 attempt per instance). Trajectories were generated by running SWE-agent + Claude 3.7 Sonnet on task instances from the SWE-smith dataset.
<div align="center"> <a href="https://swesmith.com"> <img src="https://avatars.githubusercontent.com/u/189315905?s=200&v=4" alt="Logo" width="200"> <h1 align="center">SWE-smith Trajectories</h1> </a> </div> <p align="center"> <a href="https://github.com/SWE-bench/SWE-smith">Code</a> • <a href="https://arxiv.org/abs/2504.21798">Paper</a> • <a href="https://swesmith.com/">Site</a> </p>
This dataset contains the 5017 trajectories we fine-tuned Qwen 2.5 Coder Instruct on, leading to SWE-agent-LM-32B, a coding LM agent that achieve 40.2% on SWE-bench Verified (no verifiers or multiple rollouts, just 1 attempt per instance).
Trajectories were generated by running SWE-agent + Claude 3.7 Sonnet on task instances from the SWE-smith dataset.
