CoolFace
Datasetpublic

SWE-bench/SWE-smith-trajectories

SWE-smith Trajectories Code • Paper • Site This dataset contains the 5017 trajectories we fine-tuned Qwen 2.5 Coder Instruct on, leading to SWE-agent-LM-32B, a coding LM agent that achieve 40.2% on SWE-bench Verified (no verifiers or multiple rollouts, just 1 attempt per instance). Trajectories were generated by running SWE-agent + Claude 3.7 Sonnet on task instances from the SWE-smith dataset.

sourceHugging Facemitupdated 1y agoView on Hugging Face
79likes14kdownloads
Dataset Card

<div align="center"> <a href="https://swesmith.com"> <img src="https://avatars.githubusercontent.com/u/189315905?s=200&v=4" alt="Logo" width="200"> <h1 align="center">SWE-smith Trajectories</h1> </a> </div> <p align="center"> <a href="https://github.com/SWE-bench/SWE-smith">Code</a> • <a href="https://arxiv.org/abs/2504.21798">Paper</a> • <a href="https://swesmith.com/">Site</a> </p>

This dataset contains the 5017 trajectories we fine-tuned Qwen 2.5 Coder Instruct on, leading to SWE-agent-LM-32B, a coding LM agent that achieve 40.2% on SWE-bench Verified (no verifiers or multiple rollouts, just 1 attempt per instance).

Trajectories were generated by running SWE-agent + Claude 3.7 Sonnet on task instances from the SWE-smith dataset.