CoolFace
Datasetpublic

Self-Improving-Coding-Agents/SI2CA-Training-Trajectories

Dataset Card for SI2CA-Training-Trajectories [🌐 Website] β€’ [πŸ€— Dataset] β€’ [πŸ“œ Paper] β€’ [🐱 GitHub] πŸ’‘ Introduction This dataset consists of 32,340 coding-agent trajectories generated by Qwen3.5-122B-A10B on the same 10,780 executable Python SWE tasks under the three trajectory-curation settings of Section 4.4 of the paper: standard sampling, full self-judgement, and an efficient discovered strategy found by the recursive self-improvement framework. Each task is… See the full description on the dataset page: https://huggingface.co/datasets/Self-Improving-Coding-Agents/SI2CA-Training-Trajectories.

sourceHugging Facecc-by-4.0updated 3d agoView on Hugging Face
0likes81downloads
7 commits on main
69931be3d ago

Update README.md

MasterVito
c128e523d ago

Dataset card: rename to SI2CA-Training-Trajectories, drop upstream arXiv links

MasterVito
aca454d4d ago

Dataset card: centered title and links, justified paragraphs, emoji section headers

MasterVito
636b0bb4d ago

Dataset card: concise rewrite (links, settings, two tables, field list, single citation)

MasterVito
06dbacb4d ago

Dataset card: field-level corrections (timeout notice text, fail_to_pass median, judge-log edge cases), window definition, weighted-total formula

MasterVito
9ed97274d ago

Add files using upload-large-folder tool

MasterVito
5adc30b4d ago

initial commit

MasterVito