DCAgent2/g1_gptlong_plus_diverse_tezos_glm47_traces
DCAgent2/g1_gptlong_plus_diverse_tezos_glm47_traces 132,259 rows. Concatenation of: DCAgent/g1_min_episodes_e1_gpt_long_top8_glm47_traces (37,925 rows) — top8 + extra long swegym traces DCAgent/g1_diverse_tezos_top4_100k_glm47_traces (94,334 rows) — balanced 4-way mix of swesmith/issue/superuser/tezos No deduplication; shuffled with seed=42. Schema: intersection columns (agent, conversations, date, episode, model, model_provider, result, run_id, task, trial_name).… See the full description on the dataset page: https://huggingface.co/datasets/DCAgent2/g1_gptlong_plus_diverse_tezos_glm47_traces.
DCAgent2/g1gptlongplusdiversetezosglm47traces
132,259 rows. Concatenation of:
DCAgent/g1_min_episodes_e1_gpt_long_top8_glm47_traces(37,925 rows) — top8 + extra long swegym tracesDCAgent/g1_diverse_tezos_top4_100k_glm47_traces(94,334 rows) — balanced 4-way mix of swesmith/issue/superuser/tezos
No deduplication; shuffled with seed=42. Schema: intersection columns (agent, conversations, date, episode, model, model_provider, result, run_id, task, trial_name).
Subset breakdown
Source models
- gptlong → trained
DCAgent/g1_gptlong_top8_32b(Qwen3-32B, lr=4e-5, wd=0.04, gn=1e-3, 5 epochs) - diversetezos → trained `DCAgent/g1diversetezos100k_32b`
Generated for combining the two recipes' data.
Training config
The recommended SFT recipe for this dataset is the gptlong/opt100k config — included in this repo at configs/32k_base_bs96_opt100k.yaml.
Key hyperparameters:
Suggested cluster: 24 nodes × 4 GPUs = 96 GPUs, gradientaccumulationsteps=1.
