eewer/qwen3-4b-thinking-sft-v54-raw2030-strictpassed-processed
Qwen3 4B Thinking SFT v54 Processed Training View This dataset is the processed and filtered training view used by the v54 Qwen3-4B-Thinking SFT recipe. It starts from eewer/swerebench-traces-raw-source-targeted-limitations-compaction-full-20260616-2030 and uses the strict-passed raw2030 mini-swe aligned view. Rows are compressed JSONL.zst files under data/. Each row contains a top-level messages column, optional tools, and scalar source mapping fields such as source_uuid… See the full description on the dataset page: https://huggingface.co/datasets/eewer/qwen3-4b-thinking-sft-v54-raw2030-strictpassed-processed.
Qwen3 4B Thinking SFT v54 Processed Training View
This dataset is the processed and filtered training view used by the v54 Qwen3-4B-Thinking SFT recipe. It starts from eewer/swerebench-traces-raw-source-targeted-limitations-compaction-full-20260616-2030 and uses the strict-passed raw2030 mini-swe aligned view.
Rows are compressed JSONL.zst files under data/. Each row contains a top-level messages column, optional tools, and scalar source mapping fields such as source_uuid, task_id, source_shard, and source_row_index when available. Assistant turns that v54 did not train on have loss: false.
Important counts:
- Rows: 8,261
- Input training shards: 64
- Assistant turns: 316,104
- Trainable assistant turns: 311,073
- Masked assistant turns: 5,031
Default repo id: eewer/qwen3-4b-thinking-sft-v54-raw2030-strictpassed-processed.
