CoolFace
Datasetpublic

eewer/qwen3-4b-thinking-sft-v54-raw2030-strictpassed-processed

Qwen3 4B Thinking SFT v54 Processed Training View This dataset is the processed and filtered training view used by the v54 Qwen3-4B-Thinking SFT recipe. It starts from eewer/swerebench-traces-raw-source-targeted-limitations-compaction-full-20260616-2030 and uses the strict-passed raw2030 mini-swe aligned view. Rows are compressed JSONL.zst files under data/. Each row contains a top-level messages column, optional tools, and scalar source mapping fields such as source_uuid… See the full description on the dataset page: https://huggingface.co/datasets/eewer/qwen3-4b-thinking-sft-v54-raw2030-strictpassed-processed.

sourceHugging Faceotherupdated 3mo agoView on Hugging Face
0likes44downloads
Dataset Card

Qwen3 4B Thinking SFT v54 Processed Training View

This dataset is the processed and filtered training view used by the v54 Qwen3-4B-Thinking SFT recipe. It starts from eewer/swerebench-traces-raw-source-targeted-limitations-compaction-full-20260616-2030 and uses the strict-passed raw2030 mini-swe aligned view.

Rows are compressed JSONL.zst files under data/. Each row contains a top-level messages column, optional tools, and scalar source mapping fields such as source_uuid, task_id, source_shard, and source_row_index when available. Assistant turns that v54 did not train on have loss: false.

Important counts:

  • —Rows: 8,261
  • —Input training shards: 64
  • —Assistant turns: 316,104
  • —Trainable assistant turns: 311,073
  • —Masked assistant turns: 5,031

Default repo id: eewer/qwen3-4b-thinking-sft-v54-raw2030-strictpassed-processed.