gemma3-27b
python4-gemma3-27b-eft-v2-logsgemma-3-27b-it-eval-logs-and-scorespython4-gemma3-27b-eft-v2-evalDAPO-Gemma3-27B-PT-RL-step40-seed43-SFT-Data-32k-n4
DAPO-Gemma3-27B-PT-RL-step40-seed43-SFT-Data-32k-n4
Teacher-generated SFT/distillation data for Gemma 3 math distillation.
Source
Teacher: JWei05/dapo-gemma3-27b-pt-from-step40-seed43, subfolder step_000040
Prompts: JWei05/DAPO-OpenMathInstruct2-34k, train split
Rows: 128,000
Unique prompts: 32,000
Responses per prompt: 4
Sampling: temperature=1.0, top_p=1.0, top_k=-1, max_tokens=20480
Columns
Column
Description
messages
User prompt and teacher… See the full description on the dataset page: https://huggingface.co/datasets/JWei05/DAPO-Gemma3-27B-PT-RL-step40-seed43-SFT-Data-32k-n4.gemma-3-27b-it_lm_sys_responses_rot13_clip1024DAPO-Gemma3-27B-IT-RL-SFT-Data-correct
DAPO-Gemma3-27B-IT-RL-SFT-Data-correct
Filtered subset of
JWei05/DAPO-Gemma3-27B-IT-RL-SFT-Data:
only the teacher responses whose final answer is math_verify-correct against
the original DAPO-Math-17k ground truth.
Stats
Source rows: 69,592 (17,398 prompts × 4 teacher responses)
Kept rows: 41,831 (60.1%)
Prompts with ≥1 correct response: 13,062 / 17,398 (75.1%)
Prompts with 4/4 correct responses: 7,492 (43.1%)
Scoring
Same function as used during RL… See the full description on the dataset page: https://huggingface.co/datasets/JWei05/DAPO-Gemma3-27B-IT-RL-SFT-Data-correct.
