datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
simplemath-400kUnroll-Qwen2.5-7B-Instruct_1754915847_eval_6a28_math500_simple-avg_num_prune_ffn_5_run-002
chengfu0118/Unroll-Qwen2.5-7B-Instruct_1754915847_eval_6a28_math500_simple-avg_num_prune_ffn_5_run-002
Precomputed model outputs for evaluation.
Evaluation Results
MATH500
Accuracy: 30.00%
Accuracy
Questions Solved
Total Questions
30.00%
150
500
Unroll-Qwen2.5-7B-Instruct_1754915889_eval_6a28_math500_simple-avg_num_prune_ffn_6_run-002
chengfu0118/Unroll-Qwen2.5-7B-Instruct_1754915889_eval_6a28_math500_simple-avg_num_prune_ffn_6_run-002
Precomputed model outputs for evaluation.
Evaluation Results
MATH500
Accuracy: 4.00%
Accuracy
Questions Solved
Total Questions
4.00%
20
500
Unroll-Qwen2.5-7B-Instruct_1754915759_eval_6a28_math500_simple-avg_num_prune_ffn_3_run-002
chengfu0118/Unroll-Qwen2.5-7B-Instruct_1754915759_eval_6a28_math500_simple-avg_num_prune_ffn_3_run-002
Precomputed model outputs for evaluation.
Evaluation Results
MATH500
Accuracy: 42.20%
Accuracy
Questions Solved
Total Questions
42.20%
211
500
Unroll-Qwen2.5-7B-Instruct_1754915716_eval_6a28_math500_simple-avg_num_prune_ffn_2_run-002
chengfu0118/Unroll-Qwen2.5-7B-Instruct_1754915716_eval_6a28_math500_simple-avg_num_prune_ffn_2_run-002
Precomputed model outputs for evaluation.
Evaluation Results
MATH500
Accuracy: 46.80%
Accuracy
Questions Solved
Total Questions
46.80%
234
500
Unroll-Qwen2.5-7B-Instruct_1754915801_eval_6a28_math500_simple-avg_num_prune_ffn_4_run-002
chengfu0118/Unroll-Qwen2.5-7B-Instruct_1754915801_eval_6a28_math500_simple-avg_num_prune_ffn_4_run-002
Precomputed model outputs for evaluation.
Evaluation Results
MATH500
Accuracy: 38.60%
Accuracy
Questions Solved
Total Questions
38.60%
193
500
simple-math-Qwen-1.5B-rolloutssimplemath
