CoolFace
Modelpublic

cesun/ThinkEdit-deepseek-qwen-32b

sourceHugging Facemitupdated 1y agoView on Hugging Face
2likes22downloads
Model Card

Repository for:

ThinkEdit-deepseek-qwen-32b

(We also release ThinkEdit versions for ThinkEdit-deepseek-qwen-1.5b, ThinkEdit-deepseek-llama3-8b, and ThinkEdit-deepseek-qwen-14b.)

Authors: Chung-En Sun, Ge Yan, Tsui-Wei Weng Paper: ThinkEdit: Interpretable Weight Editing to Mitigate Overly Short Thinking in Reasoning Models

Github: https://github.com/Trustworthy-ML-Lab/ThinkEdit


Introduction

Reasoning-augmented models sometimes fail by generating overly short, abstract chain-of-thought (CoT) reasoning, hurting their accuracy.

ThinkEdit is a lightweight weight-editing method that:

  • —Identifies ~4% of "short reasoning" attention heads
  • —Edits only ~0.2% of total parameters
  • —Removes the "short reasoning" direction from their output
  • —Boosts performance, especially on cases with short reasoning traces

Full Performance Results

1. Overall Accuracy

ModelGSM8KMMLU Elementary MathMATH-Level1MATH-Level5MATH-500
deepseek-qwen-32b92.97 ± 0.3995.93 ± 0.8396.41 ± 0.4591.27 ± 0.5391.62 ± 0.58
ThinkEdit-deepseek-qwen-32b95.25 ± 0.2598.02 ± 0.3196.02 ± 0.4291.31 ± 0.5091.60 ± 0.65
deepseek-qwen-14b90.80 ± 0.3695.08 ± 0.6596.32 ± 0.3590.25 ± 0.7291.48 ± 0.55
ThinkEdit-deepseek-qwen-14b93.78 ± 0.5096.56 ± 0.8496.38 ± 0.5291.03 ± 0.4491.92 ± 0.63
deepseek-llama3-8b82.26 ± 0.9196.01 ± 0.6293.46 ± 0.8485.49 ± 0.8387.26 ± 1.16
ThinkEdit-deepseek-llama3-8b89.44 ± 0.5596.19 ± 0.7394.44 ± 0.3186.49 ± 0.5488.06 ± 1.09
deepseek-qwen-1.5b79.15 ± 1.0868.52 ± 1.5693.00 ± 0.3375.48 ± 0.9082.22 ± 1.29
ThinkEdit-deepseek-qwen-1.5b84.56 ± 0.7990.66 ± 0.9793.66 ± 0.6275.05 ± 0.8282.24 ± 0.89

2. Accuracy on Short Reasoning Cases (Top 5% / 10% / 20%)

ModelGSM8KMMLU Elementary MathMATH-Level1MATH-Level5MATH-500
deepseek-qwen-32b98.31 / 97.18 / 96.2097.78 / 97.03 / 95.87100.00 / 100.00 / 98.9793.03 / 96.36 / 97.3586.40 / 92.00 / 94.00
ThinkEdit-deepseek-qwen-32b98.92 / 97.71 / 97.8397.78 / 97.57 / 97.20100.00 / 100.00 / 98.7498.03 / 98.64 / 97.9992.00 / 94.40 / 95.80
deepseek-qwen-14b96.31 / 95.65 / 92.9393.89 / 96.22 / 95.6099.52 / 99.30 / 97.7089.39 / 94.32 / 96.2586.40 / 91.40 / 93.50
ThinkEdit-deepseek-qwen-14b96.31 / 96.18 / 96.7797.78 / 95.14 / 96.5399.53 / 98.62 / 98.6796.67 / 97.88 / 98.1191.20 / 93.20 / 95.00
deepseek-llama3-8b88.92 / 87.18 / 85.8297.22 / 96.49 / 96.8097.14 / 94.88 / 94.8378.64 / 88.79 / 93.4182.00 / 81.40 / 88.30
ThinkEdit-deepseek-llama3-8b97.08 / 95.27 / 93.9597.78 / 98.65 / 97.87100.00 / 99.30 / 98.6295.61 / 96.89 / 97.1292.80 / 93.60 / 94.40
deepseek-qwen-1.5b88.46 / 87.48 / 85.0262.78 / 62.16 / 60.5397.62 / 95.12 / 93.9191.52 / 95.00 / 95.7282.40 / 89.80 / 93.40
ThinkEdit-deepseek-qwen-1.5b92.62 / 92.90 / 92.3287.78 / 88.11 / 88.6795.71 / 95.58 / 96.4495.15 / 96.59 / 97.2790.80 / 92.00 / 94.20

Usage

The usage of ThinkEdit models is exactly the same as the original deepseek-distilled models.

Citation

bibtex
@misc{sun2025thinkedit,
      title={ThinkEdit: Interpretable Weight Editing to Mitigate Overly Short Thinking in Reasoning Models}, 
      author={Chung-En Sun and Ge Yan and Tsui-Wei Weng},
      year={2025},
      eprint={2503.22048},
      archivePrefix={arXiv},
      primaryClass={cs.CL},
      url={https://arxiv.org/abs/2503.22048}, 
}