ThinkEdit: Interpretable Weight Editing to Mitigate Overly Short Thinking in Reasoning Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Sun, Chung-En, Yan, Ge, Weng, Tsui-Wei |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Crafting Large Language Models for Enhanced Interpretability
von: Sun, Chung-En, et al.
Veröffentlicht: (2024)
von: Sun, Chung-En, et al.
Veröffentlicht: (2024)
Steer2Edit: From Activation Steering to Component-Level Editing
von: Sun, Chung-En, et al.
Veröffentlicht: (2026)
von: Sun, Chung-En, et al.
Veröffentlicht: (2026)
Effective Skill Unlearning through Intervention and Abstention
von: Li, Yongce, et al.
Veröffentlicht: (2025)
von: Li, Yongce, et al.
Veröffentlicht: (2025)
Concept Bottleneck Large Language Models
von: Sun, Chung-En, et al.
Veröffentlicht: (2024)
von: Sun, Chung-En, et al.
Veröffentlicht: (2024)
ReFIne: A Framework for Trustworthy Large Reasoning Models with Reliability, Faithfulness, and Interpretability
von: Sun, Chung-En, et al.
Veröffentlicht: (2025)
von: Sun, Chung-En, et al.
Veröffentlicht: (2025)
Interpretable Generative Models through Post-hoc Concept Bottlenecks
von: Kulkarni, Akshay, et al.
Veröffentlicht: (2025)
von: Kulkarni, Akshay, et al.
Veröffentlicht: (2025)
AdapThink: Adaptive Thinking Preferences for Reasoning Language Model
von: Wan, Xu, et al.
Veröffentlicht: (2025)
von: Wan, Xu, et al.
Veröffentlicht: (2025)
Think Right: Learning to Mitigate Under-Over Thinking via Adaptive, Attentive Compression
von: Singh, Joykirat, et al.
Veröffentlicht: (2025)
von: Singh, Joykirat, et al.
Veröffentlicht: (2025)
Resolving UnderEdit & OverEdit with Iterative & Neighbor-Assisted Model Editing
von: Baghel, Bhiman Kumar, et al.
Veröffentlicht: (2025)
von: Baghel, Bhiman Kumar, et al.
Veröffentlicht: (2025)
To Think or Not to Think: Exploring the Unthinking Vulnerability in Large Reasoning Models
von: Zhu, Zihao, et al.
Veröffentlicht: (2025)
von: Zhu, Zihao, et al.
Veröffentlicht: (2025)
AdaptThink: Reasoning Models Can Learn When to Think
von: Zhang, Jiajie, et al.
Veröffentlicht: (2025)
von: Zhang, Jiajie, et al.
Veröffentlicht: (2025)
Train Long, Think Short: Curriculum Learning for Efficient Reasoning
von: Hammoud, Hasan Abed Al Kader, et al.
Veröffentlicht: (2025)
von: Hammoud, Hasan Abed Al Kader, et al.
Veröffentlicht: (2025)
LLM Agents Already Know When to Call Tools -- Even Without Reasoning
von: Sun, Chung-En, et al.
Veröffentlicht: (2026)
von: Sun, Chung-En, et al.
Veröffentlicht: (2026)
OptimalThinkingBench: Evaluating Over and Underthinking in LLMs
von: Aggarwal, Pranjal, et al.
Veröffentlicht: (2025)
von: Aggarwal, Pranjal, et al.
Veröffentlicht: (2025)
Iterative Self-Tuning LLMs for Enhanced Jailbreaking Capabilities
von: Sun, Chung-En, et al.
Veröffentlicht: (2024)
von: Sun, Chung-En, et al.
Veröffentlicht: (2024)
Thinking About Thinking: SAGE-nano's Inverse Reasoning for Self-Aware Language Models
von: Jha, Basab, et al.
Veröffentlicht: (2025)
von: Jha, Basab, et al.
Veröffentlicht: (2025)
Dynamic Thinking-Token Selection for Efficient Reasoning in Large Reasoning Models
von: Guo, Zhenyuan, et al.
Veröffentlicht: (2026)
von: Guo, Zhenyuan, et al.
Veröffentlicht: (2026)
Reasoning Models Don't Just Think Longer, They Move Differently
von: Gjølbye, Anders, et al.
Veröffentlicht: (2026)
von: Gjølbye, Anders, et al.
Veröffentlicht: (2026)
MeTHanol: Modularized Thinking Language Models with Intermediate Layer Thinking, Decoding and Bootstrapping Reasoning
von: Xi, Ningyuan, et al.
Veröffentlicht: (2024)
von: Xi, Ningyuan, et al.
Veröffentlicht: (2024)
Efficient Reasoning with Hidden Thinking
von: Shen, Xuan, et al.
Veröffentlicht: (2025)
von: Shen, Xuan, et al.
Veröffentlicht: (2025)
Efficient Reasoning with Balanced Thinking
von: Li, Yulin, et al.
Veröffentlicht: (2026)
von: Li, Yulin, et al.
Veröffentlicht: (2026)
Asynchronous Reasoning: Training-Free Interactive Thinking LLMs
von: Yakushev, George, et al.
Veröffentlicht: (2025)
von: Yakushev, George, et al.
Veröffentlicht: (2025)
Missing Premise exacerbates Overthinking: Are Reasoning Models losing Critical Thinking Skill?
von: Fan, Chenrui, et al.
Veröffentlicht: (2025)
von: Fan, Chenrui, et al.
Veröffentlicht: (2025)
Thinking-Free Policy Initialization Makes Distilled Reasoning Models More Effective and Efficient Reasoners
von: Xu, Xin, et al.
Veröffentlicht: (2025)
von: Xu, Xin, et al.
Veröffentlicht: (2025)
Effectively Controlling Reasoning Models through Thinking Intervention
von: Wu, Tong, et al.
Veröffentlicht: (2025)
von: Wu, Tong, et al.
Veröffentlicht: (2025)
Reasoning Models Don't Always Say What They Think
von: Chen, Yanda, et al.
Veröffentlicht: (2025)
von: Chen, Yanda, et al.
Veröffentlicht: (2025)
More Edits, More Stable: Understanding the Lifelong Normalization in Sequential Model Editing
von: Ma, Xin, et al.
Veröffentlicht: (2026)
von: Ma, Xin, et al.
Veröffentlicht: (2026)
ThinkRouter: Efficient Reasoning via Routing Thinking between Latent and Discrete Spaces
von: Xu, Xin, et al.
Veröffentlicht: (2026)
von: Xu, Xin, et al.
Veröffentlicht: (2026)
Thinking Augmented Pre-training
von: Wang, Liang, et al.
Veröffentlicht: (2025)
von: Wang, Liang, et al.
Veröffentlicht: (2025)
Loop, Think, & Generalize: Implicit Reasoning in Recurrent-Depth Transformers
von: Kohli, Harsh, et al.
Veröffentlicht: (2026)
von: Kohli, Harsh, et al.
Veröffentlicht: (2026)
Think Beyond Size: Adaptive Prompting for More Effective Reasoning
von: R, Kamesh
Veröffentlicht: (2024)
von: R, Kamesh
Veröffentlicht: (2024)
CoRT: Code-integrated Reasoning within Thinking
von: Li, Chengpeng, et al.
Veröffentlicht: (2025)
von: Li, Chengpeng, et al.
Veröffentlicht: (2025)
DeepReview: Improving LLM-based Paper Review with Human-like Deep Thinking Process
von: Zhu, Minjun, et al.
Veröffentlicht: (2025)
von: Zhu, Minjun, et al.
Veröffentlicht: (2025)
Multiplex Thinking: Reasoning via Token-wise Branch-and-Merge
von: Tang, Yao, et al.
Veröffentlicht: (2026)
von: Tang, Yao, et al.
Veröffentlicht: (2026)
Reverse Thinking Makes LLMs Stronger Reasoners
von: Chen, Justin Chih-Yao, et al.
Veröffentlicht: (2024)
von: Chen, Justin Chih-Yao, et al.
Veröffentlicht: (2024)
Better Think Thrice: Learning to Reason Causally with Double Counterfactual Consistency
von: Lin, Victoria, et al.
Veröffentlicht: (2026)
von: Lin, Victoria, et al.
Veröffentlicht: (2026)
Think Dense, Not Long: Dynamic Decoupled Conditional Advantage for Efficient Reasoning
von: Peng, Keqin, et al.
Veröffentlicht: (2026)
von: Peng, Keqin, et al.
Veröffentlicht: (2026)
Thinking with Knowledge Graphs: Enhancing LLM Reasoning Through Structured Data
von: Wu, Xue, et al.
Veröffentlicht: (2024)
von: Wu, Xue, et al.
Veröffentlicht: (2024)
FLEKE: Federated Locate-then-Edit Knowledge Editing
von: Zhao, Zongkai, et al.
Veröffentlicht: (2025)
von: Zhao, Zongkai, et al.
Veröffentlicht: (2025)
BiasEdit: Debiasing Stereotyped Language Models via Model Editing
von: Xu, Xin, et al.
Veröffentlicht: (2025)
von: Xu, Xin, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Crafting Large Language Models for Enhanced Interpretability
von: Sun, Chung-En, et al.
Veröffentlicht: (2024) -
Steer2Edit: From Activation Steering to Component-Level Editing
von: Sun, Chung-En, et al.
Veröffentlicht: (2026) -
Effective Skill Unlearning through Intervention and Abstention
von: Li, Yongce, et al.
Veröffentlicht: (2025) -
Concept Bottleneck Large Language Models
von: Sun, Chung-En, et al.
Veröffentlicht: (2024) -
ReFIne: A Framework for Trustworthy Large Reasoning Models with Reliability, Faithfulness, and Interpretability
von: Sun, Chung-En, et al.
Veröffentlicht: (2025)