CyclicReflex: Improving Reasoning Models via Cyclical Reflection Token Scheduling
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Fan, Chongyu, Zhang, Yihua, Jia, Jinghan, Hero, Alfred, Liu, Sijia |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Towards LLM Unlearning Resilient to Relearning Attacks: A Sharpness-Aware Minimization Perspective and Beyond
von: Fan, Chongyu, et al.
Veröffentlicht: (2025)
von: Fan, Chongyu, et al.
Veröffentlicht: (2025)
Challenging Forgets: Unveiling the Worst-Case Forget Sets in Machine Unlearning
von: Fan, Chongyu, et al.
Veröffentlicht: (2024)
von: Fan, Chongyu, et al.
Veröffentlicht: (2024)
Beyond SFT: Reinforcement Learning for Safer Large Reasoning Models with Better Reasoning Ability
von: Jia, Jinghan, et al.
Veröffentlicht: (2025)
von: Jia, Jinghan, et al.
Veröffentlicht: (2025)
Simplicity Prevails: Rethinking Negative Preference Optimization for LLM Unlearning
von: Fan, Chongyu, et al.
Veröffentlicht: (2024)
von: Fan, Chongyu, et al.
Veröffentlicht: (2024)
SEUF: Is Unlearning One Expert Enough for Mixture-of-Experts LLMs?
von: Zhuang, Haomin, et al.
Veröffentlicht: (2024)
von: Zhuang, Haomin, et al.
Veröffentlicht: (2024)
SOUL: Unlocking the Power of Second-Order Optimization for LLM Unlearning
von: Jia, Jinghan, et al.
Veröffentlicht: (2024)
von: Jia, Jinghan, et al.
Veröffentlicht: (2024)
Reasoning Model Unlearning: Forgetting Traces, Not Just Answers, While Preserving Reasoning Skills
von: Wang, Changsheng, et al.
Veröffentlicht: (2025)
von: Wang, Changsheng, et al.
Veröffentlicht: (2025)
One Token Embedding Is Enough to Deadlock Your Large Reasoning Model
von: Zhang, Mohan, et al.
Veröffentlicht: (2025)
von: Zhang, Mohan, et al.
Veröffentlicht: (2025)
Defensive Unlearning with Adversarial Training for Robust Concept Erasure in Diffusion Models
von: Zhang, Yimeng, et al.
Veröffentlicht: (2024)
von: Zhang, Yimeng, et al.
Veröffentlicht: (2024)
Downgrade to Upgrade: Optimizer Simplification Enhances Robustness in LLM Unlearning
von: Lang, Yicheng, et al.
Veröffentlicht: (2025)
von: Lang, Yicheng, et al.
Veröffentlicht: (2025)
LLM Unlearning Under the Microscope: A Full-Stack View on Methods and Metrics
von: Fan, Chongyu, et al.
Veröffentlicht: (2025)
von: Fan, Chongyu, et al.
Veröffentlicht: (2025)
Zero Token-Driven Deep Thinking in LLMs: Unlocking the Full Potential of Existing Parameters via Cyclic Refinement
von: Li, Guanghao, et al.
Veröffentlicht: (2025)
von: Li, Guanghao, et al.
Veröffentlicht: (2025)
Improved Neural Protoform Reconstruction via Reflex Prediction
von: Lu, Liang, et al.
Veröffentlicht: (2024)
von: Lu, Liang, et al.
Veröffentlicht: (2024)
A Multiscale Geometric Method for Capturing Relational Topic Alignment
von: Hougen, Conrad D., et al.
Veröffentlicht: (2025)
von: Hougen, Conrad D., et al.
Veröffentlicht: (2025)
UnlearnCanvas: Stylized Image Dataset for Enhanced Machine Unlearning Evaluation in Diffusion Models
von: Zhang, Yihua, et al.
Veröffentlicht: (2024)
von: Zhang, Yihua, et al.
Veröffentlicht: (2024)
Vision-Language Models Can Self-Improve Reasoning via Reflection
von: Cheng, Kanzhi, et al.
Veröffentlicht: (2024)
von: Cheng, Kanzhi, et al.
Veröffentlicht: (2024)
Token Assorted: Mixing Latent and Text Tokens for Improved Language Model Reasoning
von: Su, DiJia, et al.
Veröffentlicht: (2025)
von: Su, DiJia, et al.
Veröffentlicht: (2025)
Faithfulness as Information Flow: Evaluating and Training Faithful Chain-of-Thought Reasoning
von: Jia, Jinghan, et al.
Veröffentlicht: (2026)
von: Jia, Jinghan, et al.
Veröffentlicht: (2026)
Improve Decoding Factuality by Token-wise Cross Layer Entropy of Large Language Models
von: Wu, Jialiang, et al.
Veröffentlicht: (2025)
von: Wu, Jialiang, et al.
Veröffentlicht: (2025)
Forgetting to Forget: Attention Sink as A Gateway for Backdooring LLM Unlearning
von: Shang, Bingqi, et al.
Veröffentlicht: (2025)
von: Shang, Bingqi, et al.
Veröffentlicht: (2025)
EPiC: Towards Lossless Speedup for Reasoning Training through Edge-Preserving CoT Condensation
von: Jia, Jinghan, et al.
Veröffentlicht: (2025)
von: Jia, Jinghan, et al.
Veröffentlicht: (2025)
Bring Reason to Vision: Understanding Perception and Reasoning through Model Merging
von: Chen, Shiqi, et al.
Veröffentlicht: (2025)
von: Chen, Shiqi, et al.
Veröffentlicht: (2025)
Cyclic Proofs in Hoare Logic and its Reverse
von: Brotherston, James, et al.
Veröffentlicht: (2025)
von: Brotherston, James, et al.
Veröffentlicht: (2025)
Disentangling Reasoning Tokens and Boilerplate Tokens For Language Model Fine-tuning
von: Ye, Ziang, et al.
Veröffentlicht: (2024)
von: Ye, Ziang, et al.
Veröffentlicht: (2024)
CyclicJudge: Mitigating Judge Bias Efficiently in LLM-based Evaluation
von: Zhu, Ziyi, et al.
Veröffentlicht: (2026)
von: Zhu, Ziyi, et al.
Veröffentlicht: (2026)
Complete the Cycle: Reachability Types with Expressive Cyclic References (Extended Version)
von: Deng, Haotian, et al.
Veröffentlicht: (2025)
von: Deng, Haotian, et al.
Veröffentlicht: (2025)
Instruct-of-Reflection: Enhancing Large Language Models Iterative Reflection Capabilities via Dynamic-Meta Instruction
von: Liu, Liping, et al.
Veröffentlicht: (2025)
von: Liu, Liping, et al.
Veröffentlicht: (2025)
FINEREASON: Evaluating and Improving LLMs' Deliberate Reasoning through Reflective Puzzle Solving
von: Chen, Guizhen, et al.
Veröffentlicht: (2025)
von: Chen, Guizhen, et al.
Veröffentlicht: (2025)
Learn Beyond The Answer: Training Language Models with Reflection for Mathematical Reasoning
von: Zhang, Zhihan, et al.
Veröffentlicht: (2024)
von: Zhang, Zhihan, et al.
Veröffentlicht: (2024)
Transitivity Meets Cyclicity: Explicit Preference Decomposition for Dynamic Large Language Model Alignment
von: Huang, Yucong, et al.
Veröffentlicht: (2026)
von: Huang, Yucong, et al.
Veröffentlicht: (2026)
Think Clearly: Improving Reasoning via Redundant Token Pruning
von: Choi, Daewon, et al.
Veröffentlicht: (2025)
von: Choi, Daewon, et al.
Veröffentlicht: (2025)
iCLP: Large Language Model Reasoning with Implicit Cognition Latent Planning
von: Chen, Sijia, et al.
Veröffentlicht: (2025)
von: Chen, Sijia, et al.
Veröffentlicht: (2025)
Blind Spot Navigation in Large Language Model Reasoning with Thought Space Explorer
von: Zhang, Jinghan, et al.
Veröffentlicht: (2024)
von: Zhang, Jinghan, et al.
Veröffentlicht: (2024)
Enhancing Large Language Model Reasoning via Selective Critical Token Fine-Tuning
von: Ruan, Zhiwen, et al.
Veröffentlicht: (2025)
von: Ruan, Zhiwen, et al.
Veröffentlicht: (2025)
One Token Is Enough: Improving Diffusion Language Models with a Sink Token
von: Zhang, Zihou, et al.
Veröffentlicht: (2026)
von: Zhang, Zihou, et al.
Veröffentlicht: (2026)
Efficient Reasoning Through Suppression of Self-Affirmation Reflections in Large Reasoning Models
von: Liu, Kaiyuan, et al.
Veröffentlicht: (2025)
von: Liu, Kaiyuan, et al.
Veröffentlicht: (2025)
Arithmetic in the Wild: Llama uses Base-10 Addition to Reason About Cyclic Concepts
von: Feucht, Sheridan, et al.
Veröffentlicht: (2026)
von: Feucht, Sheridan, et al.
Veröffentlicht: (2026)
RATT: A Thought Structure for Coherent and Correct LLM Reasoning
von: Zhang, Jinghan, et al.
Veröffentlicht: (2024)
von: Zhang, Jinghan, et al.
Veröffentlicht: (2024)
SEED: Accelerating Reasoning Tree Construction via Scheduled Speculative Decoding
von: Wang, Zhenglin, et al.
Veröffentlicht: (2024)
von: Wang, Zhenglin, et al.
Veröffentlicht: (2024)
Order Matters in Hallucination: Reasoning Order as Benchmark and Reflexive Prompting for Large-Language-Models
von: Xie, Zikai
Veröffentlicht: (2024)
von: Xie, Zikai
Veröffentlicht: (2024)
Ähnliche Einträge
-
Towards LLM Unlearning Resilient to Relearning Attacks: A Sharpness-Aware Minimization Perspective and Beyond
von: Fan, Chongyu, et al.
Veröffentlicht: (2025) -
Challenging Forgets: Unveiling the Worst-Case Forget Sets in Machine Unlearning
von: Fan, Chongyu, et al.
Veröffentlicht: (2024) -
Beyond SFT: Reinforcement Learning for Safer Large Reasoning Models with Better Reasoning Ability
von: Jia, Jinghan, et al.
Veröffentlicht: (2025) -
Simplicity Prevails: Rethinking Negative Preference Optimization for LLM Unlearning
von: Fan, Chongyu, et al.
Veröffentlicht: (2024) -
SEUF: Is Unlearning One Expert Enough for Mixture-of-Experts LLMs?
von: Zhuang, Haomin, et al.
Veröffentlicht: (2024)