Gespeichert in:
| Hauptverfasser: | Jia, Jinghan, Reisizadeh, Hadi, Fan, Chongyu, Baracaldo, Nathalie, Hong, Mingyi, Liu, Sijia |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2506.04205 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Reasoning Model Unlearning: Forgetting Traces, Not Just Answers, While Preserving Reasoning Skills
von: Wang, Changsheng, et al.
Veröffentlicht: (2025)
von: Wang, Changsheng, et al.
Veröffentlicht: (2025)
Beyond SFT: Reinforcement Learning for Safer Large Reasoning Models with Better Reasoning Ability
von: Jia, Jinghan, et al.
Veröffentlicht: (2025)
von: Jia, Jinghan, et al.
Veröffentlicht: (2025)
Towards LLM Unlearning Resilient to Relearning Attacks: A Sharpness-Aware Minimization Perspective and Beyond
von: Fan, Chongyu, et al.
Veröffentlicht: (2025)
von: Fan, Chongyu, et al.
Veröffentlicht: (2025)
EPiC: Efficient Video Camera Control Learning with Precise Anchor-Video Guidance
von: Wang, Zun, et al.
Veröffentlicht: (2025)
von: Wang, Zun, et al.
Veröffentlicht: (2025)
Leak@$k$: Unlearning Does Not Make LLMs Forget Under Probabilistic Decoding
von: Reisizadeh, Hadi, et al.
Veröffentlicht: (2025)
von: Reisizadeh, Hadi, et al.
Veröffentlicht: (2025)
BLUR: A Bi-Level Optimization Approach for LLM Unlearning
von: Reisizadeh, Hadi, et al.
Veröffentlicht: (2025)
von: Reisizadeh, Hadi, et al.
Veröffentlicht: (2025)
CyclicReflex: Improving Reasoning Models via Cyclical Reflection Token Scheduling
von: Fan, Chongyu, et al.
Veröffentlicht: (2025)
von: Fan, Chongyu, et al.
Veröffentlicht: (2025)
Defensive Unlearning with Adversarial Training for Robust Concept Erasure in Diffusion Models
von: Zhang, Yimeng, et al.
Veröffentlicht: (2024)
von: Zhang, Yimeng, et al.
Veröffentlicht: (2024)
WAGLE: Strategic Weight Attribution for Effective and Modular Unlearning in Large Language Models
von: Jia, Jinghan, et al.
Veröffentlicht: (2024)
von: Jia, Jinghan, et al.
Veröffentlicht: (2024)
Rethinking Muon Beyond Pretraining: Spectral Failures and High-Pass Remedies for VLA and RLVR
von: Fan, Chongyu, et al.
Veröffentlicht: (2026)
von: Fan, Chongyu, et al.
Veröffentlicht: (2026)
Downgrade to Upgrade: Optimizer Simplification Enhances Robustness in LLM Unlearning
von: Lang, Yicheng, et al.
Veröffentlicht: (2025)
von: Lang, Yicheng, et al.
Veröffentlicht: (2025)
Simplicity Prevails: Rethinking Negative Preference Optimization for LLM Unlearning
von: Fan, Chongyu, et al.
Veröffentlicht: (2024)
von: Fan, Chongyu, et al.
Veröffentlicht: (2024)
Invariance Makes LLM Unlearning Resilient Even to Unanticipated Downstream Fine-Tuning
von: Wang, Changsheng, et al.
Veröffentlicht: (2025)
von: Wang, Changsheng, et al.
Veröffentlicht: (2025)
Parrot: A Training Pipeline Enhances Both Program CoT and Natural Language CoT for Reasoning
von: Jin, Senjie, et al.
Veröffentlicht: (2025)
von: Jin, Senjie, et al.
Veröffentlicht: (2025)
A Statistical Physics of Language Model Reasoning
von: Carson, Jack David, et al.
Veröffentlicht: (2025)
von: Carson, Jack David, et al.
Veröffentlicht: (2025)
CoT-based Synthesizer: Enhancing LLM Performance through Answer Synthesis
von: Zhang, Bohan, et al.
Veröffentlicht: (2025)
von: Zhang, Bohan, et al.
Veröffentlicht: (2025)
Challenging Forgets: Unveiling the Worst-Case Forget Sets in Machine Unlearning
von: Fan, Chongyu, et al.
Veröffentlicht: (2024)
von: Fan, Chongyu, et al.
Veröffentlicht: (2024)
C$^{2}$TC: A Training-Free Framework for Efficient Tabular Data Condensation
von: Xu, Sijia, et al.
Veröffentlicht: (2026)
von: Xu, Sijia, et al.
Veröffentlicht: (2026)
Subspace Control: Turning Constrained Model Steering into Controllable Spectral Optimization
von: Huang, Yancheng, et al.
Veröffentlicht: (2026)
von: Huang, Yancheng, et al.
Veröffentlicht: (2026)
Reasoning as an Attack Surface: Adaptive Evolutionary CoT Jailbreaks for LLMs
von: Li, Jianan, et al.
Veröffentlicht: (2026)
von: Li, Jianan, et al.
Veröffentlicht: (2026)
Faithfulness as Information Flow: Evaluating and Training Faithful Chain-of-Thought Reasoning
von: Jia, Jinghan, et al.
Veröffentlicht: (2026)
von: Jia, Jinghan, et al.
Veröffentlicht: (2026)
Galaxy: A Cognition-Centered Framework for Proactive, Privacy-Preserving, and Self-Evolving LLM Agents
von: Bao, Chongyu, et al.
Veröffentlicht: (2025)
von: Bao, Chongyu, et al.
Veröffentlicht: (2025)
One Token Embedding Is Enough to Deadlock Your Large Reasoning Model
von: Zhang, Mohan, et al.
Veröffentlicht: (2025)
von: Zhang, Mohan, et al.
Veröffentlicht: (2025)
LLM Unlearning Under the Microscope: A Full-Stack View on Methods and Metrics
von: Fan, Chongyu, et al.
Veröffentlicht: (2025)
von: Fan, Chongyu, et al.
Veröffentlicht: (2025)
From Explicit CoT to Implicit CoT: Learning to Internalize CoT Step by Step
von: Deng, Yuntian, et al.
Veröffentlicht: (2024)
von: Deng, Yuntian, et al.
Veröffentlicht: (2024)
Towards a Re-evaluation of Data Forging Attacks in Practice
von: Suliman, Mohamed, et al.
Veröffentlicht: (2024)
von: Suliman, Mohamed, et al.
Veröffentlicht: (2024)
Distribution-Aligned Sequence Distillation for Superior Long-CoT Reasoning
von: Yan, Shaotian, et al.
Veröffentlicht: (2026)
von: Yan, Shaotian, et al.
Veröffentlicht: (2026)
Mol-R1: Towards Explicit Long-CoT Reasoning in Molecule Discovery
von: Li, Jiatong, et al.
Veröffentlicht: (2025)
von: Li, Jiatong, et al.
Veröffentlicht: (2025)
VG-CoT: Towards Trustworthy Visual Reasoning via Grounded Chain-of-Thought
von: Lim, Byeonggeuk, et al.
Veröffentlicht: (2026)
von: Lim, Byeonggeuk, et al.
Veröffentlicht: (2026)
UnlearnCanvas: Stylized Image Dataset for Enhanced Machine Unlearning Evaluation in Diffusion Models
von: Zhang, Yihua, et al.
Veröffentlicht: (2024)
von: Zhang, Yihua, et al.
Veröffentlicht: (2024)
Exploring the Limitations of Mamba in COPY and CoT Reasoning
von: Ren, Ruifeng, et al.
Veröffentlicht: (2024)
von: Ren, Ruifeng, et al.
Veröffentlicht: (2024)
Evaluating LLM Reasoning Beyond Correctness and CoT
von: Abbasloo, Soheil
Veröffentlicht: (2025)
von: Abbasloo, Soheil
Veröffentlicht: (2025)
Investigating CoT Monitorability in Large Reasoning Models
von: Yang, Shu, et al.
Veröffentlicht: (2025)
von: Yang, Shu, et al.
Veröffentlicht: (2025)
Improve Student's Reasoning Generalizability through Cascading Decomposed CoTs Distillation
von: Dai, Chengwei, et al.
Veröffentlicht: (2024)
von: Dai, Chengwei, et al.
Veröffentlicht: (2024)
Navigating Complexity: Toward Lossless Graph Condensation via Expanding Window Matching
von: Zhang, Yuchen, et al.
Veröffentlicht: (2024)
von: Zhang, Yuchen, et al.
Veröffentlicht: (2024)
Reconsidering Overthinking: Penalizing Internal and External Redundancy in CoT Reasoning
von: Hong, Jialiang, et al.
Veröffentlicht: (2025)
von: Hong, Jialiang, et al.
Veröffentlicht: (2025)
D-SCoRE: Document-Centric Segmentation and CoT Reasoning with Structured Export for QA-CoT Data Generation
von: Zhou, Weibo, et al.
Veröffentlicht: (2025)
von: Zhou, Weibo, et al.
Veröffentlicht: (2025)
CondenseGraph: Communication-Efficient Distributed GNN Training via On-the-Fly Graph Condensation
von: Zhang, Zizhao, et al.
Veröffentlicht: (2026)
von: Zhang, Zizhao, et al.
Veröffentlicht: (2026)
Distilling Long-CoT Reasoning through Collaborative Step-wise Multi-Teacher Decoding
von: Yun, Taewon, et al.
Veröffentlicht: (2026)
von: Yun, Taewon, et al.
Veröffentlicht: (2026)
Powering Up Zeroth-Order Training via Subspace Gradient Orthogonalization
von: Lang, Yicheng, et al.
Veröffentlicht: (2026)
von: Lang, Yicheng, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Reasoning Model Unlearning: Forgetting Traces, Not Just Answers, While Preserving Reasoning Skills
von: Wang, Changsheng, et al.
Veröffentlicht: (2025) -
Beyond SFT: Reinforcement Learning for Safer Large Reasoning Models with Better Reasoning Ability
von: Jia, Jinghan, et al.
Veröffentlicht: (2025) -
Towards LLM Unlearning Resilient to Relearning Attacks: A Sharpness-Aware Minimization Perspective and Beyond
von: Fan, Chongyu, et al.
Veröffentlicht: (2025) -
EPiC: Efficient Video Camera Control Learning with Precise Anchor-Video Guidance
von: Wang, Zun, et al.
Veröffentlicht: (2025) -
Leak@$k$: Unlearning Does Not Make LLMs Forget Under Probabilistic Decoding
von: Reisizadeh, Hadi, et al.
Veröffentlicht: (2025)