MetaGDPO: Alleviating Catastrophic Forgetting with Metacognitive Knowledge through Group Direct Preference Optimization
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhang, Lanxue, Xie, Yuqiang, Fang, Fang, Dong, Fanglong, Liu, Rui, Cao, Yanan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
GDPO: Group reward-Decoupled Normalization Policy Optimization for Multi-reward RL Optimization
von: Liu, Shih-Yang, et al.
Veröffentlicht: (2026)
von: Liu, Shih-Yang, et al.
Veröffentlicht: (2026)
Revisiting Catastrophic Forgetting in Large Language Model Tuning
von: Li, Hongyu, et al.
Veröffentlicht: (2024)
von: Li, Hongyu, et al.
Veröffentlicht: (2024)
Revisiting Catastrophic Forgetting in Continual Knowledge Graph Embedding
von: Pons, Gerard, et al.
Veröffentlicht: (2026)
von: Pons, Gerard, et al.
Veröffentlicht: (2026)
GDPO: Learning to Directly Align Language Models with Diversity Using GFlowNets
von: Kwon, Oh Joon, et al.
Veröffentlicht: (2024)
von: Kwon, Oh Joon, et al.
Veröffentlicht: (2024)
How to Alleviate Catastrophic Forgetting in LLMs Finetuning? Hierarchical Layer-Wise and Element-Wise Regularization
von: Song, Shezheng, et al.
Veröffentlicht: (2025)
von: Song, Shezheng, et al.
Veröffentlicht: (2025)
Can LLMs Alleviate Catastrophic Forgetting in Graph Continual Learning? A Systematic Study
von: Cheng, Ziyang, et al.
Veröffentlicht: (2025)
von: Cheng, Ziyang, et al.
Veröffentlicht: (2025)
Keep the General, Inject the Specific: Structured Dialogue Fine-Tuning for Knowledge Injection without Catastrophic Forgetting
von: Hong, Yijie, et al.
Veröffentlicht: (2025)
von: Hong, Yijie, et al.
Veröffentlicht: (2025)
Catastrophic Forgetting in Kolmogorov-Arnold Networks
von: Rahman, Mohammad Marufur, et al.
Veröffentlicht: (2025)
von: Rahman, Mohammad Marufur, et al.
Veröffentlicht: (2025)
CORE: Mitigating Catastrophic Forgetting in Continual Learning through Cognitive Replay
von: Zhang, Jianshu, et al.
Veröffentlicht: (2024)
von: Zhang, Jianshu, et al.
Veröffentlicht: (2024)
Sequencing to Mitigate Catastrophic Forgetting in Continual Learning
von: Moussa, Hesham G., et al.
Veröffentlicht: (2025)
von: Moussa, Hesham G., et al.
Veröffentlicht: (2025)
Stable Preference Optimization: A Bilevel Approach to Catastrophic Preference Shift
von: Jian, Chengtao, et al.
Veröffentlicht: (2025)
von: Jian, Chengtao, et al.
Veröffentlicht: (2025)
Addressing Loss of Plasticity and Catastrophic Forgetting in Continual Learning
von: Elsayed, Mohamed, et al.
Veröffentlicht: (2024)
von: Elsayed, Mohamed, et al.
Veröffentlicht: (2024)
Quantifying Catastrophic Forgetting in IoT Intrusion Detection Systems
von: Banerjee, Sourasekhar, et al.
Veröffentlicht: (2026)
von: Banerjee, Sourasekhar, et al.
Veröffentlicht: (2026)
A Conformal Predictive Measure for Assessing Catastrophic Forgetting
von: Pitsiorlas, Ioannis, et al.
Veröffentlicht: (2025)
von: Pitsiorlas, Ioannis, et al.
Veröffentlicht: (2025)
On the Implicit Adversariality of Catastrophic Forgetting in Deep Continual Learning
von: Peng, Ze, et al.
Veröffentlicht: (2025)
von: Peng, Ze, et al.
Veröffentlicht: (2025)
Continual Learning and Catastrophic Forgetting
von: van de Ven, Gido M., et al.
Veröffentlicht: (2024)
von: van de Ven, Gido M., et al.
Veröffentlicht: (2024)
Meta-R1: Empowering Large Reasoning Models with Metacognition
von: Dong, Haonan, et al.
Veröffentlicht: (2025)
von: Dong, Haonan, et al.
Veröffentlicht: (2025)
Evolutionary Strategies lead to Catastrophic Forgetting in LLMs
von: Abdi, Immanuel, et al.
Veröffentlicht: (2026)
von: Abdi, Immanuel, et al.
Veröffentlicht: (2026)
CoT2-Meta: Budgeted Metacognitive Control for Test-Time Reasoning
von: Ma, Siyuan, et al.
Veröffentlicht: (2026)
von: Ma, Siyuan, et al.
Veröffentlicht: (2026)
Explaining Robustness to Catastrophic Forgetting Through Incremental Concept Formation
von: Barari, Nicki, et al.
Veröffentlicht: (2025)
von: Barari, Nicki, et al.
Veröffentlicht: (2025)
Catastrophic Forgetting Mitigation Through Plateau Phase Activity Profiling
von: Mashiach, Idan, et al.
Veröffentlicht: (2025)
von: Mashiach, Idan, et al.
Veröffentlicht: (2025)
Model Editing at Scale leads to Gradual and Catastrophic Forgetting
von: Gupta, Akshat, et al.
Veröffentlicht: (2024)
von: Gupta, Akshat, et al.
Veröffentlicht: (2024)
Analyzing and Reducing Catastrophic Forgetting in Parameter Efficient Tuning
von: Ren, Weijieying, et al.
Veröffentlicht: (2024)
von: Ren, Weijieying, et al.
Veröffentlicht: (2024)
MAD-Fact: A Multi-Agent Debate Framework for Long-Form Factuality Evaluation in LLMs
von: Ning, Yucheng, et al.
Veröffentlicht: (2025)
von: Ning, Yucheng, et al.
Veröffentlicht: (2025)
MoE-CT: A Novel Approach For Large Language Models Training With Resistance To Catastrophic Forgetting
von: Li, Tianhao, et al.
Veröffentlicht: (2024)
von: Li, Tianhao, et al.
Veröffentlicht: (2024)
Negative Preference Optimization: From Catastrophic Collapse to Effective Unlearning
von: Zhang, Ruiqi, et al.
Veröffentlicht: (2024)
von: Zhang, Ruiqi, et al.
Veröffentlicht: (2024)
Token-Importance Guided Direct Preference Optimization
von: Yang, Ning, et al.
Veröffentlicht: (2025)
von: Yang, Ning, et al.
Veröffentlicht: (2025)
Delayed Bottlenecking: Alleviating Forgetting in Pre-trained Graph Neural Networks
von: Zhao, Zhe, et al.
Veröffentlicht: (2024)
von: Zhao, Zhe, et al.
Veröffentlicht: (2024)
Mitigating Catastrophic Forgetting in Large Language Models with Self-Synthesized Rehearsal
von: Huang, Jianheng, et al.
Veröffentlicht: (2024)
von: Huang, Jianheng, et al.
Veröffentlicht: (2024)
Overcoming Catastrophic Forgetting by Exemplar Selection in Task-oriented Dialogue System
von: Chen, Chen, et al.
Veröffentlicht: (2024)
von: Chen, Chen, et al.
Veröffentlicht: (2024)
Autoregressive Direct Preference Optimization
von: Oi, Masanari, et al.
Veröffentlicht: (2026)
von: Oi, Masanari, et al.
Veröffentlicht: (2026)
Do LLMs Know Tool Irrelevance? Demystifying Structural Alignment Bias in Tool Invocations
von: Liu, Yilong, et al.
Veröffentlicht: (2026)
von: Liu, Yilong, et al.
Veröffentlicht: (2026)
Adaptive Collaboration with Humans: Metacognitive Policy Optimization for Multi-Agent LLMs with Continual Learning
von: Yang, Wei, et al.
Veröffentlicht: (2026)
von: Yang, Wei, et al.
Veröffentlicht: (2026)
Intelligent Learning Rate Distribution to reduce Catastrophic Forgetting in Transformers
von: Kenneweg, Philip, et al.
Veröffentlicht: (2024)
von: Kenneweg, Philip, et al.
Veröffentlicht: (2024)
VAGPO: Vision-augmented Asymmetric Group Preference Optimization for Graph Routing Problems
von: Liu, Shiyan, et al.
Veröffentlicht: (2025)
von: Liu, Shiyan, et al.
Veröffentlicht: (2025)
Direct Preference Optimization with Unobserved Preference Heterogeneity: The Necessity of Ternary Preferences
von: Chidambaram, Keertana, et al.
Veröffentlicht: (2025)
von: Chidambaram, Keertana, et al.
Veröffentlicht: (2025)
Gradient Correlation Subspace Learning against Catastrophic Forgetting
von: Dubnov, Tammuz, et al.
Veröffentlicht: (2024)
von: Dubnov, Tammuz, et al.
Veröffentlicht: (2024)
DeMansia: Mamba Never Forgets Any Tokens
von: Fang, Ricky
Veröffentlicht: (2024)
von: Fang, Ricky
Veröffentlicht: (2024)
PKG-DPO: Optimizing Domain-Specific AI systems with Physics Knowledge Graphs and Direct Preference Optimization
von: Kulkarni, Nitin Nagesh, et al.
Veröffentlicht: (2025)
von: Kulkarni, Nitin Nagesh, et al.
Veröffentlicht: (2025)
Knowledge Editing in Language Models via Adapted Direct Preference Optimization
von: Rozner, Amit, et al.
Veröffentlicht: (2024)
von: Rozner, Amit, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
GDPO: Group reward-Decoupled Normalization Policy Optimization for Multi-reward RL Optimization
von: Liu, Shih-Yang, et al.
Veröffentlicht: (2026) -
Revisiting Catastrophic Forgetting in Large Language Model Tuning
von: Li, Hongyu, et al.
Veröffentlicht: (2024) -
Revisiting Catastrophic Forgetting in Continual Knowledge Graph Embedding
von: Pons, Gerard, et al.
Veröffentlicht: (2026) -
GDPO: Learning to Directly Align Language Models with Diversity Using GFlowNets
von: Kwon, Oh Joon, et al.
Veröffentlicht: (2024) -
How to Alleviate Catastrophic Forgetting in LLMs Finetuning? Hierarchical Layer-Wise and Element-Wise Regularization
von: Song, Shezheng, et al.
Veröffentlicht: (2025)