Robust Policy Optimization to Prevent Catastrophic Forgetting
Fuente:
arXiv
Saved in:
| Main Authors: | Sabbaghi, Mahdi, Pappas, George, Javanmard, Adel, Hassani, Hamed |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
InfoSFT: Learn More and Forget Less with Information-Aware Token Weighting
by: Sabbaghi, Mahdi, et al.
Published: (2026)
by: Sabbaghi, Mahdi, et al.
Published: (2026)
Explicitly Encoding Structural Symmetry is Key to Length Generalization in Arithmetic Tasks
by: Sabbaghi, Mahdi, et al.
Published: (2024)
by: Sabbaghi, Mahdi, et al.
Published: (2024)
The curse of overparametrization in adversarial training: Precise analysis of robust generalization for random features regression
by: Hassani, Hamed, et al.
Published: (2022)
by: Hassani, Hamed, et al.
Published: (2022)
Adversarial Reasoning at Jailbreaking Time
by: Sabbaghi, Mahdi, et al.
Published: (2025)
by: Sabbaghi, Mahdi, et al.
Published: (2025)
Length Optimization in Conformal Prediction
by: Kiyani, Shayan, et al.
Published: (2024)
by: Kiyani, Shayan, et al.
Published: (2024)
Robust Decision Making with Partially Calibrated Forecasts
by: Kiyani, Shayan, et al.
Published: (2025)
by: Kiyani, Shayan, et al.
Published: (2025)
Conformal Prediction with Learned Features
by: Kiyani, Shayan, et al.
Published: (2024)
by: Kiyani, Shayan, et al.
Published: (2024)
Multi-Round Human-AI Collaboration with User-Specified Requirements
by: Noorani, Sima, et al.
Published: (2026)
by: Noorani, Sima, et al.
Published: (2026)
When to Trust the Cheap Check: Weak and Strong Verification for Reasoning
by: Kiyani, Shayan, et al.
Published: (2026)
by: Kiyani, Shayan, et al.
Published: (2026)
Decision Theoretic Foundations for Conformal Prediction: Optimal Uncertainty Quantification for Risk-Averse Agents
by: Kiyani, Shayan, et al.
Published: (2025)
by: Kiyani, Shayan, et al.
Published: (2025)
Conformal Prediction Beyond the Seen: A Missing Mass Perspective for Uncertainty Quantification in Generative Models
by: Noorani, Sima, et al.
Published: (2025)
by: Noorani, Sima, et al.
Published: (2025)
Temporal Difference Learning with Compressed Updates: Error-Feedback meets Reinforcement Learning
by: Mitra, Aritra, et al.
Published: (2023)
by: Mitra, Aritra, et al.
Published: (2023)
SmoothLLM: Defending Large Language Models Against Jailbreaking Attacks
by: Robey, Alexander, et al.
Published: (2023)
by: Robey, Alexander, et al.
Published: (2023)
Human-AI Collaborative Uncertainty Quantification
by: Noorani, Sima, et al.
Published: (2025)
by: Noorani, Sima, et al.
Published: (2025)
Differentially Private Model-X Knockoffs via Johnson-Lindenstrauss Transform
by: Tao, Yuxuan, et al.
Published: (2025)
by: Tao, Yuxuan, et al.
Published: (2025)
Conformal Inference under High-Dimensional Covariate Shifts via Likelihood-Ratio Regularization
by: Joshi, Sunay, et al.
Published: (2025)
by: Joshi, Sunay, et al.
Published: (2025)
Robust Feature Learning for Multi-Index Models in High Dimensions
by: Mousavi-Hosseini, Alireza, et al.
Published: (2024)
by: Mousavi-Hosseini, Alireza, et al.
Published: (2024)
Preventing Catastrophic Forgetting through Memory Networks in Continuous Detection
by: Bhatt, Gaurav, et al.
Published: (2024)
by: Bhatt, Gaurav, et al.
Published: (2024)
Conformal Risk Minimization with Variance Reduction
by: Noorani, Sima, et al.
Published: (2024)
by: Noorani, Sima, et al.
Published: (2024)
Adversarial Training Should Be Cast as a Non-Zero-Sum Game
by: Robey, Alexander, et al.
Published: (2023)
by: Robey, Alexander, et al.
Published: (2023)
Federated Temporal Difference Learning with Linear Function Approximation under Environmental Heterogeneity
by: Wang, Han, et al.
Published: (2023)
by: Wang, Han, et al.
Published: (2023)
Continual Deep Reinforcement Learning to Prevent Catastrophic Forgetting in Jamming Mitigation
by: Davaslioglu, Kemal, et al.
Published: (2024)
by: Davaslioglu, Kemal, et al.
Published: (2024)
Explaining Robustness to Catastrophic Forgetting Through Incremental Concept Formation
by: Barari, Nicki, et al.
Published: (2025)
by: Barari, Nicki, et al.
Published: (2025)
Predicting the Susceptibility of Examples to Catastrophic Forgetting
by: Hacohen, Guy, et al.
Published: (2024)
by: Hacohen, Guy, et al.
Published: (2024)
Chordal Sparsity for Lipschitz Constant Estimation of Deep Neural Networks
by: Xue, Anton, et al.
Published: (2022)
by: Xue, Anton, et al.
Published: (2022)
Mitigating Catastrophic Forgetting in Large Language Models with Forgetting-aware Pruning
by: Huang, Wei, et al.
Published: (2025)
by: Huang, Wei, et al.
Published: (2025)
Theoretical Perspectives on Data Quality and Synergistic Effects in Pre- and Post-Training Reasoning Models
by: Javanmard, Adel, et al.
Published: (2026)
by: Javanmard, Adel, et al.
Published: (2026)
Understanding the Role of Training Data in Test-Time Scaling
by: Javanmard, Adel, et al.
Published: (2025)
by: Javanmard, Adel, et al.
Published: (2025)
Multi-Task Dynamic Pricing in Credit Market with Contextual Information
by: Javanmard, Adel, et al.
Published: (2024)
by: Javanmard, Adel, et al.
Published: (2024)
Catastrophic Forgetting in Kolmogorov-Arnold Networks
by: Rahman, Mohammad Marufur, et al.
Published: (2025)
by: Rahman, Mohammad Marufur, et al.
Published: (2025)
Jailbreaking Black Box Large Language Models in Twenty Queries
by: Chao, Patrick, et al.
Published: (2023)
by: Chao, Patrick, et al.
Published: (2023)
PriorBoost: An Adaptive Algorithm for Learning from Aggregate Responses
by: Javanmard, Adel, et al.
Published: (2024)
by: Javanmard, Adel, et al.
Published: (2024)
Eidetic Learning: an Efficient and Provable Solution to Catastrophic Forgetting
by: Dronen, Nicholas, et al.
Published: (2025)
by: Dronen, Nicholas, et al.
Published: (2025)
Are Time Series Foundation Models Susceptible to Catastrophic Forgetting?
by: Karaouli, Nouha, et al.
Published: (2025)
by: Karaouli, Nouha, et al.
Published: (2025)
Sharpness-Aware Pretraining Mitigates Catastrophic Forgetting
by: Watts, Ishaan, et al.
Published: (2026)
by: Watts, Ishaan, et al.
Published: (2026)
Sequencing to Mitigate Catastrophic Forgetting in Continual Learning
by: Moussa, Hesham G., et al.
Published: (2025)
by: Moussa, Hesham G., et al.
Published: (2025)
Feature Learning in Linear-Width Two-Layer Networks: Two vs. One Step of Gradient Descent
by: Moniri, Behrad, et al.
Published: (2026)
by: Moniri, Behrad, et al.
Published: (2026)
On the Mechanisms of Weak-to-Strong Generalization: A Theoretical Perspective
by: Moniri, Behrad, et al.
Published: (2025)
by: Moniri, Behrad, et al.
Published: (2025)
Provable Multi-Task Representation Learning by Two-Layer ReLU Neural Networks
by: Collins, Liam, et al.
Published: (2023)
by: Collins, Liam, et al.
Published: (2023)
Mitigating Catastrophic Forgetting in Language Transfer via Model Merging
by: Alexandrov, Anton, et al.
Published: (2024)
by: Alexandrov, Anton, et al.
Published: (2024)
Similar Items
-
InfoSFT: Learn More and Forget Less with Information-Aware Token Weighting
by: Sabbaghi, Mahdi, et al.
Published: (2026) -
Explicitly Encoding Structural Symmetry is Key to Length Generalization in Arithmetic Tasks
by: Sabbaghi, Mahdi, et al.
Published: (2024) -
The curse of overparametrization in adversarial training: Precise analysis of robust generalization for random features regression
by: Hassani, Hamed, et al.
Published: (2022) -
Adversarial Reasoning at Jailbreaking Time
by: Sabbaghi, Mahdi, et al.
Published: (2025) -
Length Optimization in Conformal Prediction
by: Kiyani, Shayan, et al.
Published: (2024)