A Retention-Centric Framework for Continual Learning with Guaranteed Model Developmental Safety
Fuente:
arXiv
Guardado en:
| Autores principales: | Li, Gang, Yu, Wendi, Yao, Yao, Tong, Wei, Liang, Yingbin, Lin, Qihang, Yang, Tianbao |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Agentic Transformers Provably Learn to Search via Reinforcement Learning
por: Yang, Tong, et al.
Publicado: (2026)
por: Yang, Tong, et al.
Publicado: (2026)
A Note on Complexity for Two Classes of Structured Non-Smooth Non-Convex Compositional Optimization
por: Yao, Yao, et al.
Publicado: (2024)
por: Yao, Yao, et al.
Publicado: (2024)
Deterministic and Stochastic Accelerated Gradient Method for Convex Semi-Infinite Optimization
por: Yao, Yao, et al.
Publicado: (2023)
por: Yao, Yao, et al.
Publicado: (2023)
Multi-head Transformers Provably Learn Symbolic Multi-step Reasoning via Gradient Descent
por: Yang, Tong, et al.
Publicado: (2025)
por: Yang, Tong, et al.
Publicado: (2025)
The Implicit Curriculum: Learning Dynamics in RL with Verifiable Rewards
por: Huang, Yu, et al.
Publicado: (2026)
por: Huang, Yu, et al.
Publicado: (2026)
Distributionally Robust Control with End-to-End Statistically Guaranteed Metric Learning
por: Wu, Jingyi, et al.
Publicado: (2025)
por: Wu, Jingyi, et al.
Publicado: (2025)
Non-Smooth Weakly-Convex Finite-sum Coupled Compositional Optimization
por: Hu, Quanqi, et al.
Publicado: (2023)
por: Hu, Quanqi, et al.
Publicado: (2023)
Score as Action: Fine-Tuning Diffusion Generative Models by Continuous-time Reinforcement Learning
por: Zhao, Hanyang, et al.
Publicado: (2025)
por: Zhao, Hanyang, et al.
Publicado: (2025)
To Cool or not to Cool? Temperature Network Meets Large Foundation Models via DRO
por: Qiu, Zi-Hao, et al.
Publicado: (2024)
por: Qiu, Zi-Hao, et al.
Publicado: (2024)
Using Laplace Transform To Optimize the Hallucination of Generation Models
por: Kang, Cheng, et al.
Publicado: (2026)
por: Kang, Cheng, et al.
Publicado: (2026)
Quantization through Piecewise-Affine Regularization: Optimization and Statistical Guarantees
por: Ma, Jianhao, et al.
Publicado: (2025)
por: Ma, Jianhao, et al.
Publicado: (2025)
Mixed variable structural optimization using mixed variable system Monte Carlo tree search formulation
por: Ko, Fu-Yao, et al.
Publicado: (2023)
por: Ko, Fu-Yao, et al.
Publicado: (2023)
Safety-Aware Reinforcement Learning for Electric Vehicle Charging Station Management in Distribution Network
por: Fan, Jiarong, et al.
Publicado: (2024)
por: Fan, Jiarong, et al.
Publicado: (2024)
Model Predictive Control and Reinforcement Learning: A Unified Framework Based on Dynamic Programming
por: Bertsekas, Dimitri P.
Publicado: (2024)
por: Bertsekas, Dimitri P.
Publicado: (2024)
R2L: Reliable Reinforcement Learning: Guaranteed Return & Reliable Policies in Reinforcement Learning
por: Farhi, Nadir
Publicado: (2025)
por: Farhi, Nadir
Publicado: (2025)
FMIP: Joint Continuous-Integer Flow For Mixed-Integer Linear Programming
por: Li, Hongpei, et al.
Publicado: (2025)
por: Li, Hongpei, et al.
Publicado: (2025)
T-SKM-Net: Trainable Neural Network Framework for Linear Constraint Satisfaction via Sampling Kaczmarz-Motzkin Method
por: Zhu, Haoyu, et al.
Publicado: (2025)
por: Zhu, Haoyu, et al.
Publicado: (2025)
Optimization and Generalization Guarantees for Weight Normalization
por: Cisneros-Velarde, Pedro, et al.
Publicado: (2024)
por: Cisneros-Velarde, Pedro, et al.
Publicado: (2024)
Learning to Cut: Reinforcement Learning for Benders Decomposition
por: Cai, Haochen, et al.
Publicado: (2026)
por: Cai, Haochen, et al.
Publicado: (2026)
Client-Centric Federated Adaptive Optimization
por: Sun, Jianhui, et al.
Publicado: (2025)
por: Sun, Jianhui, et al.
Publicado: (2025)
A Clustering-Based Variable Ordering Framework for Relaxed Decision Diagrams for Maximum Weighted Independent Set Problem
por: Nafar, Mohsen, et al.
Publicado: (2025)
por: Nafar, Mohsen, et al.
Publicado: (2025)
WANCO: Weak Adversarial Networks for Constrained Optimization problems
por: Bao, Gang, et al.
Publicado: (2024)
por: Bao, Gang, et al.
Publicado: (2024)
ResearchEVO: An End-to-End Framework for Automated Scientific Discovery and Documentation
por: Zhao, Zhe, et al.
Publicado: (2026)
por: Zhao, Zhe, et al.
Publicado: (2026)
Sublinear Regret for a Class of Continuous-Time Linear-Quadratic Reinforcement Learning Problems
por: Huang, Yilie, et al.
Publicado: (2024)
por: Huang, Yilie, et al.
Publicado: (2024)
Feasible Pairings for Decentralized Integral Controllability of Non-Square Systems
por: Tong, Yuhao, et al.
Publicado: (2026)
por: Tong, Yuhao, et al.
Publicado: (2026)
Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems
por: Huang, Yilie, et al.
Publicado: (2025)
por: Huang, Yilie, et al.
Publicado: (2025)
Omni-scale Learning-based Sequential Decision Framework for Order Fulfillment of Tote-handling Robotic Systems
por: Liu, Jiaxin, et al.
Publicado: (2026)
por: Liu, Jiaxin, et al.
Publicado: (2026)
Global Convergence Guarantees for Federated Policy Gradient Methods with Adversaries
por: Ganesh, Swetha, et al.
Publicado: (2024)
por: Ganesh, Swetha, et al.
Publicado: (2024)
Stochastic Subgradient Methods with Guaranteed Global Stability in Nonsmooth Nonconvex Optimization
por: Xiao, Nachuan, et al.
Publicado: (2023)
por: Xiao, Nachuan, et al.
Publicado: (2023)
Federated Dynamical Low-Rank Training with Global Loss Convergence Guarantees
por: Schotthöfer, Steffen, et al.
Publicado: (2024)
por: Schotthöfer, Steffen, et al.
Publicado: (2024)
Wasserstein gradient flow for optimal probability measure decomposition
por: Han, Jiangze, et al.
Publicado: (2024)
por: Han, Jiangze, et al.
Publicado: (2024)
StoSignSGD: Unbiased Structural Stochasticity Fixes SignSGD for Training Large Language Models
por: Yu, Dingzhi, et al.
Publicado: (2026)
por: Yu, Dingzhi, et al.
Publicado: (2026)
Understanding Sampler Stochasticity in Training Diffusion Models for RLHF
por: Sheng, Jiayuan, et al.
Publicado: (2025)
por: Sheng, Jiayuan, et al.
Publicado: (2025)
Deterministic Policy Gradient Primal-Dual Methods for Continuous-Space Constrained MDPs
por: Rozada, Sergio, et al.
Publicado: (2024)
por: Rozada, Sergio, et al.
Publicado: (2024)
Boosting Gradient Ascent for Continuous DR-submodular Maximization
por: Zhang, Qixin, et al.
Publicado: (2024)
por: Zhang, Qixin, et al.
Publicado: (2024)
Energy Management for Renewable-Colocated Artificial Intelligence Data Centers
por: Li, Siying, et al.
Publicado: (2025)
por: Li, Siying, et al.
Publicado: (2025)
Deep Reinforcement Learning for Flexible Job Shop Scheduling with Random Job Arrivals
por: Tang, Yu, et al.
Publicado: (2026)
por: Tang, Yu, et al.
Publicado: (2026)
PySCIPOpt-ML: Embedding Trained Machine Learning Models into Mixed-Integer Programs
por: Turner, Mark, et al.
Publicado: (2023)
por: Turner, Mark, et al.
Publicado: (2023)
GBO:AMulti-Granularity Optimization Algorithm via Granular-ball for Continuous Problems
por: Xia, Shuyin, et al.
Publicado: (2023)
por: Xia, Shuyin, et al.
Publicado: (2023)
Multi-CALF: A Policy Combination Approach with Statistical Guarantees
por: Malaniya, Georgiy, et al.
Publicado: (2025)
por: Malaniya, Georgiy, et al.
Publicado: (2025)
Ejemplares similares
-
Agentic Transformers Provably Learn to Search via Reinforcement Learning
por: Yang, Tong, et al.
Publicado: (2026) -
A Note on Complexity for Two Classes of Structured Non-Smooth Non-Convex Compositional Optimization
por: Yao, Yao, et al.
Publicado: (2024) -
Deterministic and Stochastic Accelerated Gradient Method for Convex Semi-Infinite Optimization
por: Yao, Yao, et al.
Publicado: (2023) -
Multi-head Transformers Provably Learn Symbolic Multi-step Reasoning via Gradient Descent
por: Yang, Tong, et al.
Publicado: (2025) -
The Implicit Curriculum: Learning Dynamics in RL with Verifiable Rewards
por: Huang, Yu, et al.
Publicado: (2026)