Rethinking the Foundations for Continual Reinforcement Learning
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Elelimy, Esraa, Szepesvari, David, White, Martha, Bowling, Michael |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Real-Time Recurrent Learning using Trace Units in Reinforcement Learning
von: Elelimy, Esraa, et al.
Veröffentlicht: (2024)
von: Elelimy, Esraa, et al.
Veröffentlicht: (2024)
Deep Reinforcement Learning with Gradient Eligibility Traces
von: Elelimy, Esraa, et al.
Veröffentlicht: (2025)
von: Elelimy, Esraa, et al.
Veröffentlicht: (2025)
AGaLiTe: Approximate Gated Linear Transformers for Online Reinforcement Learning
von: Pramanik, Subhojeet, et al.
Veröffentlicht: (2023)
von: Pramanik, Subhojeet, et al.
Veröffentlicht: (2023)
Investigating the Histogram Loss in Regression
von: Imani, Ehsan, et al.
Veröffentlicht: (2024)
von: Imani, Ehsan, et al.
Veröffentlicht: (2024)
A Method for Evaluating Hyperparameter Sensitivity in Reinforcement Learning
von: Adkins, Jacob, et al.
Veröffentlicht: (2024)
von: Adkins, Jacob, et al.
Veröffentlicht: (2024)
Forager: a lightweight testbed for continual learning with partial observability in RL
von: Tang, Steven, et al.
Veröffentlicht: (2026)
von: Tang, Steven, et al.
Veröffentlicht: (2026)
On the Interplay Between Sparsity and Training in Deep Reinforcement Learning
von: Davelouis, Fatima, et al.
Veröffentlicht: (2025)
von: Davelouis, Fatima, et al.
Veröffentlicht: (2025)
Empirical Design in Reinforcement Learning
von: Patterson, Andrew, et al.
Veröffentlicht: (2023)
von: Patterson, Andrew, et al.
Veröffentlicht: (2023)
A Generalized Projected Bellman Error for Off-policy Value Estimation in Reinforcement Learning
von: Patterson, Andrew, et al.
Veröffentlicht: (2021)
von: Patterson, Andrew, et al.
Veröffentlicht: (2021)
Mitigating Value Hallucination in Dyna Planning via Multistep Predecessor Models
von: Aminmansour, Farzane, et al.
Veröffentlicht: (2020)
von: Aminmansour, Farzane, et al.
Veröffentlicht: (2020)
Value Bonuses using Ensemble Errors for Exploration in Reinforcement Learning
von: Wahab, Abdul, et al.
Veröffentlicht: (2026)
von: Wahab, Abdul, et al.
Veröffentlicht: (2026)
Proper Laplacian Representation Learning
von: Gomez, Diego, et al.
Veröffentlicht: (2023)
von: Gomez, Diego, et al.
Veröffentlicht: (2023)
Harnessing Discrete Representations For Continual Reinforcement Learning
von: Meyer, Edan, et al.
Veröffentlicht: (2023)
von: Meyer, Edan, et al.
Veröffentlicht: (2023)
A New View on Planning in Online Reinforcement Learning
von: Roice, Kevin, et al.
Veröffentlicht: (2024)
von: Roice, Kevin, et al.
Veröffentlicht: (2024)
Beyond Statistical Learning: Exact Learning Is Essential for General Intelligence
von: György, András, et al.
Veröffentlicht: (2025)
von: György, András, et al.
Veröffentlicht: (2025)
Generalized Munchausen Reinforcement Learning using Tsallis KL Divergence
von: Zhu, Lingwei, et al.
Veröffentlicht: (2023)
von: Zhu, Lingwei, et al.
Veröffentlicht: (2023)
When is Offline Policy Selection Sample Efficient for Reinforcement Learning?
von: Liu, Vincent, et al.
Veröffentlicht: (2023)
von: Liu, Vincent, et al.
Veröffentlicht: (2023)
Learning to Be Cautious
von: Mohammedalamen, Montaser, et al.
Veröffentlicht: (2021)
von: Mohammedalamen, Montaser, et al.
Veröffentlicht: (2021)
Balancing optimism and pessimism in offline-to-online learning
von: Sentenac, Flore, et al.
Veröffentlicht: (2025)
von: Sentenac, Flore, et al.
Veröffentlicht: (2025)
Rethinking Plasticity in Deep Reinforcement Learning
von: He, Zhiqiang
Veröffentlicht: (2026)
von: He, Zhiqiang
Veröffentlicht: (2026)
Ergodic Risk Measures: Towards a Risk-Aware Foundation for Continual Reinforcement Learning
von: Rojas, Juan Sebastian, et al.
Veröffentlicht: (2025)
von: Rojas, Juan Sebastian, et al.
Veröffentlicht: (2025)
Parseval Regularization for Continual Reinforcement Learning
von: Chung, Wesley, et al.
Veröffentlicht: (2024)
von: Chung, Wesley, et al.
Veröffentlicht: (2024)
Rethinking Momentum Knowledge Distillation in Online Continual Learning
von: Michel, Nicolas, et al.
Veröffentlicht: (2023)
von: Michel, Nicolas, et al.
Veröffentlicht: (2023)
Agnostic Reinforcement Learning: Foundations and Algorithms
von: Li, Gene
Veröffentlicht: (2025)
von: Li, Gene
Veröffentlicht: (2025)
Meta-Gradient Search Control: A Method for Improving the Efficiency of Dyna-style Planning
von: Burega, Bradley, et al.
Veröffentlicht: (2024)
von: Burega, Bradley, et al.
Veröffentlicht: (2024)
Knowledge Retention for Continual Model-Based Reinforcement Learning
von: Sun, Yixiang, et al.
Veröffentlicht: (2025)
von: Sun, Yixiang, et al.
Veröffentlicht: (2025)
Cog-Rethinker: Hierarchical Metacognitive Reinforcement Learning for LLM Reasoning
von: Sun, Zexu, et al.
Veröffentlicht: (2025)
von: Sun, Zexu, et al.
Veröffentlicht: (2025)
Fine-Tuning without Performance Degradation
von: Wang, Han, et al.
Veröffentlicht: (2025)
von: Wang, Han, et al.
Veröffentlicht: (2025)
Demystifying the Recency Heuristic in Temporal-Difference Learning
von: Daley, Brett, et al.
Veröffentlicht: (2024)
von: Daley, Brett, et al.
Veröffentlicht: (2024)
Rethinking Adversarial Attacks in Reinforcement Learning from Policy Distribution Perspective
von: Duan, Tianyang, et al.
Veröffentlicht: (2025)
von: Duan, Tianyang, et al.
Veröffentlicht: (2025)
Rethinking Inverse Reinforcement Learning: from Data Alignment to Task Alignment
von: Zhou, Weichao, et al.
Veröffentlicht: (2024)
von: Zhou, Weichao, et al.
Veröffentlicht: (2024)
Topological Foundations of Reinforcement Learning
von: Kadurha, David Krame
Veröffentlicht: (2024)
von: Kadurha, David Krame
Veröffentlicht: (2024)
Continual Learning as Computationally Constrained Reinforcement Learning
von: Kumar, Saurabh, et al.
Veröffentlicht: (2023)
von: Kumar, Saurabh, et al.
Veröffentlicht: (2023)
Diffusion Models for Reinforcement Learning: Foundations, Taxonomy, and Development
von: Xu, Changfu, et al.
Veröffentlicht: (2025)
von: Xu, Changfu, et al.
Veröffentlicht: (2025)
A Survey of Continual Reinforcement Learning
von: Pan, Chaofan, et al.
Veröffentlicht: (2025)
von: Pan, Chaofan, et al.
Veröffentlicht: (2025)
Parameter Importance-Driven Continual Learning for Foundation Models
von: Wang, Lingxiang, et al.
Veröffentlicht: (2025)
von: Wang, Lingxiang, et al.
Veröffentlicht: (2025)
VL-Rethinker: Incentivizing Self-Reflection of Vision-Language Models with Reinforcement Learning
von: Wang, Haozhe, et al.
Veröffentlicht: (2025)
von: Wang, Haozhe, et al.
Veröffentlicht: (2025)
Rethinking Out-of-Distribution Detection for Reinforcement Learning: Advancing Methods for Evaluation and Detection
von: Nasvytis, Linas, et al.
Veröffentlicht: (2024)
von: Nasvytis, Linas, et al.
Veröffentlicht: (2024)
Counteractive RL: Rethinking Core Principles for Efficient and Scalable Deep Reinforcement Learning
von: Korkmaz, Ezgi
Veröffentlicht: (2026)
von: Korkmaz, Ezgi
Veröffentlicht: (2026)
Operator-Guided Invariance Learning for Continuous Reinforcement Learning
von: Zhang, Zuyuan, et al.
Veröffentlicht: (2026)
von: Zhang, Zuyuan, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Real-Time Recurrent Learning using Trace Units in Reinforcement Learning
von: Elelimy, Esraa, et al.
Veröffentlicht: (2024) -
Deep Reinforcement Learning with Gradient Eligibility Traces
von: Elelimy, Esraa, et al.
Veröffentlicht: (2025) -
AGaLiTe: Approximate Gated Linear Transformers for Online Reinforcement Learning
von: Pramanik, Subhojeet, et al.
Veröffentlicht: (2023) -
Investigating the Histogram Loss in Regression
von: Imani, Ehsan, et al.
Veröffentlicht: (2024) -
A Method for Evaluating Hyperparameter Sensitivity in Reinforcement Learning
von: Adkins, Jacob, et al.
Veröffentlicht: (2024)