On Double Descent in Reinforcement Learning with LSTD and Random Features
Fuente:
arXiv
Saved in:
| Main Authors: | Brellmann, David, Berthier, Eloïse, Filliat, David, Frehse, Goran |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
On The Presence of Double-Descent in Deep Reinforcement Learning
by: Veselý, Viktor, et al.
Published: (2025)
by: Veselý, Viktor, et al.
Published: (2025)
Double Descent Meets Out-of-Distribution Detection: Theoretical Insights and Empirical Analysis on the role of model complexity
by: Ammar, Mouïn Ben, et al.
Published: (2024)
by: Ammar, Mouïn Ben, et al.
Published: (2024)
Concept-Based Mechanistic Interpretability Using Structured Knowledge Graphs
by: Chorna, Sofiia, et al.
Published: (2025)
by: Chorna, Sofiia, et al.
Published: (2025)
Adaptive Reinforcement Learning for Unobservable Random Delays
by: Wikman, John, et al.
Published: (2025)
by: Wikman, John, et al.
Published: (2025)
Benchmarking XAI Explanations with Human-Aligned Evaluations
by: Kazmierczak, Rémi, et al.
Published: (2024)
by: Kazmierczak, Rémi, et al.
Published: (2024)
Two Speeds of Learning: A Representation-Readout Decomposition of Grokking and Double Descent
by: Chou, Chi-Ning, et al.
Published: (2026)
by: Chou, Chi-Ning, et al.
Published: (2026)
Randomness and Interpolation Improve Gradient Descent
by: Li, Jiawen, et al.
Published: (2025)
by: Li, Jiawen, et al.
Published: (2025)
Reinforcement Learning for Durable Algorithmic Recourse
by: Ceccon, Marina, et al.
Published: (2025)
by: Ceccon, Marina, et al.
Published: (2025)
Corruption Robust Offline Reinforcement Learning with Human Feedback
by: Mandal, Debmalya, et al.
Published: (2024)
by: Mandal, Debmalya, et al.
Published: (2024)
Distributionally Robust Reinforcement Learning with Human Feedback
by: Mandal, Debmalya, et al.
Published: (2025)
by: Mandal, Debmalya, et al.
Published: (2025)
Opinion-Guided Reinforcement Learning
by: Dagenais, Kyanna, et al.
Published: (2024)
by: Dagenais, Kyanna, et al.
Published: (2024)
Double Successive Over-Relaxation Q-Learning with an Extension to Deep Reinforcement Learning
by: R, Shreyas S
Published: (2024)
by: R, Shreyas S
Published: (2024)
Double Actor-Critic with TD Error-Driven Regularization in Reinforcement Learning
by: Chen, Haohui, et al.
Published: (2024)
by: Chen, Haohui, et al.
Published: (2024)
Natural Gradient Descent for Online Continual Learning
by: Khawand, Joe, et al.
Published: (2026)
by: Khawand, Joe, et al.
Published: (2026)
Memory Allocation in Resource-Constrained Reinforcement Learning
by: Tamborski, Massimiliano, et al.
Published: (2025)
by: Tamborski, Massimiliano, et al.
Published: (2025)
Three Dogmas of Reinforcement Learning
by: Abel, David, et al.
Published: (2024)
by: Abel, David, et al.
Published: (2024)
Learning Associative Memories with Gradient Descent
by: Cabannes, Vivien, et al.
Published: (2024)
by: Cabannes, Vivien, et al.
Published: (2024)
Local Entropy Search over Descent Sequences for Bayesian Optimization
by: Stenger, David, et al.
Published: (2025)
by: Stenger, David, et al.
Published: (2025)
On the Effective Horizon of Inverse Reinforcement Learning
by: Xu, Yiqing, et al.
Published: (2023)
by: Xu, Yiqing, et al.
Published: (2023)
Exploiting Estimation Bias in Clipped Double Q-Learning for Continous Control Reinforcement Learning Tasks
by: Turcato, Niccolò, et al.
Published: (2024)
by: Turcato, Niccolò, et al.
Published: (2024)
Cross-domain Random Pre-training with Prototypes for Reinforcement Learning
by: Liu, Xin, et al.
Published: (2023)
by: Liu, Xin, et al.
Published: (2023)
Model-Based Reinforcement Learning under Random Observation Delays
by: Karamzade, Armin, et al.
Published: (2025)
by: Karamzade, Armin, et al.
Published: (2025)
Trajectory Modeling via Random Utility Inverse Reinforcement Learning
by: Pitombeira-Neto, Anselmo R., et al.
Published: (2021)
by: Pitombeira-Neto, Anselmo R., et al.
Published: (2021)
Reinforcement Learning via Conservative Agent for Environments with Random Delays
by: Lee, Jongsoo, et al.
Published: (2025)
by: Lee, Jongsoo, et al.
Published: (2025)
Rethinking the Foundations for Continual Reinforcement Learning
by: Elelimy, Esraa, et al.
Published: (2025)
by: Elelimy, Esraa, et al.
Published: (2025)
Parseval Regularization for Continual Reinforcement Learning
by: Chung, Wesley, et al.
Published: (2024)
by: Chung, Wesley, et al.
Published: (2024)
Reinforcement Learning-based Feature Generation Algorithm for Scientific Data
by: Xiao, Meng, et al.
Published: (2025)
by: Xiao, Meng, et al.
Published: (2025)
Stacked Universal Successor Feature Approximators for Safety in Reinforcement Learning
by: Cannon, Ian, et al.
Published: (2024)
by: Cannon, Ian, et al.
Published: (2024)
Evaluating Feature Dependent Noise in Preference-based Reinforcement Learning
by: Li, Yuxuan, et al.
Published: (2026)
by: Li, Yuxuan, et al.
Published: (2026)
Learning with Density Matrices and Random Features
by: González, Fabio A., et al.
Published: (2021)
by: González, Fabio A., et al.
Published: (2021)
Gradient Free Deep Reinforcement Learning With TabPFN
by: Schiff, David, et al.
Published: (2025)
by: Schiff, David, et al.
Published: (2025)
Regret-Based Defense in Adversarial Reinforcement Learning
by: Belaire, Roman, et al.
Published: (2023)
by: Belaire, Roman, et al.
Published: (2023)
Random Policy Enables In-Context Reinforcement Learning within Trust Horizons
by: Chen, Weiqin, et al.
Published: (2024)
by: Chen, Weiqin, et al.
Published: (2024)
More Efficient Randomized Exploration for Reinforcement Learning via Approximate Sampling
by: Ishfaq, Haque, et al.
Published: (2024)
by: Ishfaq, Haque, et al.
Published: (2024)
Elastic Multi-Gradient Descent for Parallel Continual Learning
by: Lyu, Fan, et al.
Published: (2024)
by: Lyu, Fan, et al.
Published: (2024)
The Initialization Determines Whether In-Context Learning Is Gradient Descent
by: Xie, Shifeng, et al.
Published: (2025)
by: Xie, Shifeng, et al.
Published: (2025)
Conflict-Averse Gradient Descent for Multi-task Learning
by: Liu, Bo, et al.
Published: (2021)
by: Liu, Bo, et al.
Published: (2021)
All Random Features Representations are Equivalent
by: Sernau, Luke, et al.
Published: (2024)
by: Sernau, Luke, et al.
Published: (2024)
Beyond-Expert Performance with Limited Demonstrations: Efficient Imitation Learning with Double Exploration
by: Zhao, Heyang, et al.
Published: (2025)
by: Zhao, Heyang, et al.
Published: (2025)
Combining Pre-Trained Models for Enhanced Feature Representation in Reinforcement Learning
by: Piccoli, Elia, et al.
Published: (2025)
by: Piccoli, Elia, et al.
Published: (2025)
Similar Items
-
On The Presence of Double-Descent in Deep Reinforcement Learning
by: Veselý, Viktor, et al.
Published: (2025) -
Double Descent Meets Out-of-Distribution Detection: Theoretical Insights and Empirical Analysis on the role of model complexity
by: Ammar, Mouïn Ben, et al.
Published: (2024) -
Concept-Based Mechanistic Interpretability Using Structured Knowledge Graphs
by: Chorna, Sofiia, et al.
Published: (2025) -
Adaptive Reinforcement Learning for Unobservable Random Delays
by: Wikman, John, et al.
Published: (2025) -
Benchmarking XAI Explanations with Human-Aligned Evaluations
by: Kazmierczak, Rémi, et al.
Published: (2024)