Use the Online Network If You Can: Towards Fast and Stable Reinforcement Learning
Fuente:
arXiv
Saved in:
| Main Authors: | Hendawy, Ahmed, Metternich, Henrik, Vincent, Théo, Kallel, Mahdi, Peters, Jan, D'Eramo, Carlo |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Multi-Task Reinforcement Learning with Mixture of Orthogonal Experts
by: Hendawy, Ahmed, et al.
Published: (2023)
by: Hendawy, Ahmed, et al.
Published: (2023)
Do Not Imitate, Reinforce: Iterative Classification via Belief Refinement
by: Kallel, Mahdi, et al.
Published: (2026)
by: Kallel, Mahdi, et al.
Published: (2026)
Eau De $Q$-Network: Adaptive Distillation of Neural Networks in Deep Reinforcement Learning
by: Vincent, Théo, et al.
Published: (2025)
by: Vincent, Théo, et al.
Published: (2025)
Adaptive $Q$-Network: On-the-fly Target Selection for Deep Reinforcement Learning
by: Vincent, Théo, et al.
Published: (2024)
by: Vincent, Théo, et al.
Published: (2024)
Iterated $Q$-Network: Beyond One-Step Bellman Updates in Deep Reinforcement Learning
by: Vincent, Théo, et al.
Published: (2024)
by: Vincent, Théo, et al.
Published: (2024)
Augmented Bayesian Policy Search
by: Kallel, Mahdi, et al.
Published: (2024)
by: Kallel, Mahdi, et al.
Published: (2024)
On the Benefit of Optimal Transport for Curriculum Reinforcement Learning
by: Klink, Pascal, et al.
Published: (2023)
by: Klink, Pascal, et al.
Published: (2023)
Streaming Reinforcement Learning under Partial Observability with Real-Time Recurrent Learning
by: Farr, Noah, et al.
Published: (2026)
by: Farr, Noah, et al.
Published: (2026)
$K$-Level Policy Gradients for Multi-Agent Reinforcement Learning
by: Reddi, Aryaman, et al.
Published: (2025)
by: Reddi, Aryaman, et al.
Published: (2025)
Sharing Knowledge in Multi-Task Deep Reinforcement Learning
by: D'Eramo, Carlo, et al.
Published: (2024)
by: D'Eramo, Carlo, et al.
Published: (2024)
Bridging the Performance Gap Between Target-Free and Target-Based Reinforcement Learning
by: Vincent, Théo, et al.
Published: (2025)
by: Vincent, Théo, et al.
Published: (2025)
Learning to Explore in Diverse Reward Settings via Temporal-Difference-Error Maximization
by: Griesbach, Sebastian, et al.
Published: (2025)
by: Griesbach, Sebastian, et al.
Published: (2025)
Parameterized Projected Bellman Operator
by: Vincent, Théo, et al.
Published: (2023)
by: Vincent, Théo, et al.
Published: (2023)
Deterministic Exploration via Stationary Bellman Error Maximization
by: Griesbach, Sebastian, et al.
Published: (2024)
by: Griesbach, Sebastian, et al.
Published: (2024)
Gradient Iterated Temporal-Difference Learning
by: Vincent, Théo, et al.
Published: (2026)
by: Vincent, Théo, et al.
Published: (2026)
Dynamic Obstacle Avoidance with Bounded Rationality Adversarial Reinforcement Learning
by: Holgado-Alvarez, Jose-Luis, et al.
Published: (2025)
by: Holgado-Alvarez, Jose-Luis, et al.
Published: (2025)
Machine Learning with Physics Knowledge for Prediction: A Survey
by: Watson, Joe, et al.
Published: (2024)
by: Watson, Joe, et al.
Published: (2024)
Domain Randomization via Entropy Maximization
by: Tiboni, Gabriele, et al.
Published: (2023)
by: Tiboni, Gabriele, et al.
Published: (2023)
Light to Heavy, Brief to Eternal: An Axion for Every Occasion (in the Early Universe)
by: D'Eramo, Francesco
Published: (2026)
by: D'Eramo, Francesco
Published: (2026)
Axion Portal to Scalar Dark Matter: Unveiling Stabilizing Symmetry Footprints
by: D'Eramo, Francesco, et al.
Published: (2025)
by: D'Eramo, Francesco, et al.
Published: (2025)
Back to the phase space: thermal axion dark radiation via couplings to standard model fermions
by: D'Eramo, Francesco, et al.
Published: (2024)
by: D'Eramo, Francesco, et al.
Published: (2024)
Probing Non-Minimal Dark Sectors via the 21 cm Line at Cosmic Dawn
by: Cima, Federico, et al.
Published: (2025)
by: Cima, Federico, et al.
Published: (2025)
Dark Matter Freeze-In and Small-Scale Observables: Novel Mass Bounds and Viable Particle Candidates
by: D'Eramo, Francesco, et al.
Published: (2025)
by: D'Eramo, Francesco, et al.
Published: (2025)
Ultralight Dark Matter from the Edge of Field Space
by: Becker, Mathias, et al.
Published: (2025)
by: Becker, Mathias, et al.
Published: (2025)
Seasons of Dark Matter Freeze-In Shaped by the Weather of the Early Universe
by: D'Eramo, Francesco, et al.
Published: (2025)
by: D'Eramo, Francesco, et al.
Published: (2025)
Cosmic-Ray Signatures of Annihilating and Semi-Annihilating Dark Matter via One-Step Cascades
by: D'Eramo, Francesco, et al.
Published: (2026)
by: D'Eramo, Francesco, et al.
Published: (2026)
Irreducible cosmological backgrounds of a real scalar with a broken symmetry
by: D'Eramo, Francesco, et al.
Published: (2024)
by: D'Eramo, Francesco, et al.
Published: (2024)
Dark Radiation from the Primordial Thermal Bath in Momentum Space
by: D'Eramo, Francesco, et al.
Published: (2023)
by: D'Eramo, Francesco, et al.
Published: (2023)
Contact Energy Based Hindsight Experience Prioritization
by: Sayar, Erdi, et al.
Published: (2023)
by: Sayar, Erdi, et al.
Published: (2023)
Deep Reinforcement Learning Agents are not even close to Human Intelligence
by: Delfosse, Quentin, et al.
Published: (2025)
by: Delfosse, Quentin, et al.
Published: (2025)
Towards Fast Safe Online Reinforcement Learning via Policy Finetuning
by: Chen, Keru, et al.
Published: (2024)
by: Chen, Keru, et al.
Published: (2024)
FlashSAC: Fast and Stable Off-Policy Reinforcement Learning for High-Dimensional Robot Control
by: Kim, Donghu, et al.
Published: (2026)
by: Kim, Donghu, et al.
Published: (2026)
Lift What You Can: Green Online Learning with Heterogeneous Ensembles
by: Köbschall, Kirsten, et al.
Published: (2025)
by: Köbschall, Kirsten, et al.
Published: (2025)
Apple: Toward General Active Perception via Reinforcement Learning
by: Schneider, Tim, et al.
Published: (2025)
by: Schneider, Tim, et al.
Published: (2025)
Squint: Fast Visual Reinforcement Learning for Sim-to-Real Robotics
by: Almuzairee, Abdulaziz, et al.
Published: (2026)
by: Almuzairee, Abdulaziz, et al.
Published: (2026)
Discovering What You Can Control: Interventional Boundary Discovery for Reinforcement Learning
by: Liu, Jiaxin, et al.
Published: (2026)
by: Liu, Jiaxin, et al.
Published: (2026)
Stable Port-Hamiltonian Neural Networks
by: Roth, Fabian J., et al.
Published: (2025)
by: Roth, Fabian J., et al.
Published: (2025)
HyperController: A Hyperparameter Controller for Fast and Stable Training of Reinforcement Learning Neural Networks
by: Gornet, Jonathan, et al.
Published: (2025)
by: Gornet, Jonathan, et al.
Published: (2025)
Maximum Total Correlation Reinforcement Learning
by: You, Bang, et al.
Published: (2025)
by: You, Bang, et al.
Published: (2025)
Seer: Online Context Learning for Fast Synchronous LLM Reinforcement Learning
by: Qin, Ruoyu, et al.
Published: (2025)
by: Qin, Ruoyu, et al.
Published: (2025)
Similar Items
-
Multi-Task Reinforcement Learning with Mixture of Orthogonal Experts
by: Hendawy, Ahmed, et al.
Published: (2023) -
Do Not Imitate, Reinforce: Iterative Classification via Belief Refinement
by: Kallel, Mahdi, et al.
Published: (2026) -
Eau De $Q$-Network: Adaptive Distillation of Neural Networks in Deep Reinforcement Learning
by: Vincent, Théo, et al.
Published: (2025) -
Adaptive $Q$-Network: On-the-fly Target Selection for Deep Reinforcement Learning
by: Vincent, Théo, et al.
Published: (2024) -
Iterated $Q$-Network: Beyond One-Step Bellman Updates in Deep Reinforcement Learning
by: Vincent, Théo, et al.
Published: (2024)