Bridging the Performance Gap Between Target-Free and Target-Based Reinforcement Learning
Fuente:
arXiv
Saved in:
| Main Authors: | Vincent, Théo, Tripathi, Yogesh, Faust, Tim, Akgül, Abdullah, Oren, Yaniv, Kandemir, Melih, Peters, Jan, D'Eramo, Carlo |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Eau De $Q$-Network: Adaptive Distillation of Neural Networks in Deep Reinforcement Learning
by: Vincent, Théo, et al.
Published: (2025)
by: Vincent, Théo, et al.
Published: (2025)
Adaptive $Q$-Network: On-the-fly Target Selection for Deep Reinforcement Learning
by: Vincent, Théo, et al.
Published: (2024)
by: Vincent, Théo, et al.
Published: (2024)
Deterministic Uncertainty Propagation for Improved Model-Based Offline Reinforcement Learning
by: Akgül, Abdullah, et al.
Published: (2024)
by: Akgül, Abdullah, et al.
Published: (2024)
Iterated $Q$-Network: Beyond One-Step Bellman Updates in Deep Reinforcement Learning
by: Vincent, Théo, et al.
Published: (2024)
by: Vincent, Théo, et al.
Published: (2024)
Multi-Task Reinforcement Learning with Mixture of Orthogonal Experts
by: Hendawy, Ahmed, et al.
Published: (2023)
by: Hendawy, Ahmed, et al.
Published: (2023)
Gradient Iterated Temporal-Difference Learning
by: Vincent, Théo, et al.
Published: (2026)
by: Vincent, Théo, et al.
Published: (2026)
Use the Online Network If You Can: Towards Fast and Stable Reinforcement Learning
by: Hendawy, Ahmed, et al.
Published: (2025)
by: Hendawy, Ahmed, et al.
Published: (2025)
On the Benefit of Optimal Transport for Curriculum Reinforcement Learning
by: Klink, Pascal, et al.
Published: (2023)
by: Klink, Pascal, et al.
Published: (2023)
Continual Learning of Multi-modal Dynamics with External Memory
by: Akgül, Abdullah, et al.
Published: (2022)
by: Akgül, Abdullah, et al.
Published: (2022)
$K$-Level Policy Gradients for Multi-Agent Reinforcement Learning
by: Reddi, Aryaman, et al.
Published: (2025)
by: Reddi, Aryaman, et al.
Published: (2025)
Streaming Reinforcement Learning under Partial Observability with Real-Time Recurrent Learning
by: Farr, Noah, et al.
Published: (2026)
by: Farr, Noah, et al.
Published: (2026)
Sharing Knowledge in Multi-Task Deep Reinforcement Learning
by: D'Eramo, Carlo, et al.
Published: (2024)
by: D'Eramo, Carlo, et al.
Published: (2024)
Parameterized Projected Bellman Operator
by: Vincent, Théo, et al.
Published: (2023)
by: Vincent, Théo, et al.
Published: (2023)
Overcoming Non-stationary Dynamics with Evidential Proximal Policy Optimization
by: Akgül, Abdullah, et al.
Published: (2025)
by: Akgül, Abdullah, et al.
Published: (2025)
Learning to Explore in Diverse Reward Settings via Temporal-Difference-Error Maximization
by: Griesbach, Sebastian, et al.
Published: (2025)
by: Griesbach, Sebastian, et al.
Published: (2025)
Deterministic Exploration via Stationary Bellman Error Maximization
by: Griesbach, Sebastian, et al.
Published: (2024)
by: Griesbach, Sebastian, et al.
Published: (2024)
Weighted Sequential Bayesian Inference for Non-Stationary Linear Contextual Bandits
by: Werge, Nicklas, et al.
Published: (2023)
by: Werge, Nicklas, et al.
Published: (2023)
Light to Heavy, Brief to Eternal: An Axion for Every Occasion (in the Early Universe)
by: D'Eramo, Francesco
Published: (2026)
by: D'Eramo, Francesco
Published: (2026)
Hitting Time Isomorphism for Multi-Stage Planning with Foundation Policies
by: Boock, Magnus Victor, et al.
Published: (2026)
by: Boock, Magnus Victor, et al.
Published: (2026)
Calibrating Bayesian UNet++ for Sub-Seasonal Forecasting
by: Asan, Busra, et al.
Published: (2024)
by: Asan, Busra, et al.
Published: (2024)
Dynamic Obstacle Avoidance with Bounded Rationality Adversarial Reinforcement Learning
by: Holgado-Alvarez, Jose-Luis, et al.
Published: (2025)
by: Holgado-Alvarez, Jose-Luis, et al.
Published: (2025)
Do Not Imitate, Reinforce: Iterative Classification via Belief Refinement
by: Kallel, Mahdi, et al.
Published: (2026)
by: Kallel, Mahdi, et al.
Published: (2026)
Distributional Active Inference
by: Akgül, Abdullah, et al.
Published: (2026)
by: Akgül, Abdullah, et al.
Published: (2026)
PAC-Bayesian Soft Actor-Critic Learning
by: Tasdighi, Bahareh, et al.
Published: (2023)
by: Tasdighi, Bahareh, et al.
Published: (2023)
ObjectRL: An Object-Oriented Reinforcement Learning Codebase
by: Baykal, Gulcin, et al.
Published: (2025)
by: Baykal, Gulcin, et al.
Published: (2025)
Domain Randomization via Entropy Maximization
by: Tiboni, Gabriele, et al.
Published: (2023)
by: Tiboni, Gabriele, et al.
Published: (2023)
Axion Portal to Scalar Dark Matter: Unveiling Stabilizing Symmetry Footprints
by: D'Eramo, Francesco, et al.
Published: (2025)
by: D'Eramo, Francesco, et al.
Published: (2025)
Back to the phase space: thermal axion dark radiation via couplings to standard model fermions
by: D'Eramo, Francesco, et al.
Published: (2024)
by: D'Eramo, Francesco, et al.
Published: (2024)
Probing Non-Minimal Dark Sectors via the 21 cm Line at Cosmic Dawn
by: Cima, Federico, et al.
Published: (2025)
by: Cima, Federico, et al.
Published: (2025)
Improved Algorithms for Stochastic Linear Bandits Using Tail Bounds for Martingale Mixtures
by: Flynn, Hamish, et al.
Published: (2023)
by: Flynn, Hamish, et al.
Published: (2023)
Contact Energy Based Hindsight Experience Prioritization
by: Sayar, Erdi, et al.
Published: (2023)
by: Sayar, Erdi, et al.
Published: (2023)
Augmented Bayesian Policy Search
by: Kallel, Mahdi, et al.
Published: (2024)
by: Kallel, Mahdi, et al.
Published: (2024)
Bridging the Gap Between Target Networks and Functional Regularization
by: Piche, Alexandre, et al.
Published: (2022)
by: Piche, Alexandre, et al.
Published: (2022)
Dark Matter Freeze-In and Small-Scale Observables: Novel Mass Bounds and Viable Particle Candidates
by: D'Eramo, Francesco, et al.
Published: (2025)
by: D'Eramo, Francesco, et al.
Published: (2025)
Ultralight Dark Matter from the Edge of Field Space
by: Becker, Mathias, et al.
Published: (2025)
by: Becker, Mathias, et al.
Published: (2025)
Seasons of Dark Matter Freeze-In Shaped by the Weather of the Early Universe
by: D'Eramo, Francesco, et al.
Published: (2025)
by: D'Eramo, Francesco, et al.
Published: (2025)
Cosmic-Ray Signatures of Annihilating and Semi-Annihilating Dark Matter via One-Step Cascades
by: D'Eramo, Francesco, et al.
Published: (2026)
by: D'Eramo, Francesco, et al.
Published: (2026)
Irreducible cosmological backgrounds of a real scalar with a broken symmetry
by: D'Eramo, Francesco, et al.
Published: (2024)
by: D'Eramo, Francesco, et al.
Published: (2024)
Dark Radiation from the Primordial Thermal Bath in Momentum Space
by: D'Eramo, Francesco, et al.
Published: (2023)
by: D'Eramo, Francesco, et al.
Published: (2023)
Bridging Domain Gaps with Target-Aligned Generation for Offline Reinforcement Learning
by: Kim, Minung, et al.
Published: (2026)
by: Kim, Minung, et al.
Published: (2026)
Similar Items
-
Eau De $Q$-Network: Adaptive Distillation of Neural Networks in Deep Reinforcement Learning
by: Vincent, Théo, et al.
Published: (2025) -
Adaptive $Q$-Network: On-the-fly Target Selection for Deep Reinforcement Learning
by: Vincent, Théo, et al.
Published: (2024) -
Deterministic Uncertainty Propagation for Improved Model-Based Offline Reinforcement Learning
by: Akgül, Abdullah, et al.
Published: (2024) -
Iterated $Q$-Network: Beyond One-Step Bellman Updates in Deep Reinforcement Learning
by: Vincent, Théo, et al.
Published: (2024) -
Multi-Task Reinforcement Learning with Mixture of Orthogonal Experts
by: Hendawy, Ahmed, et al.
Published: (2023)