Sharing Knowledge in Multi-Task Deep Reinforcement Learning
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | D'Eramo, Carlo, Tateo, Davide, Bonarini, Andrea, Restelli, Marcello, Peters, Jan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Multi-Task Reinforcement Learning with Mixture of Orthogonal Experts
von: Hendawy, Ahmed, et al.
Veröffentlicht: (2023)
von: Hendawy, Ahmed, et al.
Veröffentlicht: (2023)
$K$-Level Policy Gradients for Multi-Agent Reinforcement Learning
von: Reddi, Aryaman, et al.
Veröffentlicht: (2025)
von: Reddi, Aryaman, et al.
Veröffentlicht: (2025)
On the Benefit of Optimal Transport for Curriculum Reinforcement Learning
von: Klink, Pascal, et al.
Veröffentlicht: (2023)
von: Klink, Pascal, et al.
Veröffentlicht: (2023)
Streaming Reinforcement Learning under Partial Observability with Real-Time Recurrent Learning
von: Farr, Noah, et al.
Veröffentlicht: (2026)
von: Farr, Noah, et al.
Veröffentlicht: (2026)
Adaptive $Q$-Network: On-the-fly Target Selection for Deep Reinforcement Learning
von: Vincent, Théo, et al.
Veröffentlicht: (2024)
von: Vincent, Théo, et al.
Veröffentlicht: (2024)
Parameterized Projected Bellman Operator
von: Vincent, Théo, et al.
Veröffentlicht: (2023)
von: Vincent, Théo, et al.
Veröffentlicht: (2023)
Iterated $Q$-Network: Beyond One-Step Bellman Updates in Deep Reinforcement Learning
von: Vincent, Théo, et al.
Veröffentlicht: (2024)
von: Vincent, Théo, et al.
Veröffentlicht: (2024)
Eau De $Q$-Network: Adaptive Distillation of Neural Networks in Deep Reinforcement Learning
von: Vincent, Théo, et al.
Veröffentlicht: (2025)
von: Vincent, Théo, et al.
Veröffentlicht: (2025)
Learning to Explore in Diverse Reward Settings via Temporal-Difference-Error Maximization
von: Griesbach, Sebastian, et al.
Veröffentlicht: (2025)
von: Griesbach, Sebastian, et al.
Veröffentlicht: (2025)
Use the Online Network If You Can: Towards Fast and Stable Reinforcement Learning
von: Hendawy, Ahmed, et al.
Veröffentlicht: (2025)
von: Hendawy, Ahmed, et al.
Veröffentlicht: (2025)
Deterministic Exploration via Stationary Bellman Error Maximization
von: Griesbach, Sebastian, et al.
Veröffentlicht: (2024)
von: Griesbach, Sebastian, et al.
Veröffentlicht: (2024)
Dynamic Obstacle Avoidance with Bounded Rationality Adversarial Reinforcement Learning
von: Holgado-Alvarez, Jose-Luis, et al.
Veröffentlicht: (2025)
von: Holgado-Alvarez, Jose-Luis, et al.
Veröffentlicht: (2025)
Do Not Imitate, Reinforce: Iterative Classification via Belief Refinement
von: Kallel, Mahdi, et al.
Veröffentlicht: (2026)
von: Kallel, Mahdi, et al.
Veröffentlicht: (2026)
Handling Long-Term Safety and Uncertainty in Safe Reinforcement Learning
von: Günster, Jonas, et al.
Veröffentlicht: (2024)
von: Günster, Jonas, et al.
Veröffentlicht: (2024)
Bridging the Performance Gap Between Target-Free and Target-Based Reinforcement Learning
von: Vincent, Théo, et al.
Veröffentlicht: (2025)
von: Vincent, Théo, et al.
Veröffentlicht: (2025)
Domain Randomization via Entropy Maximization
von: Tiboni, Gabriele, et al.
Veröffentlicht: (2023)
von: Tiboni, Gabriele, et al.
Veröffentlicht: (2023)
Safe Reinforcement Learning on the Constraint Manifold: Theory and Applications
von: Liu, Puze, et al.
Veröffentlicht: (2024)
von: Liu, Puze, et al.
Veröffentlicht: (2024)
Learning in Markov Decision Processes with Exogenous Dynamics
von: Maran, Davide, et al.
Veröffentlicht: (2026)
von: Maran, Davide, et al.
Veröffentlicht: (2026)
Augmented Bayesian Policy Search
von: Kallel, Mahdi, et al.
Veröffentlicht: (2024)
von: Kallel, Mahdi, et al.
Veröffentlicht: (2024)
Scalable Multi-Agent Offline Reinforcement Learning and the Role of Information
von: Zamboni, Riccardo, et al.
Veröffentlicht: (2025)
von: Zamboni, Riccardo, et al.
Veröffentlicht: (2025)
Finite Sample Bounds for Non-Parametric Regression: Optimal Sample Efficiency and Space Complexity
von: Maran, Davide, et al.
Veröffentlicht: (2024)
von: Maran, Davide, et al.
Veröffentlicht: (2024)
Online Market Making and the Value of Observing the Order Book
von: Maran, Davide, et al.
Veröffentlicht: (2026)
von: Maran, Davide, et al.
Veröffentlicht: (2026)
Towards Principled Unsupervised Multi-Agent Reinforcement Learning
von: Zamboni, Riccardo, et al.
Veröffentlicht: (2025)
von: Zamboni, Riccardo, et al.
Veröffentlicht: (2025)
"So, Tell Me About Your Policy...": Distillation of interpretable policies from Deep Reinforcement Learning agents
von: Dispoto, Giovanni, et al.
Veröffentlicht: (2025)
von: Dispoto, Giovanni, et al.
Veröffentlicht: (2025)
A Reinforcement Learning Approach for Optimal Control in Microgrids
von: Salaorni, Davide, et al.
Veröffentlicht: (2025)
von: Salaorni, Davide, et al.
Veröffentlicht: (2025)
Gradient Iterated Temporal-Difference Learning
von: Vincent, Théo, et al.
Veröffentlicht: (2026)
von: Vincent, Théo, et al.
Veröffentlicht: (2026)
Local Linearity: the Key for No-regret Reinforcement Learning in Continuous MDPs
von: Maran, Davide, et al.
Veröffentlicht: (2024)
von: Maran, Davide, et al.
Veröffentlicht: (2024)
Interpetable Target-Feature Aggregation for Multi-Task Learning based on Bias-Variance Analysis
von: Bonetti, Paolo, et al.
Veröffentlicht: (2024)
von: Bonetti, Paolo, et al.
Veröffentlicht: (2024)
Bridging the gap between Learning-to-plan, Motion Primitives and Safe Reinforcement Learning
von: Kicki, Piotr, et al.
Veröffentlicht: (2024)
von: Kicki, Piotr, et al.
Veröffentlicht: (2024)
Projection by Convolution: Optimal Sample Complexity for Reinforcement Learning in Continuous-Space MDPs
von: Maran, Davide, et al.
Veröffentlicht: (2024)
von: Maran, Davide, et al.
Veröffentlicht: (2024)
Power Grid Control with Graph-Based Distributed Reinforcement Learning
von: Fabrizio, Carlo, et al.
Veröffentlicht: (2025)
von: Fabrizio, Carlo, et al.
Veröffentlicht: (2025)
Building surrogate models using trajectories of agents trained by Reinforcement Learning
von: Cestero, Julen, et al.
Veröffentlicht: (2025)
von: Cestero, Julen, et al.
Veröffentlicht: (2025)
Unsupervised Behavioral Compression: Learning Low-Dimensional Policy Manifolds through State-Occupancy Matching
von: Fraschini, Andrea, et al.
Veröffentlicht: (2026)
von: Fraschini, Andrea, et al.
Veröffentlicht: (2026)
Trust Region Inverse Reinforcement Learning: Explicit Dual Ascent using Local Policy Updates
von: Diwan, Anish, et al.
Veröffentlicht: (2026)
von: Diwan, Anish, et al.
Veröffentlicht: (2026)
One Policy to Run Them All: an End-to-end Learning Approach to Multi-Embodiment Locomotion
von: Bohlinger, Nico, et al.
Veröffentlicht: (2024)
von: Bohlinger, Nico, et al.
Veröffentlicht: (2024)
Online Dynamic Pricing of Complementary Products
von: Mussi, Marco, et al.
Veröffentlicht: (2025)
von: Mussi, Marco, et al.
Veröffentlicht: (2025)
Inverse Reinforcement Learning with Sub-optimal Experts
von: Poiani, Riccardo, et al.
Veröffentlicht: (2024)
von: Poiani, Riccardo, et al.
Veröffentlicht: (2024)
Exciting Action: Investigating Efficient Exploration for Learning Musculoskeletal Humanoid Locomotion
von: Geiß, Henri-Jacques, et al.
Veröffentlicht: (2024)
von: Geiß, Henri-Jacques, et al.
Veröffentlicht: (2024)
Machine Learning with Physics Knowledge for Prediction: A Survey
von: Watson, Joe, et al.
Veröffentlicht: (2024)
von: Watson, Joe, et al.
Veröffentlicht: (2024)
Exploiting Causal Graph Priors with Posterior Sampling for Reinforcement Learning
von: Mutti, Mirco, et al.
Veröffentlicht: (2023)
von: Mutti, Mirco, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
Multi-Task Reinforcement Learning with Mixture of Orthogonal Experts
von: Hendawy, Ahmed, et al.
Veröffentlicht: (2023) -
$K$-Level Policy Gradients for Multi-Agent Reinforcement Learning
von: Reddi, Aryaman, et al.
Veröffentlicht: (2025) -
On the Benefit of Optimal Transport for Curriculum Reinforcement Learning
von: Klink, Pascal, et al.
Veröffentlicht: (2023) -
Streaming Reinforcement Learning under Partial Observability with Real-Time Recurrent Learning
von: Farr, Noah, et al.
Veröffentlicht: (2026) -
Adaptive $Q$-Network: On-the-fly Target Selection for Deep Reinforcement Learning
von: Vincent, Théo, et al.
Veröffentlicht: (2024)