Diverse Projection Ensembles for Distributional Reinforcement Learning
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zanger, Moritz A., Böhmer, Wendelin, Spaan, Matthijs T. J. |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2023
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
How Ensembles of Distilled Policies Improve Generalisation in Reinforcement Learning
von: Weltevrede, Max, et al.
Veröffentlicht: (2025)
von: Weltevrede, Max, et al.
Veröffentlicht: (2025)
Contextual Similarity Distillation: Ensemble Uncertainties with a Single Model
von: Zanger, Moritz A., et al.
Veröffentlicht: (2025)
von: Zanger, Moritz A., et al.
Veröffentlicht: (2025)
On the Equivalence of Random Network Distillation, Deep Ensembles, and Bayesian Inference
von: Zanger, Moritz A., et al.
Veröffentlicht: (2026)
von: Zanger, Moritz A., et al.
Veröffentlicht: (2026)
Explore-Go: Leveraging Exploration for Generalisation in Deep Reinforcement Learning
von: Weltevrede, Max, et al.
Veröffentlicht: (2024)
von: Weltevrede, Max, et al.
Veröffentlicht: (2024)
Epistemic Monte Carlo Tree Search
von: Oren, Yaniv, et al.
Veröffentlicht: (2022)
von: Oren, Yaniv, et al.
Veröffentlicht: (2022)
Exploration Implies Data Augmentation: Reachability and Generalisation in Contextual MDPs
von: Weltevrede, Max, et al.
Veröffentlicht: (2024)
von: Weltevrede, Max, et al.
Veröffentlicht: (2024)
Value Improved Actor Critic Algorithms
von: Oren, Yaniv, et al.
Veröffentlicht: (2024)
von: Oren, Yaniv, et al.
Veröffentlicht: (2024)
Universal Value-Function Uncertainties
von: Zanger, Moritz A., et al.
Veröffentlicht: (2025)
von: Zanger, Moritz A., et al.
Veröffentlicht: (2025)
Twice Sequential Monte Carlo for Tree Search
von: Oren, Yaniv, et al.
Veröffentlicht: (2025)
von: Oren, Yaniv, et al.
Veröffentlicht: (2025)
Improving Robustness of AlphaZero Algorithms to Test-Time Environment Changes
von: Tamassia, Isidoro, et al.
Veröffentlicht: (2025)
von: Tamassia, Isidoro, et al.
Veröffentlicht: (2025)
TransZero: Parallel Tree Expansion in MuZero using Transformer Networks
von: Malmsten, Emil, et al.
Veröffentlicht: (2025)
von: Malmsten, Emil, et al.
Veröffentlicht: (2025)
Reinforcement Learning by Guided Safe Exploration
von: Yang, Qisong, et al.
Veröffentlicht: (2023)
von: Yang, Qisong, et al.
Veröffentlicht: (2023)
Sparse Masked Attention Policies for Reliable Generalization
von: Horsch, Caroline, et al.
Veröffentlicht: (2026)
von: Horsch, Caroline, et al.
Veröffentlicht: (2026)
Positive Experience Reflection for Agents in Interactive Text Environments
von: Lippmann, Philip, et al.
Veröffentlicht: (2024)
von: Lippmann, Philip, et al.
Veröffentlicht: (2024)
Priors Matter: Addressing Misspecification in Bayesian Deep Q-Learning
von: van der Vaart, Pascal R., et al.
Veröffentlicht: (2025)
von: van der Vaart, Pascal R., et al.
Veröffentlicht: (2025)
Generalisation to unseen topologies: Towards control of biological neural network activity
von: Engwegen, Laurens, et al.
Veröffentlicht: (2024)
von: Engwegen, Laurens, et al.
Veröffentlicht: (2024)
EfficientTDMPC: Improved MPC Objectives for Sample-Efficient Continuous Control
von: Evers, Thomas, et al.
Veröffentlicht: (2026)
von: Evers, Thomas, et al.
Veröffentlicht: (2026)
PMCTS: Particle Monte Carlo Tree Search for Principled Parallelized Inference Time Scaling
von: Oren, Yaniv, et al.
Veröffentlicht: (2026)
von: Oren, Yaniv, et al.
Veröffentlicht: (2026)
RecBayes: Recurrent Bayesian Ad Hoc Teamwork in Large Partially Observable Domains
von: Ribeiro, João G., et al.
Veröffentlicht: (2025)
von: Ribeiro, João G., et al.
Veröffentlicht: (2025)
To the Max: Reinventing Reward in Reinforcement Learning
von: Veviurko, Grigorii, et al.
Veröffentlicht: (2024)
von: Veviurko, Grigorii, et al.
Veröffentlicht: (2024)
Rethinking Policy Diversity in Ensemble Policy Gradient in Large-Scale Reinforcement Learning
von: Shitanda, Naoki, et al.
Veröffentlicht: (2026)
von: Shitanda, Naoki, et al.
Veröffentlicht: (2026)
A Unified Theory of Diversity in Ensemble Learning
von: Wood, Danny, et al.
Veröffentlicht: (2023)
von: Wood, Danny, et al.
Veröffentlicht: (2023)
Pessimistic Iterative Planning with RNNs for Robust POMDPs
von: Galesloot, Maris F. L., et al.
Veröffentlicht: (2024)
von: Galesloot, Maris F. L., et al.
Veröffentlicht: (2024)
Federated Ensemble-Directed Offline Reinforcement Learning
von: Rengarajan, Desik, et al.
Veröffentlicht: (2023)
von: Rengarajan, Desik, et al.
Veröffentlicht: (2023)
RLAE: Reinforcement Learning-Assisted Ensemble for LLMs
von: Fu, Yuqian, et al.
Veröffentlicht: (2025)
von: Fu, Yuqian, et al.
Veröffentlicht: (2025)
Value Bonuses using Ensemble Errors for Exploration in Reinforcement Learning
von: Wahab, Abdul, et al.
Veröffentlicht: (2026)
von: Wahab, Abdul, et al.
Veröffentlicht: (2026)
Advancing Robustness in Deep Reinforcement Learning with an Ensemble Defense Approach
von: Mohan, Adithya, et al.
Veröffentlicht: (2025)
von: Mohan, Adithya, et al.
Veröffentlicht: (2025)
Reinforcement Learning Agent for a 2D Shooter Game
von: Ackermann, Thomas, et al.
Veröffentlicht: (2025)
von: Ackermann, Thomas, et al.
Veröffentlicht: (2025)
Meta-Ensemble Learning with Diverse Data Splits for Improved Respiratory Sound Classification
von: Kim, June-Woo, et al.
Veröffentlicht: (2026)
von: Kim, June-Woo, et al.
Veröffentlicht: (2026)
Ensemble Successor Representations for Task Generalization in Offline-to-Online Reinforcement Learning
von: Wang, Changhong, et al.
Veröffentlicht: (2024)
von: Wang, Changhong, et al.
Veröffentlicht: (2024)
FORLER: Federated Offline Reinforcement Learning with Q-Ensemble and Actor Rectification
von: Qiao, Nan, et al.
Veröffentlicht: (2026)
von: Qiao, Nan, et al.
Veröffentlicht: (2026)
RLLTE: Long-Term Evolution Project of Reinforcement Learning
von: Yuan, Mingqi, et al.
Veröffentlicht: (2023)
von: Yuan, Mingqi, et al.
Veröffentlicht: (2023)
UACER: An Uncertainty-Adaptive Critic Ensemble Framework for Robust Adversarial Reinforcement Learning
von: Wu, Jiaxi, et al.
Veröffentlicht: (2025)
von: Wu, Jiaxi, et al.
Veröffentlicht: (2025)
FineFT: Efficient and Risk-Aware Ensemble Reinforcement Learning for Futures Trading
von: Qin, Molei, et al.
Veröffentlicht: (2025)
von: Qin, Molei, et al.
Veröffentlicht: (2025)
Ensemble Distributionally Robust Bayesian Optimisation
von: Ramazyan, Tigran, et al.
Veröffentlicht: (2026)
von: Ramazyan, Tigran, et al.
Veröffentlicht: (2026)
ENOTO: Improving Offline-to-Online Reinforcement Learning with Q-Ensembles
von: Zhao, Kai, et al.
Veröffentlicht: (2023)
von: Zhao, Kai, et al.
Veröffentlicht: (2023)
Semantic-based Distributed Learning for Diverse and Discriminative Representations
von: Tian, Zhuojun, et al.
Veröffentlicht: (2026)
von: Tian, Zhuojun, et al.
Veröffentlicht: (2026)
Jumanji: a Diverse Suite of Scalable Reinforcement Learning Environments in JAX
von: Bonnet, Clément, et al.
Veröffentlicht: (2023)
von: Bonnet, Clément, et al.
Veröffentlicht: (2023)
Off-Policy Safe Reinforcement Learning with Constrained Optimistic Exploration
von: Li, Guopeng, et al.
Veröffentlicht: (2026)
von: Li, Guopeng, et al.
Veröffentlicht: (2026)
Proximal Policy Gradient Arborescence for Quality Diversity Reinforcement Learning
von: Batra, Sumeet, et al.
Veröffentlicht: (2023)
von: Batra, Sumeet, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
How Ensembles of Distilled Policies Improve Generalisation in Reinforcement Learning
von: Weltevrede, Max, et al.
Veröffentlicht: (2025) -
Contextual Similarity Distillation: Ensemble Uncertainties with a Single Model
von: Zanger, Moritz A., et al.
Veröffentlicht: (2025) -
On the Equivalence of Random Network Distillation, Deep Ensembles, and Bayesian Inference
von: Zanger, Moritz A., et al.
Veröffentlicht: (2026) -
Explore-Go: Leveraging Exploration for Generalisation in Deep Reinforcement Learning
von: Weltevrede, Max, et al.
Veröffentlicht: (2024) -
Epistemic Monte Carlo Tree Search
von: Oren, Yaniv, et al.
Veröffentlicht: (2022)