Ensemble Elastic DQN: A novel multi-step ensemble approach to address overestimation in deep value-based reinforcement learning
Fuente:
arXiv
Saved in:
| Main Authors: | Ly, Adrian, Dazeley, Richard, Vamplew, Peter, Cruz, Francisco, Aryal, Sunil |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Issues with Value-Based Multi-objective Reinforcement Learning: Value Function Interference and Overestimation Sensitivity
by: Vamplew, Peter, et al.
Published: (2024)
by: Vamplew, Peter, et al.
Published: (2024)
An Empirical Investigation of Value-Based Multi-objective Reinforcement Learning for Stochastic Environments
by: Ding, Kewen, et al.
Published: (2024)
by: Ding, Kewen, et al.
Published: (2024)
Adaptive Alignment: Dynamic Preference Adjustments via Multi-Objective Reinforcement Learning for Pluralistic AI
by: Harland, Hadassah, et al.
Published: (2024)
by: Harland, Hadassah, et al.
Published: (2024)
AI Apology: A Critical Review of Apology in AI Systems
by: Harland, Hadassah, et al.
Published: (2024)
by: Harland, Hadassah, et al.
Published: (2024)
Learning the Value Systems of Societies with Preference-based Multi-objective Reinforcement Learning
by: Holgado-Sánchez, Andrés, et al.
Published: (2026)
by: Holgado-Sánchez, Andrés, et al.
Published: (2026)
Learning Rewards, Not Labels: Adversarial Inverse Reinforcement Learning for Machinery Fault Detection
by: Neupane, Dhiraj, et al.
Published: (2026)
by: Neupane, Dhiraj, et al.
Published: (2026)
Data-driven Machinery Fault Diagnosis: A Comprehensive Review
by: Neupane, Dhiraj, et al.
Published: (2024)
by: Neupane, Dhiraj, et al.
Published: (2024)
Handling Out-of-Distribution Data: A Survey
by: Tamang, Lakpa, et al.
Published: (2025)
by: Tamang, Lakpa, et al.
Published: (2025)
Margin-bounded Confidence Scores for Out-of-Distribution Detection
by: Tamang, Lakpa D., et al.
Published: (2024)
by: Tamang, Lakpa D., et al.
Published: (2024)
Multi-objective Reinforcement Learning: A Tool for Pluralistic Alignment
by: Vamplew, Peter, et al.
Published: (2024)
by: Vamplew, Peter, et al.
Published: (2024)
Modified Double DQN: addressing stability
by: Halat, Shervin, et al.
Published: (2021)
by: Halat, Shervin, et al.
Published: (2021)
Deep Q-Network (DQN) multi-agent reinforcement learning (MARL) for Stock Trading
by: Tidwell, John Christopher, et al.
Published: (2025)
by: Tidwell, John Christopher, et al.
Published: (2025)
A novel multi-agent dynamic portfolio optimization learning system based on hierarchical deep reinforcement learning
by: Sun, Ruoyu, et al.
Published: (2025)
by: Sun, Ruoyu, et al.
Published: (2025)
Deep Learning for Sports Video Event Detection: Tasks, Datasets, Methods, and Challenges
by: Xu, Hao, et al.
Published: (2025)
by: Xu, Hao, et al.
Published: (2025)
An approach of deep reinforcement learning for maximizing the net present value of stochastic projects
by: Xu, Wei, et al.
Published: (2025)
by: Xu, Wei, et al.
Published: (2025)
On the consistency of hyper-parameter selection in value-based deep reinforcement learning
by: Obando-Ceron, Johan, et al.
Published: (2024)
by: Obando-Ceron, Johan, et al.
Published: (2024)
Decoding non-invasive brain activity with novel deep-learning approaches
by: Csaky, Richard
Published: (2025)
by: Csaky, Richard
Published: (2025)
História do esporte no cenário internacional: visão geral
by: Wray Vamplew
Published: (2013)
by: Wray Vamplew
Published: (2013)
Multi-objective Reinforcement Learning With Augmented States Requires Rewards After Deployment
by: Vamplew, Peter, et al.
Published: (2026)
by: Vamplew, Peter, et al.
Published: (2026)
Adaptive traffic signal safety and efficiency improvement by multi objective deep reinforcement learning approach
by: Mirbakhsh, Shahin, et al.
Published: (2024)
by: Mirbakhsh, Shahin, et al.
Published: (2024)
In value-based deep reinforcement learning, a pruned network is a good network
by: Obando-Ceron, Johan, et al.
Published: (2024)
by: Obando-Ceron, Johan, et al.
Published: (2024)
Generalized Bayesian deep reinforcement learning
by: Roy, Shreya Sinha, et al.
Published: (2024)
by: Roy, Shreya Sinha, et al.
Published: (2024)
A computational approach to visual ecology with deep reinforcement learning
by: Sokoloski, Sacha, et al.
Published: (2024)
by: Sokoloski, Sacha, et al.
Published: (2024)
Rescue path planning for urban flood: A deep reinforcement learning–based approach
by: Xiao‐Yan Li, et al.
Published: (2024)
by: Xiao‐Yan Li, et al.
Published: (2024)
Contrastive learning-based agent modeling for deep reinforcement learning
by: Ma, Wenhao, et al.
Published: (2023)
by: Ma, Wenhao, et al.
Published: (2023)
An Arbitration Control for an Ensemble of Diversified DQN variants in Continual Reinforcement Learning
by: Jang, Wonseo, et al.
Published: (2025)
by: Jang, Wonseo, et al.
Published: (2025)
An ensemble deep learning approach to detect tumors on Mohs micrographic surgery slides
by: Yilmaz, Abdurrahim, et al.
Published: (2025)
by: Yilmaz, Abdurrahim, et al.
Published: (2025)
Quantile deep learning models for multi-step ahead time series prediction
by: Cheung, Jimmy, et al.
Published: (2024)
by: Cheung, Jimmy, et al.
Published: (2024)
Heating ventilation air‐conditioner system for multi‐regional commercial buildings based on deep reinforcement learning
by: Juan Yang, et al.
Published: (2024)
by: Juan Yang, et al.
Published: (2024)
Shared autonomy between human electroencephalography and TD3 deep reinforcement learning: A multi‐agent copilot approach
by: Chun‐Ren Phang, et al.
Published: (2025)
by: Chun‐Ren Phang, et al.
Published: (2025)
Breast radiation therapy fluence painting with multi‐agent deep reinforcement learning
by: Yang Dongrong, et al.
Published: (2025)
by: Yang Dongrong, et al.
Published: (2025)
On Generalization Across Environments In Multi-Objective Reinforcement Learning
by: Teoh, Jayden, et al.
Published: (2025)
by: Teoh, Jayden, et al.
Published: (2025)
ES-C51: Expected Sarsa Based C51 Distributional Reinforcement Learning Algorithm
by: Tandon, Rijul, et al.
Published: (2025)
by: Tandon, Rijul, et al.
Published: (2025)
Influence of multi‐walled carbon nanotubes on mechanical characteristics of glass fiber reinforced polymer composites: An experimental and analytical approach
by: Sunil Kumar Chaudhary, et al.
Published: (2024)
by: Sunil Kumar Chaudhary, et al.
Published: (2024)
Riddled basin geometry sets fundamental limits to predictability and reproducibility in deep learning
by: Ly, Andrew, et al.
Published: (2025)
by: Ly, Andrew, et al.
Published: (2025)
Bridging an energy system model with an ensemble deep-learning approach for electricity price forecasting
by: Amor, Souhir Ben, et al.
Published: (2024)
by: Amor, Souhir Ben, et al.
Published: (2024)
An ensemble-based approach for multi-fidelity emulation and adaptive sampling
by: Mohammadi, Hossein
Published: (2026)
by: Mohammadi, Hossein
Published: (2026)
Software bug localization based on optimized and ensembled deep learning models
by: Waqas Ali, et al.
Published: (2024)
by: Waqas Ali, et al.
Published: (2024)
Autism spectrum disorder identification using multi‐model deep ensemble classifier with transfer learning
by: Lakmini Herath, et al.
Published: (2024)
by: Lakmini Herath, et al.
Published: (2024)
Multi‐class brain tumor diagnosis using MRI: A dynamic reinforcement ensemble learning approach
by: Jia Sheng Yang, et al.
Published: (2025)
by: Jia Sheng Yang, et al.
Published: (2025)
Similar Items
-
Issues with Value-Based Multi-objective Reinforcement Learning: Value Function Interference and Overestimation Sensitivity
by: Vamplew, Peter, et al.
Published: (2024) -
An Empirical Investigation of Value-Based Multi-objective Reinforcement Learning for Stochastic Environments
by: Ding, Kewen, et al.
Published: (2024) -
Adaptive Alignment: Dynamic Preference Adjustments via Multi-Objective Reinforcement Learning for Pluralistic AI
by: Harland, Hadassah, et al.
Published: (2024) -
AI Apology: A Critical Review of Apology in AI Systems
by: Harland, Hadassah, et al.
Published: (2024) -
Learning the Value Systems of Societies with Preference-based Multi-objective Reinforcement Learning
by: Holgado-Sánchez, Andrés, et al.
Published: (2026)