CM-DQN: A Value-Based Deep Reinforcement Learning Model to Simulate Confirmation Bias
Fuente:
arXiv
Saved in:
| Main Authors: | Shen, Jiacheng, Feng, Lihan |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
GB-DQN: Gradient Boosted DQN Models for Non-stationary Reinforcement Learning
by: Lee, Chang-Hwan, et al.
Published: (2025)
by: Lee, Chang-Hwan, et al.
Published: (2025)
A Comparative Study of Deep Reinforcement Learning Models: DQN vs PPO vs A2C
by: De La Fuente, Neil, et al.
Published: (2024)
by: De La Fuente, Neil, et al.
Published: (2024)
SF-DQN: Provable Knowledge Transfer using Successor Feature for Deep Reinforcement Learning
by: Zhang, Shuai, et al.
Published: (2024)
by: Zhang, Shuai, et al.
Published: (2024)
Value of Information-Enhanced Exploration in Bootstrapped DQN
by: Plataniotis, Stergios, et al.
Published: (2025)
by: Plataniotis, Stergios, et al.
Published: (2025)
$β$-DQN: Improving Deep Q-Learning By Evolving the Behavior
by: Zhang, Hongming, et al.
Published: (2025)
by: Zhang, Hongming, et al.
Published: (2025)
An Arbitration Control for an Ensemble of Diversified DQN variants in Continual Reinforcement Learning
by: Jang, Wonseo, et al.
Published: (2025)
by: Jang, Wonseo, et al.
Published: (2025)
TreeDQN: Sample-Efficient Off-Policy Reinforcement Learning for Combinatorial Optimization
by: Sorokin, D., et al.
Published: (2023)
by: Sorokin, D., et al.
Published: (2023)
Dynamic Operating System Scheduling Using Double DQN: A Reinforcement Learning Approach to Task Optimization
by: Sun, Xiaoxuan, et al.
Published: (2025)
by: Sun, Xiaoxuan, et al.
Published: (2025)
Predicting E-commerce Purchase Behavior using a DQN-Inspired Deep Learning Model for enhanced adaptability
by: Jain, Aditi Madhusudan
Published: (2025)
by: Jain, Aditi Madhusudan
Published: (2025)
Confirmation Bias in Gaussian Mixture Models
by: Balanov, Amnon, et al.
Published: (2024)
by: Balanov, Amnon, et al.
Published: (2024)
Hybrid DQN-TD3 Reinforcement Learning for Autonomous Navigation in Dynamic Environments
by: He, Xiaoyi, et al.
Published: (2025)
by: He, Xiaoyi, et al.
Published: (2025)
A Controlled Study of Double DQN and Dueling DQN Under Cross-Environment Transfer
by: Nasir, Azkaa, et al.
Published: (2026)
by: Nasir, Azkaa, et al.
Published: (2026)
Beyond ReLU: Chebyshev-DQN for Enhanced Deep Q-Networks
by: Yazdannik, Saman, et al.
Published: (2025)
by: Yazdannik, Saman, et al.
Published: (2025)
Failing to Falsify: Evaluating and Mitigating Confirmation Bias in Language Models
by: Jhaveri, Ayush Rajesh, et al.
Published: (2026)
by: Jhaveri, Ayush Rajesh, et al.
Published: (2026)
Deep Q-Network (DQN) multi-agent reinforcement learning (MARL) for Stock Trading
by: Tidwell, John Christopher, et al.
Published: (2025)
by: Tidwell, John Christopher, et al.
Published: (2025)
Value-Based Deep Multi-Agent Reinforcement Learning with Dynamic Sparse Training
by: Hu, Pihe, et al.
Published: (2024)
by: Hu, Pihe, et al.
Published: (2024)
Does DQN Learn?
by: Gopalan, Aditya, et al.
Published: (2022)
by: Gopalan, Aditya, et al.
Published: (2022)
Unveiling Confirmation Bias in Chain-of-Thought Reasoning
by: Wan, Yue, et al.
Published: (2025)
by: Wan, Yue, et al.
Published: (2025)
Towards the Mitigation of Confirmation Bias in Semi-supervised Learning: a Debiased Training Perspective
by: Wang, Yu, et al.
Published: (2024)
by: Wang, Yu, et al.
Published: (2024)
Value-Distributional Model-Based Reinforcement Learning
by: Luis, Carlos E., et al.
Published: (2023)
by: Luis, Carlos E., et al.
Published: (2023)
Fast Value Tracking for Deep Reinforcement Learning
by: Shih, Frank, et al.
Published: (2024)
by: Shih, Frank, et al.
Published: (2024)
Secure Resource Allocation via Constrained Deep Reinforcement Learning
by: Sun, Jianfei, et al.
Published: (2025)
by: Sun, Jianfei, et al.
Published: (2025)
Solving Continuous Mean Field Games: Deep Reinforcement Learning for Non-Stationary Dynamics
by: Magnino, Lorenzo, et al.
Published: (2025)
by: Magnino, Lorenzo, et al.
Published: (2025)
DGLight: DQN-Guided GRPO Fine-Tuning of Large Language Models for Traffic Signal Control
by: Yu, Chenbo
Published: (2026)
by: Yu, Chenbo
Published: (2026)
Improving the Diversity of Bootstrapped DQN by Replacing Priors With Noise
by: Meng, Li, et al.
Published: (2022)
by: Meng, Li, et al.
Published: (2022)
Stabilizing Reinforcement Learning for Diffusion Language Models
by: Zhong, Jianyuan, et al.
Published: (2026)
by: Zhong, Jianyuan, et al.
Published: (2026)
Off-Policy Value-Based Reinforcement Learning for Large Language Models
by: Wang, Peng-Yuan, et al.
Published: (2026)
by: Wang, Peng-Yuan, et al.
Published: (2026)
A Theoretical Understanding of Gradient Bias in Meta-Reinforcement Learning
by: Feng, Xidong, et al.
Published: (2021)
by: Feng, Xidong, et al.
Published: (2021)
Issues with Value-Based Multi-objective Reinforcement Learning: Value Function Interference and Overestimation Sensitivity
by: Vamplew, Peter, et al.
Published: (2024)
by: Vamplew, Peter, et al.
Published: (2024)
SimBa: Simplicity Bias for Scaling Up Parameters in Deep Reinforcement Learning
by: Lee, Hojoon, et al.
Published: (2024)
by: Lee, Hojoon, et al.
Published: (2024)
Fisher-Guided Selective Forgetting: Mitigating The Primacy Bias in Deep Reinforcement Learning
by: Falzari, Massimiliano, et al.
Published: (2025)
by: Falzari, Massimiliano, et al.
Published: (2025)
Enhancing Few-Shot Learning with Integrated Data and GAN Model Approaches
by: Feng, Yinqiu, et al.
Published: (2024)
by: Feng, Yinqiu, et al.
Published: (2024)
Counterfactual Reward Model Training for Bias Mitigation in Multimodal Reinforcement Learning
by: Mathew, Sheryl, et al.
Published: (2025)
by: Mathew, Sheryl, et al.
Published: (2025)
Reinforcement Learning for Finite Space Mean-Field Type Games
by: Shao, Kai, et al.
Published: (2024)
by: Shao, Kai, et al.
Published: (2024)
Context-Aware Adaptive Sampling for Intelligent Data Acquisition Systems Using DQN
by: Huang, Weiqiang, et al.
Published: (2025)
by: Huang, Weiqiang, et al.
Published: (2025)
Semi-supervised learning via DQN for log anomaly detection
by: He, Yingying, et al.
Published: (2024)
by: He, Yingying, et al.
Published: (2024)
On the Reuse Bias in Off-Policy Reinforcement Learning
by: Ying, Chengyang, et al.
Published: (2022)
by: Ying, Chengyang, et al.
Published: (2022)
Factored Value Functions for Graph-Based Multi-Agent Reinforcement Learning
by: Rashwan, Ahmed, et al.
Published: (2026)
by: Rashwan, Ahmed, et al.
Published: (2026)
InterQ: A DQN Framework for Optimal Intermittent Control
by: Aggarwal, Shubham, et al.
Published: (2025)
by: Aggarwal, Shubham, et al.
Published: (2025)
Is there Value in Reinforcement Learning?
by: Fox, Lior, et al.
Published: (2025)
by: Fox, Lior, et al.
Published: (2025)
Similar Items
-
GB-DQN: Gradient Boosted DQN Models for Non-stationary Reinforcement Learning
by: Lee, Chang-Hwan, et al.
Published: (2025) -
A Comparative Study of Deep Reinforcement Learning Models: DQN vs PPO vs A2C
by: De La Fuente, Neil, et al.
Published: (2024) -
SF-DQN: Provable Knowledge Transfer using Successor Feature for Deep Reinforcement Learning
by: Zhang, Shuai, et al.
Published: (2024) -
Value of Information-Enhanced Exploration in Bootstrapped DQN
by: Plataniotis, Stergios, et al.
Published: (2025) -
$β$-DQN: Improving Deep Q-Learning By Evolving the Behavior
by: Zhang, Hongming, et al.
Published: (2025)