Bellman Error Centering
Fuente:
arXiv
Saved in:
| Main Authors: | Chen, Xingguo, Gong, Yu, Yang, Shangdong, Wang, Wenhao |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
A Variance Minimization Approach to Temporal-Difference Learning
by: Chen, Xingguo, et al.
Published: (2024)
by: Chen, Xingguo, et al.
Published: (2024)
Symmetric Q-learning: Reducing Skewness of Bellman Error in Online Reinforcement Learning
by: Omura, Motoki, et al.
Published: (2024)
by: Omura, Motoki, et al.
Published: (2024)
The Role of Inherent Bellman Error in Offline Reinforcement Learning with Linear Function Approximation
by: Golowich, Noah, et al.
Published: (2024)
by: Golowich, Noah, et al.
Published: (2024)
Regularized Centered Emphatic Temporal Difference Learning
by: Chen, Xingguo, et al.
Published: (2026)
by: Chen, Xingguo, et al.
Published: (2026)
A Generalized Projected Bellman Error for Off-policy Value Estimation in Reinforcement Learning
by: Patterson, Andrew, et al.
Published: (2021)
by: Patterson, Andrew, et al.
Published: (2021)
Behavior-Induced Mirror-Prox Temporal-Difference Learning for Faster Off-Policy Prediction
by: Chen, Xingguo, et al.
Published: (2026)
by: Chen, Xingguo, et al.
Published: (2026)
Parameterized Projected Bellman Operator
by: Vincent, Théo, et al.
Published: (2023)
by: Vincent, Théo, et al.
Published: (2023)
Gradual Transition from Bellman Optimality Operator to Bellman Operator in Online Reinforcement Learning
by: Omura, Motoki, et al.
Published: (2025)
by: Omura, Motoki, et al.
Published: (2025)
Theoretical Barriers in Bellman-Based Reinforcement Learning
by: Pinon, Brieuc, et al.
Published: (2025)
by: Pinon, Brieuc, et al.
Published: (2025)
Model-Based Offline Reinforcement Learning with Adversarial Data Augmentation
by: Cao, Hongye, et al.
Published: (2025)
by: Cao, Hongye, et al.
Published: (2025)
Behavior-Aware Auxiliary Corrections for Off-Policy Temporal-Difference Prediction
by: Chen, Xingguo, et al.
Published: (2026)
by: Chen, Xingguo, et al.
Published: (2026)
Bellman operator convergence enhancements in reinforcement learning algorithms
by: Kadurha, David Krame, et al.
Published: (2025)
by: Kadurha, David Krame, et al.
Published: (2025)
MICRO: Model-Based Offline Reinforcement Learning with a Conservative Bellman Operator
by: Liu, Xiao-Yin, et al.
Published: (2023)
by: Liu, Xiao-Yin, et al.
Published: (2023)
Adaptive Regularization of Representation Rank as an Implicit Constraint of Bellman Equation
by: He, Qiang, et al.
Published: (2024)
by: He, Qiang, et al.
Published: (2024)
Prospect-Theory Behavior from Bellman Optimality in MDPs with Catastrophic States
by: Chen, Yujiao
Published: (2026)
by: Chen, Yujiao
Published: (2026)
Linear Bellman Completeness Suffices for Efficient Online Reinforcement Learning with Few Actions
by: Golowich, Noah, et al.
Published: (2024)
by: Golowich, Noah, et al.
Published: (2024)
Path-Coupled Bellman Flows for Distributional Reinforcement Learning
by: Xu, Boyang, et al.
Published: (2026)
by: Xu, Boyang, et al.
Published: (2026)
Iterated $Q$-Network: Beyond One-Step Bellman Updates in Deep Reinforcement Learning
by: Vincent, Théo, et al.
Published: (2024)
by: Vincent, Théo, et al.
Published: (2024)
Reward Redistribution for CVaR MDPs using a Bellman Operator on L-infinity
by: Muni, Aneri, et al.
Published: (2026)
by: Muni, Aneri, et al.
Published: (2026)
Understanding the theoretical properties of projected Bellman equation, linear Q-learning, and approximate value iteration
by: Lim, Han-Dong, et al.
Published: (2025)
by: Lim, Han-Dong, et al.
Published: (2025)
Bellman Unbiasedness: Toward Provably Efficient Distributional Reinforcement Learning with General Value Function Approximation
by: Cho, Taehyun, et al.
Published: (2024)
by: Cho, Taehyun, et al.
Published: (2024)
Trajectory Bellman Residual Minimization: A Simple Value-Based Method for LLM Reasoning
by: Yuan, Yurun, et al.
Published: (2025)
by: Yuan, Yurun, et al.
Published: (2025)
Automatic Reward Shaping from Multi-Objective Human Heuristics
by: Xie, Yuqing, et al.
Published: (2025)
by: Xie, Yuqing, et al.
Published: (2025)
FedMABench: Benchmarking Mobile Agents on Decentralized Heterogeneous User Data
by: Wang, Wenhao, et al.
Published: (2025)
by: Wang, Wenhao, et al.
Published: (2025)
SABER: Small Actions, Big Errors -- Safeguarding Mutating Steps in LLM Agents
by: Cuadron, Alejandro, et al.
Published: (2025)
by: Cuadron, Alejandro, et al.
Published: (2025)
Binary Split Categorical feature with Mean Absolute Error Criteria in CART
by: Yu, Peng, et al.
Published: (2025)
by: Yu, Peng, et al.
Published: (2025)
VitalBench: A Rigorous Multi-Center Benchmark for Long-Term Vital Sign Prediction in Intraoperative Care
by: Cai, Xiuding, et al.
Published: (2025)
by: Cai, Xiuding, et al.
Published: (2025)
Error Distribution Smoothing:Advancing Low-Dimensional Imbalanced Regression
by: Chen, Donghe, et al.
Published: (2025)
by: Chen, Donghe, et al.
Published: (2025)
TextBFGS: A Case-Based Reasoning Approach to Code Optimization via Error-Operator Retrieval
by: Zhang, Zizheng, et al.
Published: (2026)
by: Zhang, Zizheng, et al.
Published: (2026)
Misspecified $Q$-Learning with Sparse Linear Function Approximation: Tight Bounds on Approximation Error
by: Du, Ally Yalei, et al.
Published: (2024)
by: Du, Ally Yalei, et al.
Published: (2024)
Towards Monotonic Improvement in In-Context Reinforcement Learning
by: Zhang, Wenhao, et al.
Published: (2025)
by: Zhang, Wenhao, et al.
Published: (2025)
ErrorEraser: Unlearning Data Bias for Improved Continual Learning
by: Cao, Xuemei, et al.
Published: (2025)
by: Cao, Xuemei, et al.
Published: (2025)
Mitigating Prior Errors in Causal Structure Learning: A Resilient Approach via Bayesian Networks
by: Chen, Lyuzhou, et al.
Published: (2023)
by: Chen, Lyuzhou, et al.
Published: (2023)
Factor Decorrelation Enhanced Data Removal from Deep Predictive Models
by: Yang, Wenhao, et al.
Published: (2025)
by: Yang, Wenhao, et al.
Published: (2025)
Mitigating Estimation Bias with Representation Learning in TD Error-Driven Regularization
by: Chen, Haohui, et al.
Published: (2025)
by: Chen, Haohui, et al.
Published: (2025)
Double Actor-Critic with TD Error-Driven Regularization in Reinforcement Learning
by: Chen, Haohui, et al.
Published: (2024)
by: Chen, Haohui, et al.
Published: (2024)
Adviser-Actor-Critic: Eliminating Steady-State Error in Reinforcement Learning Control
by: Chen, Donghe, et al.
Published: (2025)
by: Chen, Donghe, et al.
Published: (2025)
CFG-OEC: Classifier Free Guidance with Orthogonal Error Correction
by: Yang, Nakgyu, et al.
Published: (2025)
by: Yang, Nakgyu, et al.
Published: (2025)
CausalMed: Causality-Based Personalized Medication Recommendation Centered on Patient health state
by: Li, Xiang, et al.
Published: (2024)
by: Li, Xiang, et al.
Published: (2024)
Error Analysis of Shapley Value-Based Model Explanations: An Informative Perspective
by: Zhao, Ningsheng, et al.
Published: (2024)
by: Zhao, Ningsheng, et al.
Published: (2024)
Similar Items
-
A Variance Minimization Approach to Temporal-Difference Learning
by: Chen, Xingguo, et al.
Published: (2024) -
Symmetric Q-learning: Reducing Skewness of Bellman Error in Online Reinforcement Learning
by: Omura, Motoki, et al.
Published: (2024) -
The Role of Inherent Bellman Error in Offline Reinforcement Learning with Linear Function Approximation
by: Golowich, Noah, et al.
Published: (2024) -
Regularized Centered Emphatic Temporal Difference Learning
by: Chen, Xingguo, et al.
Published: (2026) -
A Generalized Projected Bellman Error for Off-policy Value Estimation in Reinforcement Learning
by: Patterson, Andrew, et al.
Published: (2021)