Stability and Sensitivity Analysis of Relative Temporal-Difference Learning: Extended Version
Fuente:
arXiv
Saved in:
| Main Authors: | Sakha, Masoud S., Kamalapurkar, Rushikesh, Meyn, Sean |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Optimistic Training and Convergence of Q-Learning -- Extended Version
by: Mehta, Prashant, et al.
Published: (2026)
by: Mehta, Prashant, et al.
Published: (2026)
Continuous-time Risk-sensitive Reinforcement Learning via Quadratic Variation Penalty
by: Jia, Yanwei
Published: (2024)
by: Jia, Yanwei
Published: (2024)
Sample Complexity of Policy Gradient for Log-Growth Control
by: Pan, Qiuhua, et al.
Published: (2026)
by: Pan, Qiuhua, et al.
Published: (2026)
On the Rate of Gaussian Approximation for Linear Regression Problems
by: Khusainov, Marat, et al.
Published: (2025)
by: Khusainov, Marat, et al.
Published: (2025)
Exploratory Randomization for Discrete-Time Linear Exponential Quadratic Gaussian (LEQG) Problem
by: Lleo, Sebastien, et al.
Published: (2025)
by: Lleo, Sebastien, et al.
Published: (2025)
Policy Gradient for Continuous-Time Mean-Field Control
by: Bayraktar, Erhan, et al.
Published: (2026)
by: Bayraktar, Erhan, et al.
Published: (2026)
Exploratory Randomization for Discrete-Time Risk-Sensitive Benchmarked Investment Management with Reinforcement Learning
by: Lleo, Sebastien, et al.
Published: (2026)
by: Lleo, Sebastien, et al.
Published: (2026)
Stochastic Control with Signatures
by: Bank, P., et al.
Published: (2024)
by: Bank, P., et al.
Published: (2024)
Beyond Bellman: High-Order Generator Regression for Continuous-Time Policy Evaluation
by: Zheng, Yaowei, et al.
Published: (2026)
by: Zheng, Yaowei, et al.
Published: (2026)
Optimal control of SDEs with merely measurable drift: an HJB approach
by: Du, Kai, et al.
Published: (2025)
by: Du, Kai, et al.
Published: (2025)
Finite-Time Analysis of Projected Two-Time-Scale Stochastic Approximation
by: Bai, Yitao, et al.
Published: (2026)
by: Bai, Yitao, et al.
Published: (2026)
Mean--Variance Portfolio Selection by Continuous-Time Reinforcement Learning: Algorithms, Regret Analysis, and Empirical Study
by: Huang, Yilie, et al.
Published: (2024)
by: Huang, Yilie, et al.
Published: (2024)
Infinite time horizon stochastic recursive control problems with jumps: dynamic programming and stochastic verification theorems
by: Luo, Sheng, et al.
Published: (2024)
by: Luo, Sheng, et al.
Published: (2024)
Gaussian Approximation and Multiplier Bootstrap for Stochastic Gradient Descent
by: Sheshukova, Marina, et al.
Published: (2025)
by: Sheshukova, Marina, et al.
Published: (2025)
Asynchronous Stochastic Approximation with Applications to Average-Reward Reinforcement Learning
by: Yu, Huizhen, et al.
Published: (2024)
by: Yu, Huizhen, et al.
Published: (2024)
A Note on Stability in Asynchronous Stochastic Approximation without Communication Delays
by: Yu, Huizhen, et al.
Published: (2023)
by: Yu, Huizhen, et al.
Published: (2023)
Stability of long run functionals with respect to stationary Markov controls
by: Stettner, Lukasz
Published: (2024)
by: Stettner, Lukasz
Published: (2024)
Markovian Foundations for Quasi-Stochastic Approximation with Applications to Extremum Seeking Control
by: Lauand, Caio Kalil, et al.
Published: (2022)
by: Lauand, Caio Kalil, et al.
Published: (2022)
Continuous time Stochastic optimal control under discrete time partial observations
by: Bayer, Christian, et al.
Published: (2024)
by: Bayer, Christian, et al.
Published: (2024)
Discrete-Time Approximations of Controlled Diffusions with Infinite Horizon Discounted and Average Cost
by: Pradhan, Somnath, et al.
Published: (2025)
by: Pradhan, Somnath, et al.
Published: (2025)
Reflected stochastic recursive control problems with jumps: dynamic programming and stochastic verification theorems
by: Liu, Lu, et al.
Published: (2025)
by: Liu, Lu, et al.
Published: (2025)
Optimal Feedback Control in Social Networks in a McKean-Vlasov-Friedkin-Johnsen System
by: Pramanik, Paramahansa
Published: (2025)
by: Pramanik, Paramahansa
Published: (2025)
Open-loop and closed-loop solvabilities for zero-sum stochastic linear quadratic differential games of Markovian regime switching system
by: Wu, Fan, et al.
Published: (2024)
by: Wu, Fan, et al.
Published: (2024)
Stochastic linear-quadratic differential game with Markovian jumps in an infinite horizon
by: Wu, Fan, et al.
Published: (2024)
by: Wu, Fan, et al.
Published: (2024)
Average-reward reinforcement learning in semi-Markov decision processes via relative value iteration
by: Yu, Huizhen, et al.
Published: (2025)
by: Yu, Huizhen, et al.
Published: (2025)
Reinforcement Learning, Optimal Control, and Bayesian Filtering in Data Assimilation
by: Hammoud, Abed
Published: (2026)
by: Hammoud, Abed
Published: (2026)
A measure-valued HJB perspective on Bayesian optimal adaptive control
by: Cox, Alexander M. G., et al.
Published: (2025)
by: Cox, Alexander M. G., et al.
Published: (2025)
Multi-Robot Relative Pose Estimation in SE(2) with Observability Analysis: A Comparison of Extended Kalman Filtering and Robust Pose Graph Optimization
by: Shin, Kihoon, et al.
Published: (2024)
by: Shin, Kihoon, et al.
Published: (2024)
Controllability and Vector Potential
by: Shankar, Shiva
Published: (2019)
by: Shankar, Shiva
Published: (2019)
Stability and performance of stochastic economic MPC - Stochastic characterization of the closed-loop asymptotics
by: Schießl, Jonas, et al.
Published: (2025)
by: Schießl, Jonas, et al.
Published: (2025)
Continuous Policy and Value Iteration for Stochastic Control Problems and Its Convergence
by: Feng, Qi, et al.
Published: (2025)
by: Feng, Qi, et al.
Published: (2025)
An optimal level of Stubbornness to win a soccer match
by: Pramanik, Paramahansa
Published: (2025)
by: Pramanik, Paramahansa
Published: (2025)
Path integral control under McKean-Vlasov dynamics
by: Bennett, Timothy
Published: (2024)
by: Bennett, Timothy
Published: (2024)
Reinforcement Learning in Real Option Models
by: Dianetti, Jodi, et al.
Published: (2026)
by: Dianetti, Jodi, et al.
Published: (2026)
Turnpike Property of a Linear-Quadratic Optimal Control Problem in Large Horizons with Regime Switching II: Non-Homogeneous Cases
by: Mei, Hongwei, et al.
Published: (2025)
by: Mei, Hongwei, et al.
Published: (2025)
Lifting partial smoothing to solve HJB equations and stochastic control problems
by: Gozzi, Fausto, et al.
Published: (2023)
by: Gozzi, Fausto, et al.
Published: (2023)
Neural Actor-Critic Methods for Hamilton-Jacobi-Bellman PDEs: Asymptotic Analysis and Numerical Studies
by: Cohen, Samuel N., et al.
Published: (2025)
by: Cohen, Samuel N., et al.
Published: (2025)
Causal Hamilton-Jacobi-Bellman Equations for Anticipative Stochastic Optimal Control
by: Bank, Peter, et al.
Published: (2025)
by: Bank, Peter, et al.
Published: (2025)
Dynamic Programming Principle and Stabilization for Mean-Field Quantum Filtering Systems
by: Chalal, Sofiane, et al.
Published: (2026)
by: Chalal, Sofiane, et al.
Published: (2026)
Reconciling Discrete-Time Mixed Policies and Continuous-Time Relaxed Controls in Reinforcement Learning and Stochastic Control
by: Carmona, Rene, et al.
Published: (2025)
by: Carmona, Rene, et al.
Published: (2025)
Similar Items
-
Optimistic Training and Convergence of Q-Learning -- Extended Version
by: Mehta, Prashant, et al.
Published: (2026) -
Continuous-time Risk-sensitive Reinforcement Learning via Quadratic Variation Penalty
by: Jia, Yanwei
Published: (2024) -
Sample Complexity of Policy Gradient for Log-Growth Control
by: Pan, Qiuhua, et al.
Published: (2026) -
On the Rate of Gaussian Approximation for Linear Regression Problems
by: Khusainov, Marat, et al.
Published: (2025) -
Exploratory Randomization for Discrete-Time Linear Exponential Quadratic Gaussian (LEQG) Problem
by: Lleo, Sebastien, et al.
Published: (2025)