Sample Complexity of Policy Gradient for Log-Growth Control
Fuente:
arXiv
Saved in:
| Main Authors: | Pan, Qiuhua, Shen, Yukai, Zhang, Liwei, Chen, Cailian, Guan, Xinping |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Stability and Sensitivity Analysis of Relative Temporal-Difference Learning: Extended Version
by: Sakha, Masoud S., et al.
Published: (2026)
by: Sakha, Masoud S., et al.
Published: (2026)
Optimistic Training and Convergence of Q-Learning -- Extended Version
by: Mehta, Prashant, et al.
Published: (2026)
by: Mehta, Prashant, et al.
Published: (2026)
Asynchronous Stochastic Approximation with Applications to Average-Reward Reinforcement Learning
by: Yu, Huizhen, et al.
Published: (2024)
by: Yu, Huizhen, et al.
Published: (2024)
Average-reward reinforcement learning in semi-Markov decision processes via relative value iteration
by: Yu, Huizhen, et al.
Published: (2025)
by: Yu, Huizhen, et al.
Published: (2025)
A Note on Stability in Asynchronous Stochastic Approximation without Communication Delays
by: Yu, Huizhen, et al.
Published: (2023)
by: Yu, Huizhen, et al.
Published: (2023)
Beyond Bellman: High-Order Generator Regression for Continuous-Time Policy Evaluation
by: Zheng, Yaowei, et al.
Published: (2026)
by: Zheng, Yaowei, et al.
Published: (2026)
Deep Relaxation of Controlled Stochastic Gradient Descent via Singular Perturbations
by: Bardi, Martino, et al.
Published: (2022)
by: Bardi, Martino, et al.
Published: (2022)
Stochastic Control with Signatures
by: Bank, P., et al.
Published: (2024)
by: Bank, P., et al.
Published: (2024)
Policy Gradient for Continuous-Time Mean-Field Control
by: Bayraktar, Erhan, et al.
Published: (2026)
by: Bayraktar, Erhan, et al.
Published: (2026)
Multi-Robot Relative Pose Estimation in SE(2) with Observability Analysis: A Comparison of Extended Kalman Filtering and Robust Pose Graph Optimization
by: Shin, Kihoon, et al.
Published: (2024)
by: Shin, Kihoon, et al.
Published: (2024)
Finite-Time Analysis of Projected Two-Time-Scale Stochastic Approximation
by: Bai, Yitao, et al.
Published: (2026)
by: Bai, Yitao, et al.
Published: (2026)
On the Rate of Gaussian Approximation for Linear Regression Problems
by: Khusainov, Marat, et al.
Published: (2025)
by: Khusainov, Marat, et al.
Published: (2025)
Continuous-time Risk-sensitive Reinforcement Learning via Quadratic Variation Penalty
by: Jia, Yanwei
Published: (2024)
by: Jia, Yanwei
Published: (2024)
Gaussian Approximation and Multiplier Bootstrap for Stochastic Gradient Descent
by: Sheshukova, Marina, et al.
Published: (2025)
by: Sheshukova, Marina, et al.
Published: (2025)
Causal Hamilton-Jacobi-Bellman Equations for Anticipative Stochastic Optimal Control
by: Bank, Peter, et al.
Published: (2025)
by: Bank, Peter, et al.
Published: (2025)
A measure-valued HJB perspective on Bayesian optimal adaptive control
by: Cox, Alexander M. G., et al.
Published: (2025)
by: Cox, Alexander M. G., et al.
Published: (2025)
Lifting partial smoothing to solve HJB equations and stochastic control problems
by: Gozzi, Fausto, et al.
Published: (2023)
by: Gozzi, Fausto, et al.
Published: (2023)
Neural Actor-Critic Methods for Hamilton-Jacobi-Bellman PDEs: Asymptotic Analysis and Numerical Studies
by: Cohen, Samuel N., et al.
Published: (2025)
by: Cohen, Samuel N., et al.
Published: (2025)
A geometric ensemble method for Bayesian inference
by: Popov, Andrey A
Published: (2025)
by: Popov, Andrey A
Published: (2025)
Optimal control of SDEs with merely measurable drift: an HJB approach
by: Du, Kai, et al.
Published: (2025)
by: Du, Kai, et al.
Published: (2025)
A Sequential Testing Problem with Signal Control
by: Campbell, Steven, et al.
Published: (2025)
by: Campbell, Steven, et al.
Published: (2025)
Reinforcement Learning, Optimal Control, and Bayesian Filtering in Data Assimilation
by: Hammoud, Abed
Published: (2026)
by: Hammoud, Abed
Published: (2026)
The Score Kalman Filter
by: Iwasaki, Kaito, et al.
Published: (2026)
by: Iwasaki, Kaito, et al.
Published: (2026)
Average Cost Optimality of Partially Observed MDPS: Contraction of Non-linear Filters, Optimal Solutions and Approximations
by: Demirci, Yunus Emre, et al.
Published: (2023)
by: Demirci, Yunus Emre, et al.
Published: (2023)
Optimal Control of a Stochastic Power System -- Algorithms and Mathematical Analysis
by: Wang, Zhen, et al.
Published: (2024)
by: Wang, Zhen, et al.
Published: (2024)
Infinite time horizon stochastic recursive control problems with jumps: dynamic programming and stochastic verification theorems
by: Luo, Sheng, et al.
Published: (2024)
by: Luo, Sheng, et al.
Published: (2024)
Discrete-Time Approximations of Controlled Diffusions with Infinite Horizon Discounted and Average Cost
by: Pradhan, Somnath, et al.
Published: (2025)
by: Pradhan, Somnath, et al.
Published: (2025)
Reflected stochastic recursive control problems with jumps: dynamic programming and stochastic verification theorems
by: Liu, Lu, et al.
Published: (2025)
by: Liu, Lu, et al.
Published: (2025)
Optimal Feedback Control in Social Networks in a McKean-Vlasov-Friedkin-Johnsen System
by: Pramanik, Paramahansa
Published: (2025)
by: Pramanik, Paramahansa
Published: (2025)
Open-loop and closed-loop solvabilities for zero-sum stochastic linear quadratic differential games of Markovian regime switching system
by: Wu, Fan, et al.
Published: (2024)
by: Wu, Fan, et al.
Published: (2024)
Stochastic linear-quadratic differential game with Markovian jumps in an infinite horizon
by: Wu, Fan, et al.
Published: (2024)
by: Wu, Fan, et al.
Published: (2024)
Continuous Policy and Value Iteration for Stochastic Control Problems and Its Convergence
by: Feng, Qi, et al.
Published: (2025)
by: Feng, Qi, et al.
Published: (2025)
Flow Matching for Efficient and Scalable Data Assimilation
by: Transue, Taos, et al.
Published: (2025)
by: Transue, Taos, et al.
Published: (2025)
Convergence of Neural Network Policies for Risk--Reward Optimization
by: Chen, Chang, et al.
Published: (2026)
by: Chen, Chang, et al.
Published: (2026)
Continuous time Stochastic optimal control under discrete time partial observations
by: Bayer, Christian, et al.
Published: (2024)
by: Bayer, Christian, et al.
Published: (2024)
Weakly-Coupled Multi-Action Restless Bandits -- Exponential Convergence in Probability
by: Fu, Jing, et al.
Published: (2026)
by: Fu, Jing, et al.
Published: (2026)
Reconciling Discrete-Time Mixed Policies and Continuous-Time Relaxed Controls in Reinforcement Learning and Stochastic Control
by: Carmona, Rene, et al.
Published: (2025)
by: Carmona, Rene, et al.
Published: (2025)
Non-Expansive Mappings in Two-Time-Scale Stochastic Approximation: Finite-Time Analysis
by: Chandak, Siddharth
Published: (2025)
by: Chandak, Siddharth
Published: (2025)
Markov Decision Processes of the Third Kind: Learning Distributions by Policy Gradient Descent
by: Bäuerle, Nicole, et al.
Published: (2026)
by: Bäuerle, Nicole, et al.
Published: (2026)
Characterizing nonconvex boundaries via scalarization
by: Ma, Jin, et al.
Published: (2025)
by: Ma, Jin, et al.
Published: (2025)
Similar Items
-
Stability and Sensitivity Analysis of Relative Temporal-Difference Learning: Extended Version
by: Sakha, Masoud S., et al.
Published: (2026) -
Optimistic Training and Convergence of Q-Learning -- Extended Version
by: Mehta, Prashant, et al.
Published: (2026) -
Asynchronous Stochastic Approximation with Applications to Average-Reward Reinforcement Learning
by: Yu, Huizhen, et al.
Published: (2024) -
Average-reward reinforcement learning in semi-Markov decision processes via relative value iteration
by: Yu, Huizhen, et al.
Published: (2025) -
A Note on Stability in Asynchronous Stochastic Approximation without Communication Delays
by: Yu, Huizhen, et al.
Published: (2023)