Beyond Bellman: High-Order Generator Regression for Continuous-Time Policy Evaluation
Fuente:
arXiv
Guardado en:
| Autores principales: | Zheng, Yaowei, Zhang, Richong, Wu, Shenxi, Bian, Shirui, Zhang, Haosong, Zeng, Li, Ma, Xingjian, Zhang, Yichi |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Policy Gradient for Continuous-Time Mean-Field Control
por: Bayraktar, Erhan, et al.
Publicado: (2026)
por: Bayraktar, Erhan, et al.
Publicado: (2026)
Infinite time horizon stochastic recursive control problems with jumps: dynamic programming and stochastic verification theorems
por: Luo, Sheng, et al.
Publicado: (2024)
por: Luo, Sheng, et al.
Publicado: (2024)
Optimal Control of Unbounded Functional Stochastic Evolution Systems in Hilbert Spaces: Second-Order Path-dependent HJB Equation
por: Tang, Shanjian, et al.
Publicado: (2024)
por: Tang, Shanjian, et al.
Publicado: (2024)
Convergence and turnpike properties of linear-quadratic mean field control problems with common noise
por: Bayraktar, Erhan, et al.
Publicado: (2026)
por: Bayraktar, Erhan, et al.
Publicado: (2026)
Ergodicity and turnpike properties of linear-quadratic mean field control problems
por: Bayraktar, Erhan, et al.
Publicado: (2025)
por: Bayraktar, Erhan, et al.
Publicado: (2025)
Mean field social optimization: feedback person-by-person optimality and the dynamic programming equation
por: Huang, Minyi, et al.
Publicado: (2025)
por: Huang, Minyi, et al.
Publicado: (2025)
Stochastic Control with Signatures
por: Bank, P., et al.
Publicado: (2024)
por: Bank, P., et al.
Publicado: (2024)
High risk aversion Merton's problem without transversality conditions
por: Biffis, Enrico, et al.
Publicado: (2025)
por: Biffis, Enrico, et al.
Publicado: (2025)
Turnpike Property of a Linear-Quadratic Optimal Control Problem in Large Horizons with Regime Switching II: Non-Homogeneous Cases
por: Mei, Hongwei, et al.
Publicado: (2025)
por: Mei, Hongwei, et al.
Publicado: (2025)
Dynamic Programming Principle and Stabilization for Mean-Field Quantum Filtering Systems
por: Chalal, Sofiane, et al.
Publicado: (2026)
por: Chalal, Sofiane, et al.
Publicado: (2026)
Risk-averse optimization under distributional uncertainty with Rockafellian relaxation
por: Antil, Harbir, et al.
Publicado: (2026)
por: Antil, Harbir, et al.
Publicado: (2026)
Rockafellian Relaxation for PDE-Constrained Optimization with Distributional Uncertainty
por: Antil, Harbir, et al.
Publicado: (2024)
por: Antil, Harbir, et al.
Publicado: (2024)
Causal Hamilton-Jacobi-Bellman Equations for Anticipative Stochastic Optimal Control
por: Bank, Peter, et al.
Publicado: (2025)
por: Bank, Peter, et al.
Publicado: (2025)
Continuous time Stochastic optimal control under discrete time partial observations
por: Bayer, Christian, et al.
Publicado: (2024)
por: Bayer, Christian, et al.
Publicado: (2024)
Stochastic Optimal Impulse Controls with Changing Running Costs
por: Cao, Yuchen, et al.
Publicado: (2025)
por: Cao, Yuchen, et al.
Publicado: (2025)
Viscosity Solutions of Second Order Path-Dependent Partial Differential Equations and Applications
por: Tang, Shanjian, et al.
Publicado: (2024)
por: Tang, Shanjian, et al.
Publicado: (2024)
A Pontryagin Maximum Principle on the Belief Space for Continuous-Time Optimal Control with Discrete Observations
por: Bayer, Christian, et al.
Publicado: (2025)
por: Bayer, Christian, et al.
Publicado: (2025)
Robust Ergodic Control of Jump-Diffusion Systems under Drift and Intensity Uncertainty
por: Azze, Abel, et al.
Publicado: (2026)
por: Azze, Abel, et al.
Publicado: (2026)
Turnpike and dissipativity in generalized discrete-time stochastic linear-quadratic optimal control
por: Schießl, Jonas, et al.
Publicado: (2023)
por: Schießl, Jonas, et al.
Publicado: (2023)
Turnpike Property of Stochastic Linear-Quadratic Optimal Control Problems in Large Horizons with Regime Switching I: Homogeneous Cases
por: Mei, Hongwei, et al.
Publicado: (2025)
por: Mei, Hongwei, et al.
Publicado: (2025)
Optimal power procurement for green cellular wireless networks under uncertainty and chance constraints
por: Rached, Nadhir Ben, et al.
Publicado: (2025)
por: Rached, Nadhir Ben, et al.
Publicado: (2025)
Viscosity-Informed Generative Actor-Critic for High-Dimensional Stochastic Optimal Control
por: Golpashin, Alen E., et al.
Publicado: (2026)
por: Golpashin, Alen E., et al.
Publicado: (2026)
Average Cost Optimality of Partially Observed MDPS: Contraction of Non-linear Filters, Optimal Solutions and Approximations
por: Demirci, Yunus Emre, et al.
Publicado: (2023)
por: Demirci, Yunus Emre, et al.
Publicado: (2023)
Sample Complexity of Policy Gradient for Log-Growth Control
por: Pan, Qiuhua, et al.
Publicado: (2026)
por: Pan, Qiuhua, et al.
Publicado: (2026)
Open-loop and closed-loop solvabilities for zero-sum stochastic linear quadratic differential games of Markovian regime switching system
por: Wu, Fan, et al.
Publicado: (2024)
por: Wu, Fan, et al.
Publicado: (2024)
Stochastic linear-quadratic differential game with Markovian jumps in an infinite horizon
por: Wu, Fan, et al.
Publicado: (2024)
por: Wu, Fan, et al.
Publicado: (2024)
Indefinite Stochastic Linear-Quadratic Optimal Control Problems with Random Coefficients and Poisson Jumps: Closed-Loop Representation of Open-Loop Optimal Controls
por: Ding, Kai, et al.
Publicado: (2026)
por: Ding, Kai, et al.
Publicado: (2026)
Stochastic internal habit formation and optimality
por: Aleandri, Michele, et al.
Publicado: (2025)
por: Aleandri, Michele, et al.
Publicado: (2025)
Non-local Hamilton-Jacobi-Bellman equations for the stochastic optimal control of path-dependent piecewise deterministic processes
por: Bandini, Elena, et al.
Publicado: (2024)
por: Bandini, Elena, et al.
Publicado: (2024)
Reconciling Discrete-Time Mixed Policies and Continuous-Time Relaxed Controls in Reinforcement Learning and Stochastic Control
por: Carmona, Rene, et al.
Publicado: (2025)
por: Carmona, Rene, et al.
Publicado: (2025)
Discrete-Time Approximations of Controlled Diffusions with Infinite Horizon Discounted and Average Cost
por: Pradhan, Somnath, et al.
Publicado: (2025)
por: Pradhan, Somnath, et al.
Publicado: (2025)
Reinforcement Learning, Optimal Control, and Bayesian Filtering in Data Assimilation
por: Hammoud, Abed
Publicado: (2026)
por: Hammoud, Abed
Publicado: (2026)
Lifting partial smoothing to solve HJB equations and stochastic control problems
por: Gozzi, Fausto, et al.
Publicado: (2023)
por: Gozzi, Fausto, et al.
Publicado: (2023)
Stochastic modeling of cyclic cancer treatments under common noise
por: Sonith, Jason
Publicado: (2024)
por: Sonith, Jason
Publicado: (2024)
Second-Order $Λ$-Sets and Extensions to Non-Smooth, Hybrid, and Stochastic Optimal Control
por: Rashid, Mohammad H. M
Publicado: (2025)
por: Rashid, Mohammad H. M
Publicado: (2025)
Weakly-Coupled Multi-Action Restless Bandits -- Exponential Convergence in Probability
por: Fu, Jing, et al.
Publicado: (2026)
por: Fu, Jing, et al.
Publicado: (2026)
A Reinforcement Learning Framework for Some Singular Stochastic Control Problems
por: Liang, Zongxia, et al.
Publicado: (2025)
por: Liang, Zongxia, et al.
Publicado: (2025)
Reinforcement learning for irreversible reinsurance problems: the randomized singular control approach
por: Liang, Zongxia, et al.
Publicado: (2025)
por: Liang, Zongxia, et al.
Publicado: (2025)
Optimal control for production inventory system with various cost criterion
por: Golui, Subrata, et al.
Publicado: (2022)
por: Golui, Subrata, et al.
Publicado: (2022)
Infinite Time Horizon Optimal Control of McKean-Vlasov SDEs
por: Rudà, Silvia
Publicado: (2025)
por: Rudà, Silvia
Publicado: (2025)
Ejemplares similares
-
Policy Gradient for Continuous-Time Mean-Field Control
por: Bayraktar, Erhan, et al.
Publicado: (2026) -
Infinite time horizon stochastic recursive control problems with jumps: dynamic programming and stochastic verification theorems
por: Luo, Sheng, et al.
Publicado: (2024) -
Optimal Control of Unbounded Functional Stochastic Evolution Systems in Hilbert Spaces: Second-Order Path-dependent HJB Equation
por: Tang, Shanjian, et al.
Publicado: (2024) -
Convergence and turnpike properties of linear-quadratic mean field control problems with common noise
por: Bayraktar, Erhan, et al.
Publicado: (2026) -
Ergodicity and turnpike properties of linear-quadratic mean field control problems
por: Bayraktar, Erhan, et al.
Publicado: (2025)