On Bellman equations for continuous-time policy evaluation I: discretization and approximation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Mou, Wenlong, Zhu, Yuhua |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Statistical guarantees for continuous-time policy evaluation: blessing of ellipticity and new tradeoffs
von: Mou, Wenlong
Veröffentlicht: (2025)
von: Mou, Wenlong
Veröffentlicht: (2025)
Numerical method for feasible and approximately optimal solutions of multi-marginal optimal transport beyond discrete measures
von: Neufeld, Ariel, et al.
Veröffentlicht: (2022)
von: Neufeld, Ariel, et al.
Veröffentlicht: (2022)
Non-asymptotic convergence analysis of the stochastic gradient Hamiltonian Monte Carlo algorithm with discontinuous stochastic gradient with applications to training of ReLU neural networks
von: Liang, Luxu, et al.
Veröffentlicht: (2024)
von: Liang, Luxu, et al.
Veröffentlicht: (2024)
Nonparametric Filtering, Estimation and Classification using Neural Jump ODEs
von: Heiss, Jakob, et al.
Veröffentlicht: (2024)
von: Heiss, Jakob, et al.
Veröffentlicht: (2024)
DeepMartingale: Duality of the Optimal Stopping Problem with Expressivity and High-Dimensional Hedging
von: Ye, Junyan, et al.
Veröffentlicht: (2025)
von: Ye, Junyan, et al.
Veröffentlicht: (2025)
Langevin dynamics based algorithm e-TH$\varepsilon$O POULA for stochastic optimization problems with discontinuous stochastic gradient
von: Lim, Dong-Young, et al.
Veröffentlicht: (2022)
von: Lim, Dong-Young, et al.
Veröffentlicht: (2022)
Stochastic Optimal Control Matching
von: Domingo-Enrich, Carles, et al.
Veröffentlicht: (2023)
von: Domingo-Enrich, Carles, et al.
Veröffentlicht: (2023)
Constrained Density Estimation via Optimal Transport
von: Hu, Yinan, et al.
Veröffentlicht: (2026)
von: Hu, Yinan, et al.
Veröffentlicht: (2026)
Generalization Bounds for Sparse Random Feature Expansions
von: Hashemi, Abolfazl, et al.
Veröffentlicht: (2021)
von: Hashemi, Abolfazl, et al.
Veröffentlicht: (2021)
Faster Computation of Entropic Optimal Transport via Stable Low Frequency Modes
von: Chhaibi, Reda, et al.
Veröffentlicht: (2025)
von: Chhaibi, Reda, et al.
Veröffentlicht: (2025)
Feasible approximation of matching equilibria for large-scale matching for teams problems
von: Neufeld, Ariel, et al.
Veröffentlicht: (2023)
von: Neufeld, Ariel, et al.
Veröffentlicht: (2023)
Numerical method for approximately optimal solutions of two-stage distributionally robust optimization with marginal constraints
von: Neufeld, Ariel, et al.
Veröffentlicht: (2022)
von: Neufeld, Ariel, et al.
Veröffentlicht: (2022)
A note on continuous-time online learning
von: Ying, Lexing
Veröffentlicht: (2024)
von: Ying, Lexing
Veröffentlicht: (2024)
Provably convergent stochastic fixed-point algorithm for free-support Wasserstein barycenter of continuous non-parametric measures
von: Chen, Zeyi, et al.
Veröffentlicht: (2025)
von: Chen, Zeyi, et al.
Veröffentlicht: (2025)
Error analysis for stochastic gradient optimization schemes using modified equations
von: Bréhier, Charles-Edouard, et al.
Veröffentlicht: (2024)
von: Bréhier, Charles-Edouard, et al.
Veröffentlicht: (2024)
Polynomial Scaling is Possible For Neural Operator Approximations of Structured Families of BSDEs
von: Furuya, Takashi, et al.
Veröffentlicht: (2024)
von: Furuya, Takashi, et al.
Veröffentlicht: (2024)
Neural Operators Can Play Dynamic Stackelberg Games
von: Alvarez, Guillermo, et al.
Veröffentlicht: (2024)
von: Alvarez, Guillermo, et al.
Veröffentlicht: (2024)
Is RL fine-tuning harder than regression? A PDE learning approach for diffusion models
von: Mou, Wenlong
Veröffentlicht: (2025)
von: Mou, Wenlong
Veröffentlicht: (2025)
Averaged Adam accelerates stochastic optimization in the training of deep neural network approximations for partial differential equation and optimal control problems
von: Dereich, Steffen, et al.
Veröffentlicht: (2025)
von: Dereich, Steffen, et al.
Veröffentlicht: (2025)
Error estimates for finite-dimensional approximations of Hamilton-Jacobi-Bellman equations on the Wasserstein space
von: Daudin, Samuel, et al.
Veröffentlicht: (2025)
von: Daudin, Samuel, et al.
Veröffentlicht: (2025)
Optimal and instance-dependent guarantees for Markovian linear stochastic approximation
von: Mou, Wenlong, et al.
Veröffentlicht: (2021)
von: Mou, Wenlong, et al.
Veröffentlicht: (2021)
Convergence analysis for an implementable scheme to solve the linear-quadratic stochastic optimal control problem with stochastic wave equation
von: Chaudhary, Abhishek
Veröffentlicht: (2025)
von: Chaudhary, Abhishek
Veröffentlicht: (2025)
Well-posedness and approximation of reflected McKean-Vlasov SDEs with applications
von: Hinds, P. D., et al.
Veröffentlicht: (2024)
von: Hinds, P. D., et al.
Veröffentlicht: (2024)
Contributions to Robust and Efficient Methods for Analysis of High Dimensional Data
von: Yang, Kai
Veröffentlicht: (2025)
von: Yang, Kai
Veröffentlicht: (2025)
Complexity of Zeroth- and First-order Stochastic Trust-Region Algorithms
von: Ha, Yunsoo, et al.
Veröffentlicht: (2024)
von: Ha, Yunsoo, et al.
Veröffentlicht: (2024)
Deep learning algorithms for FBSDEs with jumps: Applications to option pricing and a MFG model for smart grids
von: Alasseur, Clémence, et al.
Veröffentlicht: (2024)
von: Alasseur, Clémence, et al.
Veröffentlicht: (2024)
Numerical method for nonlinear Kolmogorov PDEs via sensitivity analysis
von: Bartl, Daniel, et al.
Veröffentlicht: (2024)
von: Bartl, Daniel, et al.
Veröffentlicht: (2024)
A New Look at the Ensemble Kalman Filter for Inverse Problems: Duality, Non-Asymptotic Analysis and Convergence Acceleration
von: Krishnanunni, C G, et al.
Veröffentlicht: (2026)
von: Krishnanunni, C G, et al.
Veröffentlicht: (2026)
Wasserstein Gradient Flows of the Discrepancy with Distance Kernel on the Line
von: Hertrich, Johannes, et al.
Veröffentlicht: (2023)
von: Hertrich, Johannes, et al.
Veröffentlicht: (2023)
Wasserstein Steepest Descent Flows of Discrepancies with Riesz Kernels
von: Hertrich, Johannes, et al.
Veröffentlicht: (2022)
von: Hertrich, Johannes, et al.
Veröffentlicht: (2022)
Non-exchangeable evolutionary and mean field games and their applications
von: Yoshioka, H., et al.
Veröffentlicht: (2025)
von: Yoshioka, H., et al.
Veröffentlicht: (2025)
Linking PageRank, Time Reversal, and Policy Evaluation
von: Avrachenkov, Konstantin, et al.
Veröffentlicht: (2026)
von: Avrachenkov, Konstantin, et al.
Veröffentlicht: (2026)
Accuracy of Discretely Sampled Stochastic Policies in Continuous-time Reinforcement Learning
von: Jia, Yanwei, et al.
Veröffentlicht: (2025)
von: Jia, Yanwei, et al.
Veröffentlicht: (2025)
On the existence of extremal solutions for the conjugate discrete-time Riccati equation
von: Chiang, Chun-Yueh
Veröffentlicht: (2024)
von: Chiang, Chun-Yueh
Veröffentlicht: (2024)
Learning rate adaptive stochastic gradient descent optimization methods: numerical simulations for deep learning methods for partial differential equations and convergence analyses
von: Dereich, Steffen, et al.
Veröffentlicht: (2024)
von: Dereich, Steffen, et al.
Veröffentlicht: (2024)
State-dependent temperature control in Langevin diffusions using numerical exploratory Hamiltonian-Jacobi-Bellman equations
von: Wang, Taorui, et al.
Veröffentlicht: (2026)
von: Wang, Taorui, et al.
Veröffentlicht: (2026)
Multilevel Picard scheme for solving high-dimensional drift control problems with state constraints
von: Zhong, Yuan
Veröffentlicht: (2025)
von: Zhong, Yuan
Veröffentlicht: (2025)
Real-time optimal control of high-dimensional parametrized systems by deep learning-based reduced order models
von: Tomasetto, Matteo, et al.
Veröffentlicht: (2024)
von: Tomasetto, Matteo, et al.
Veröffentlicht: (2024)
Assigning Stationary Distributions to Sparse Stochastic Matrices
von: Gillis, Nicolas, et al.
Veröffentlicht: (2023)
von: Gillis, Nicolas, et al.
Veröffentlicht: (2023)
Dynamic Proximal Gradient Algorithms for Schatten-$p$ Quasi-Norm Regularized Problems
von: Shen, Weiping, et al.
Veröffentlicht: (2026)
von: Shen, Weiping, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Statistical guarantees for continuous-time policy evaluation: blessing of ellipticity and new tradeoffs
von: Mou, Wenlong
Veröffentlicht: (2025) -
Numerical method for feasible and approximately optimal solutions of multi-marginal optimal transport beyond discrete measures
von: Neufeld, Ariel, et al.
Veröffentlicht: (2022) -
Non-asymptotic convergence analysis of the stochastic gradient Hamiltonian Monte Carlo algorithm with discontinuous stochastic gradient with applications to training of ReLU neural networks
von: Liang, Luxu, et al.
Veröffentlicht: (2024) -
Nonparametric Filtering, Estimation and Classification using Neural Jump ODEs
von: Heiss, Jakob, et al.
Veröffentlicht: (2024) -
DeepMartingale: Duality of the Optimal Stopping Problem with Expressivity and High-Dimensional Hedging
von: Ye, Junyan, et al.
Veröffentlicht: (2025)