Accuracy of Discretely Sampled Stochastic Policies in Continuous-time Reinforcement Learning
Fuente:
arXiv
Saved in:
| Main Authors: | Jia, Yanwei, Ouyang, Du, Zhang, Yufei |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Stochastic Langevin Differential Inclusions with Applications to Machine Learning
by: Difonzo, Fabio V., et al.
Published: (2022)
by: Difonzo, Fabio V., et al.
Published: (2022)
Examining Policy Entropy of Reinforcement Learning Agents for Personalization Tasks
by: Dereventsov, Anton, et al.
Published: (2022)
by: Dereventsov, Anton, et al.
Published: (2022)
Nonlinear Assimilation via Score-based Sequential Langevin Sampling
by: Ding, Zhao, et al.
Published: (2024)
by: Ding, Zhao, et al.
Published: (2024)
Policy Gradient with Second Order Momentum
by: Sun, Tianyu
Published: (2025)
by: Sun, Tianyu
Published: (2025)
Using Linearized Optimal Transport to Predict the Evolution of Stochastic Particle Systems
by: Karris, Nicholas, et al.
Published: (2024)
by: Karris, Nicholas, et al.
Published: (2024)
Adaptive Lipschitz-Free Conditional Gradient Methods for Stochastic Composite Nonconvex Optimization
by: Yuan, Ganzhao
Published: (2026)
by: Yuan, Ganzhao
Published: (2026)
UAdam: Unified Adam-Type Algorithmic Framework for Non-Convex Stochastic Optimization
by: Jiang, Yiming, et al.
Published: (2023)
by: Jiang, Yiming, et al.
Published: (2023)
A Trust-Region Interior-Point Stochastic Sequential Quadratic Programming Method
by: Fang, Yuchen, et al.
Published: (2026)
by: Fang, Yuchen, et al.
Published: (2026)
OptEMA: Adaptive Exponential Moving Average for Stochastic Optimization with Zero-Noise Optimality
by: Yuan, Ganzhao
Published: (2026)
by: Yuan, Ganzhao
Published: (2026)
On the Convergence of the Gradient Descent Method with Stochastic Fixed-point Rounding Errors under the Polyak-Lojasiewicz Inequality
by: Xia, Lu, et al.
Published: (2023)
by: Xia, Lu, et al.
Published: (2023)
Fundamental Bias in Inverting Random Sampling Matrices with Application to Sub-sampled Newton
by: Niu, Chengmei, et al.
Published: (2025)
by: Niu, Chengmei, et al.
Published: (2025)
Fast Unconstrained Optimization via Hessian Averaging and Adaptive Gradient Sampling Methods
by: O'Leary-Roseberry, Thomas, et al.
Published: (2024)
by: O'Leary-Roseberry, Thomas, et al.
Published: (2024)
Why is Normalization Preferred? A Worst-Case Complexity Theory for Stochastically Preconditioned SGD under Heavy-Tailed Noise
by: Fang, Yuchen, et al.
Published: (2026)
by: Fang, Yuchen, et al.
Published: (2026)
Gradient is All You Need? How Consensus-Based Optimization can be Interpreted as a Stochastic Relaxation of Gradient Descent
by: Riedl, Konstantin, et al.
Published: (2023)
by: Riedl, Konstantin, et al.
Published: (2023)
Stochastic Optimal Control Matching
by: Domingo-Enrich, Carles, et al.
Published: (2023)
by: Domingo-Enrich, Carles, et al.
Published: (2023)
On the Stability Connection Between Discrete-Time Algorithms and Their Resolution ODEs: Applications to Min-Max Optimisation
by: Farzin, Amir Ali, et al.
Published: (2026)
by: Farzin, Amir Ali, et al.
Published: (2026)
A note on continuous-time online learning
by: Ying, Lexing
Published: (2024)
by: Ying, Lexing
Published: (2024)
Trust-Region Sequential Quadratic Programming for Stochastic Optimization with Random Models
by: Fang, Yuchen, et al.
Published: (2024)
by: Fang, Yuchen, et al.
Published: (2024)
Learning Regularization Functionals for Inverse Problems: A Comparative Study
by: Hertrich, Johannes, et al.
Published: (2025)
by: Hertrich, Johannes, et al.
Published: (2025)
Unsupervised Ground Metric Learning
by: Auffenberg, Janis, et al.
Published: (2025)
by: Auffenberg, Janis, et al.
Published: (2025)
Real-time optimal control of high-dimensional parametrized systems by deep learning-based reduced order models
by: Tomasetto, Matteo, et al.
Published: (2024)
by: Tomasetto, Matteo, et al.
Published: (2024)
Learning incomplete factorization preconditioners for GMRES
by: Häusner, Paul, et al.
Published: (2024)
by: Häusner, Paul, et al.
Published: (2024)
Continuous Data Assimilation with Learned Surrogate Dynamics
by: Li, Wenwen, et al.
Published: (2026)
by: Li, Wenwen, et al.
Published: (2026)
Machine Learning and Control: Foundations, Advances, and Perspectives
by: Zuazua, Enrique
Published: (2025)
by: Zuazua, Enrique
Published: (2025)
Control, Optimal Transport and Neural Differential Equations in Supervised Learning
by: Phung, Minh-Nhat, et al.
Published: (2025)
by: Phung, Minh-Nhat, et al.
Published: (2025)
Learning truly monotone operators with applications to nonlinear inverse problems
by: Belkouchi, Younes, et al.
Published: (2024)
by: Belkouchi, Younes, et al.
Published: (2024)
Curvature-Aware Optimization for High-Accuracy Physics-Informed Neural Networks
by: Jnini, Anas, et al.
Published: (2026)
by: Jnini, Anas, et al.
Published: (2026)
End-to-End Mesh Optimization of a Hybrid Deep Learning Black-Box PDE Solver
by: Ma, Shaocong, et al.
Published: (2024)
by: Ma, Shaocong, et al.
Published: (2024)
Online Statistical Inference of Constrained Stochastic Optimization via Random Scaling
by: Du, Xinchen, et al.
Published: (2025)
by: Du, Xinchen, et al.
Published: (2025)
A Discrete Variational Derivation of Accelerated Methods in Optimization
by: Campos, Cédric M., et al.
Published: (2021)
by: Campos, Cédric M., et al.
Published: (2021)
Fast and Provable Tensor-Train Format Tensor Completion via Precondtioned Riemannian Gradient Descent
by: Bian, Fengmiao, et al.
Published: (2025)
by: Bian, Fengmiao, et al.
Published: (2025)
Riemannian Preconditioned LoRA for Fine-Tuning Foundation Models
by: Zhang, Fangzhao, et al.
Published: (2024)
by: Zhang, Fangzhao, et al.
Published: (2024)
Learning rate adaptive stochastic gradient descent optimization methods: numerical simulations for deep learning methods for partial differential equations and convergence analyses
by: Dereich, Steffen, et al.
Published: (2024)
by: Dereich, Steffen, et al.
Published: (2024)
WinQ: Accelerating Quantization-Aware Training of Language Models Around Saddle Points
by: Li, Dongyue, et al.
Published: (2026)
by: Li, Dongyue, et al.
Published: (2026)
Enhanced Adaptive Gradient Algorithms for Nonconvex-PL Minimax Optimization
by: Huang, Feihu, et al.
Published: (2023)
by: Huang, Feihu, et al.
Published: (2023)
On Bellman equations for continuous-time policy evaluation I: discretization and approximation
by: Mou, Wenlong, et al.
Published: (2024)
by: Mou, Wenlong, et al.
Published: (2024)
Stochastic Regret Guarantees for Online Zeroth- and First-Order Bilevel Optimization
by: Nazari, Parvin, et al.
Published: (2025)
by: Nazari, Parvin, et al.
Published: (2025)
ZORMS-LfD: Learning from Demonstrations with Zeroth-Order Random Matrix Search
by: Dry, Olivia, et al.
Published: (2025)
by: Dry, Olivia, et al.
Published: (2025)
A Family of Controllable Momentum Coefficients for Forward-Backward Accelerated Algorithms
by: Fu, Mingwei, et al.
Published: (2025)
by: Fu, Mingwei, et al.
Published: (2025)
AutoBalance: An Automatic Balancing Framework for Training Physics-Informed Neural Networks
by: An, Kang, et al.
Published: (2025)
by: An, Kang, et al.
Published: (2025)
Similar Items
-
Stochastic Langevin Differential Inclusions with Applications to Machine Learning
by: Difonzo, Fabio V., et al.
Published: (2022) -
Examining Policy Entropy of Reinforcement Learning Agents for Personalization Tasks
by: Dereventsov, Anton, et al.
Published: (2022) -
Nonlinear Assimilation via Score-based Sequential Langevin Sampling
by: Ding, Zhao, et al.
Published: (2024) -
Policy Gradient with Second Order Momentum
by: Sun, Tianyu
Published: (2025) -
Using Linearized Optimal Transport to Predict the Evolution of Stochastic Particle Systems
by: Karris, Nicholas, et al.
Published: (2024)