Statistical guarantees for continuous-time policy evaluation: blessing of ellipticity and new tradeoffs
Fuente:
arXiv
Gespeichert in:
| 1. Verfasser: | Mou, Wenlong |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Optimal and instance-dependent guarantees for Markovian linear stochastic approximation
von: Mou, Wenlong, et al.
Veröffentlicht: (2021)
von: Mou, Wenlong, et al.
Veröffentlicht: (2021)
Is RL fine-tuning harder than regression? A PDE learning approach for diffusion models
von: Mou, Wenlong
Veröffentlicht: (2025)
von: Mou, Wenlong
Veröffentlicht: (2025)
Continuous-time reinforcement learning: ellipticity enables model-free value function approximation
von: Mou, Wenlong
Veröffentlicht: (2026)
von: Mou, Wenlong
Veröffentlicht: (2026)
On Bellman equations for continuous-time policy evaluation I: discretization and approximation
von: Mou, Wenlong, et al.
Veröffentlicht: (2024)
von: Mou, Wenlong, et al.
Veröffentlicht: (2024)
Statistical Inference for Linear Functionals of Online SGD in High-dimensional Linear Regression
von: Agrawalla, Bhavya, et al.
Veröffentlicht: (2023)
von: Agrawalla, Bhavya, et al.
Veröffentlicht: (2023)
A Piecewise Lyapunov Analysis of Sub-quadratic SGD: Applications to Robust and Quantile Regression
von: Zhang, Yixuan, et al.
Veröffentlicht: (2025)
von: Zhang, Yixuan, et al.
Veröffentlicht: (2025)
Early Stopping in Contextual Bandits and Inferences
von: Cui, Zihan
Veröffentlicht: (2025)
von: Cui, Zihan
Veröffentlicht: (2025)
Fast Spawn\&Prune (FS\&P): Global convergence of stochastic conic particle gradient descent via birth/death process
von: De Castro, Yohann, et al.
Veröffentlicht: (2026)
von: De Castro, Yohann, et al.
Veröffentlicht: (2026)
High-dimensional scaling limits and fluctuations of online least-squares SGD with smooth covariance
von: Balasubramanian, Krishnakumar, et al.
Veröffentlicht: (2023)
von: Balasubramanian, Krishnakumar, et al.
Veröffentlicht: (2023)
Convergence of coordinate ascent variational inference for log-concave measures via optimal transport
von: Arnese, Manuel, et al.
Veröffentlicht: (2024)
von: Arnese, Manuel, et al.
Veröffentlicht: (2024)
Alternating minimization for generalized rank one matrix sensing: Sharp predictions from a random initialization
von: Chandrasekher, Kabir Aladin, et al.
Veröffentlicht: (2022)
von: Chandrasekher, Kabir Aladin, et al.
Veröffentlicht: (2022)
4+3 Phases of Compute-Optimal Neural Scaling Laws
von: Paquette, Elliot, et al.
Veröffentlicht: (2024)
von: Paquette, Elliot, et al.
Veröffentlicht: (2024)
Orlicz regrets to consistently bound statistics of random variables with an application to environmental indicators
von: Yoshioka, Hidekazu, et al.
Veröffentlicht: (2023)
von: Yoshioka, Hidekazu, et al.
Veröffentlicht: (2023)
Markov Kernels, Distances and Optimal Control: A Parable of Linear Quadratic Non-Gaussian Distribution Steering
von: Teter, Alexis M. H., et al.
Veröffentlicht: (2025)
von: Teter, Alexis M. H., et al.
Veröffentlicht: (2025)
Convergence rate of random scan Coordinate Ascent Variational Inference under log-concavity
von: Lavenant, Hugo, et al.
Veröffentlicht: (2024)
von: Lavenant, Hugo, et al.
Veröffentlicht: (2024)
Efficient Online Learning in Interacting Particle Systems
von: Sharrock, Louis, et al.
Veröffentlicht: (2026)
von: Sharrock, Louis, et al.
Veröffentlicht: (2026)
Non-convex matrix sensing: Breaking the quadratic rank barrier in the sample complexity
von: Stöger, Dominik, et al.
Veröffentlicht: (2024)
von: Stöger, Dominik, et al.
Veröffentlicht: (2024)
High dimensional analysis reveals conservative sharpening and a stochastic edge of stability
von: Agarwala, Atish, et al.
Veröffentlicht: (2024)
von: Agarwala, Atish, et al.
Veröffentlicht: (2024)
Propagation of Chaos in Contextual Flow Maps
von: Chen, Shi, et al.
Veröffentlicht: (2026)
von: Chen, Shi, et al.
Veröffentlicht: (2026)
Optimal estimators of cross-partial derivatives and surrogates of functions
von: Lamboni, Matieyendou
Veröffentlicht: (2024)
von: Lamboni, Matieyendou
Veröffentlicht: (2024)
The Risk Quadrangle in Optimization: An Overview with Recent Results and Extensions
von: Grechuk, Bogdan, et al.
Veröffentlicht: (2026)
von: Grechuk, Bogdan, et al.
Veröffentlicht: (2026)
State evolution beyond first-order methods I: Rigorous predictions and finite-sample guarantees
von: Celentano, Michael, et al.
Veröffentlicht: (2025)
von: Celentano, Michael, et al.
Veröffentlicht: (2025)
High-probability sample complexities for policy evaluation with linear function approximation
von: Li, Gen, et al.
Veröffentlicht: (2023)
von: Li, Gen, et al.
Veröffentlicht: (2023)
New algorithms for sampling and diffusion models
von: Zhang, Xicheng
Veröffentlicht: (2024)
von: Zhang, Xicheng
Veröffentlicht: (2024)
Contributions to Robust and Efficient Methods for Analysis of High Dimensional Data
von: Yang, Kai
Veröffentlicht: (2025)
von: Yang, Kai
Veröffentlicht: (2025)
Off-policy estimation with adaptively collected data: the power of online learning
von: Lee, Jeonghwan, et al.
Veröffentlicht: (2024)
von: Lee, Jeonghwan, et al.
Veröffentlicht: (2024)
Decentralized Sparse Linear Regression via Gradient-Tracking: Linear Convergence and Statistical Guarantees
von: Maros, Marie, et al.
Veröffentlicht: (2022)
von: Maros, Marie, et al.
Veröffentlicht: (2022)
Sharp Rates of MMD Empirical Estimation with Power Kernels
von: Colasanto, Francesco, et al.
Veröffentlicht: (2026)
von: Colasanto, Francesco, et al.
Veröffentlicht: (2026)
Central limit theorems for vector-valued composite functionals with smoothing and applications
von: Chen, Huihui, et al.
Veröffentlicht: (2024)
von: Chen, Huihui, et al.
Veröffentlicht: (2024)
Wasserstein Distributionally Robust Optimization with Heterogeneous Data Sources
von: Rychener, Yves, et al.
Veröffentlicht: (2024)
von: Rychener, Yves, et al.
Veröffentlicht: (2024)
Optimal transport and Wasserstein distances for causal models
von: Cheridito, Patrick, et al.
Veröffentlicht: (2023)
von: Cheridito, Patrick, et al.
Veröffentlicht: (2023)
Optimal distributions for randomized unbiased estimators with an infinite horizon and an adaptive algorithm
von: Zheng, Chao, et al.
Veröffentlicht: (2023)
von: Zheng, Chao, et al.
Veröffentlicht: (2023)
Optimal transport natural gradient for statistical manifolds with continuous sample space
von: Chen, Yifan, et al.
Veröffentlicht: (2018)
von: Chen, Yifan, et al.
Veröffentlicht: (2018)
Type-II Saddles and Probabilistic Stability of Stochastic Gradient Descent
von: Ziyin, Liu, et al.
Veröffentlicht: (2023)
von: Ziyin, Liu, et al.
Veröffentlicht: (2023)
Statistical Analysis of Conditional Group Distributionally Robust Optimization with Cross-Entropy Loss
von: Guo, Zijian, et al.
Veröffentlicht: (2025)
von: Guo, Zijian, et al.
Veröffentlicht: (2025)
Tensor-on-Tensor Regression: Riemannian Optimization, Over-parameterization, Statistical-computational Gap, and Their Interplay
von: Luo, Yuetian, et al.
Veröffentlicht: (2022)
von: Luo, Yuetian, et al.
Veröffentlicht: (2022)
To spike or not to spike: the whims of the Wonham filter in the strong noise regime
von: Bernardin, Cédric, et al.
Veröffentlicht: (2022)
von: Bernardin, Cédric, et al.
Veröffentlicht: (2022)
Statistical and Algorithmic Foundations of Reinforcement Learning
von: Chi, Yuejie, et al.
Veröffentlicht: (2025)
von: Chi, Yuejie, et al.
Veröffentlicht: (2025)
Regret of exploratory policy improvement and $q$-learning
von: Tang, Wenpin, et al.
Veröffentlicht: (2024)
von: Tang, Wenpin, et al.
Veröffentlicht: (2024)
Long-time dynamics and universality of nonconvex gradient descent
von: Han, Qiyang
Veröffentlicht: (2025)
von: Han, Qiyang
Veröffentlicht: (2025)
Ähnliche Einträge
-
Optimal and instance-dependent guarantees for Markovian linear stochastic approximation
von: Mou, Wenlong, et al.
Veröffentlicht: (2021) -
Is RL fine-tuning harder than regression? A PDE learning approach for diffusion models
von: Mou, Wenlong
Veröffentlicht: (2025) -
Continuous-time reinforcement learning: ellipticity enables model-free value function approximation
von: Mou, Wenlong
Veröffentlicht: (2026) -
On Bellman equations for continuous-time policy evaluation I: discretization and approximation
von: Mou, Wenlong, et al.
Veröffentlicht: (2024) -
Statistical Inference for Linear Functionals of Online SGD in High-dimensional Linear Regression
von: Agrawalla, Bhavya, et al.
Veröffentlicht: (2023)