Convergence of Batch Asynchronous Stochastic Approximation With Applications to Reinforcement Learning
Fuente:
arXiv
Guardado en:
| Autores principales: | Karandikar, Rajeeva L., Vidyasagar, M. |
|---|---|
| Formato: | Preprint |
| Publicado: |
2021
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Convergence Rates for Stochastic Approximation: Biased Noise with Unbounded Variance, and Applications
por: Karandikar, Rajeeva L., et al.
Publicado: (2023)
por: Karandikar, Rajeeva L., et al.
Publicado: (2023)
Revisiting Stochastic Approximation and Stochastic Gradient Descent
por: Karandikar, Rajeeva Laxman, et al.
Publicado: (2025)
por: Karandikar, Rajeeva Laxman, et al.
Publicado: (2025)
Gaussian Approximation and Multiplier Bootstrap for Stochastic Gradient Descent
por: Sheshukova, Marina, et al.
Publicado: (2025)
por: Sheshukova, Marina, et al.
Publicado: (2025)
Learning to reflect: A unifying approach for data-driven stochastic control strategies
por: Christensen, Sören, et al.
Publicado: (2021)
por: Christensen, Sören, et al.
Publicado: (2021)
Gaussian Approximation and Multiplier Bootstrap for Polyak-Ruppert Averaged Linear Stochastic Approximation with Applications to TD Learning
por: Samsonov, Sergey, et al.
Publicado: (2024)
por: Samsonov, Sergey, et al.
Publicado: (2024)
A geometric ensemble method for Bayesian inference
por: Popov, Andrey A
Publicado: (2025)
por: Popov, Andrey A
Publicado: (2025)
Improved Central Limit Theorem and Bootstrap Approximations for Linear Stochastic Approximation
por: Butyrin, Bogdan, et al.
Publicado: (2025)
por: Butyrin, Bogdan, et al.
Publicado: (2025)
Asynchronous Stochastic Approximation with Applications to Average-Reward Reinforcement Learning
por: Yu, Huizhen, et al.
Publicado: (2024)
por: Yu, Huizhen, et al.
Publicado: (2024)
Asymptotics of generalized Pólya urns with non-linear feedback
por: Gottfried, Thomas, et al.
Publicado: (2023)
por: Gottfried, Thomas, et al.
Publicado: (2023)
A Sequential Testing Problem with Signal Control
por: Campbell, Steven, et al.
Publicado: (2025)
por: Campbell, Steven, et al.
Publicado: (2025)
Optimality of a barrier strategy in a spectrally negative Lévy model with a level-dependent intensity of bankruptcy
por: Mata, Dante, et al.
Publicado: (2024)
por: Mata, Dante, et al.
Publicado: (2024)
Gaussian Approximation for Two-Timescale Linear Stochastic Approximation
por: Butyrin, Bogdan, et al.
Publicado: (2025)
por: Butyrin, Bogdan, et al.
Publicado: (2025)
Adaptive Multilevel Stochastic Approximation of the Value-at-Risk
por: Crépey, Stéphane, et al.
Publicado: (2024)
por: Crépey, Stéphane, et al.
Publicado: (2024)
Sample Average Approximation for Stochastic Programming with Equality Constraints
por: Lew, Thomas, et al.
Publicado: (2022)
por: Lew, Thomas, et al.
Publicado: (2022)
On the Rate of Gaussian Approximation for Linear Regression Problems
por: Khusainov, Marat, et al.
Publicado: (2025)
por: Khusainov, Marat, et al.
Publicado: (2025)
Local regression on path spaces with signature metrics
por: Bayer, Christian, et al.
Publicado: (2025)
por: Bayer, Christian, et al.
Publicado: (2025)
Asynchronous Averaging on Dynamic Graphs with Selective Neighborhood Contraction
por: Li, Hsin-Lun
Publicado: (2025)
por: Li, Hsin-Lun
Publicado: (2025)
A Multilevel Stochastic Approximation Algorithm for Value-at-Risk and Expected Shortfall Estimation
por: Crépey, Stéphane, et al.
Publicado: (2023)
por: Crépey, Stéphane, et al.
Publicado: (2023)
The ODE Method for Asymptotic Statistics in Stochastic Approximation and Reinforcement Learning
por: Borkar, Vivek, et al.
Publicado: (2021)
por: Borkar, Vivek, et al.
Publicado: (2021)
Markov approximation for controlled Hawkes Jump-Diffusions with general kernels
por: Khabou, Mahmoud, et al.
Publicado: (2025)
por: Khabou, Mahmoud, et al.
Publicado: (2025)
Approximate Transitivity of Young Translation on Rough Paths
por: Bellingeri, Carlo, et al.
Publicado: (2026)
por: Bellingeri, Carlo, et al.
Publicado: (2026)
Stochastic Volterra equations with random functional coefficients in Banach spaces
por: Kalinin, Alexander
Publicado: (2026)
por: Kalinin, Alexander
Publicado: (2026)
A measure-valued HJB perspective on Bayesian optimal adaptive control
por: Cox, Alexander M. G., et al.
Publicado: (2025)
por: Cox, Alexander M. G., et al.
Publicado: (2025)
Finite-Time Analysis of Projected Two-Time-Scale Stochastic Approximation
por: Bai, Yitao, et al.
Publicado: (2026)
por: Bai, Yitao, et al.
Publicado: (2026)
Optimistic Training and Convergence of Q-Learning -- Extended Version
por: Mehta, Prashant, et al.
Publicado: (2026)
por: Mehta, Prashant, et al.
Publicado: (2026)
Convergence of Riemannian Stochastic Gradient Descents: Varying Batch Sizes And Nonstandard Batch Forming
por: Wu, Hao
Publicado: (2026)
por: Wu, Hao
Publicado: (2026)
Mean-field games with rough common noise: the linear-quadratic case
por: Friz, Peter K., et al.
Publicado: (2026)
por: Friz, Peter K., et al.
Publicado: (2026)
Convergence of the extended Kalman filter with small and state-dependent noise
por: Njiasse, Ibrahim Mbouandi, et al.
Publicado: (2025)
por: Njiasse, Ibrahim Mbouandi, et al.
Publicado: (2025)
Distribution-free Measures of Association based on Optimal Transport
por: Deb, Nabarun, et al.
Publicado: (2024)
por: Deb, Nabarun, et al.
Publicado: (2024)
Scaled quadratic variation for controlled rough paths and parameter estimation of fractional diffusions
por: Leahy, James-Michael, et al.
Publicado: (2024)
por: Leahy, James-Michael, et al.
Publicado: (2024)
Sample Complexity of Policy Gradient for Log-Growth Control
por: Pan, Qiuhua, et al.
Publicado: (2026)
por: Pan, Qiuhua, et al.
Publicado: (2026)
On the Weak Convergence of the Function-Indexed Sequential Empirical Process and its Smoothed Analogue under Nonstationarity
por: Scholze, Florian Alexander, et al.
Publicado: (2024)
por: Scholze, Florian Alexander, et al.
Publicado: (2024)
Continuous-time Risk-sensitive Reinforcement Learning via Quadratic Variation Penalty
por: Jia, Yanwei
Publicado: (2024)
por: Jia, Yanwei
Publicado: (2024)
Inverting Poisson-Laguerre tessellations
por: van der Jagt, Thomas, et al.
Publicado: (2026)
por: van der Jagt, Thomas, et al.
Publicado: (2026)
Statistical inference for Linear Stochastic Approximation with Markovian Noise
por: Samsonov, Sergey, et al.
Publicado: (2025)
por: Samsonov, Sergey, et al.
Publicado: (2025)
Asymptotic theory and statistical inference for the samples problems with heavy-tailed data using the functional empirical process
por: Camara, Abdoulaye, et al.
Publicado: (2025)
por: Camara, Abdoulaye, et al.
Publicado: (2025)
Stochastic Control with Signatures
por: Bank, P., et al.
Publicado: (2024)
por: Bank, P., et al.
Publicado: (2024)
Pontryagin Maximum Principle for rough stochastic systems and pathwise stochastic control
por: Horst, Ulrich, et al.
Publicado: (2025)
por: Horst, Ulrich, et al.
Publicado: (2025)
Signature Methods in Stochastic Portfolio Theory
por: Cuchiero, Christa, et al.
Publicado: (2023)
por: Cuchiero, Christa, et al.
Publicado: (2023)
Sharp Convergence Rates of Empirical Unbalanced Optimal Transport for Spatio-Temporal Point Processes
por: Struleva, Marina, et al.
Publicado: (2025)
por: Struleva, Marina, et al.
Publicado: (2025)
Ejemplares similares
-
Convergence Rates for Stochastic Approximation: Biased Noise with Unbounded Variance, and Applications
por: Karandikar, Rajeeva L., et al.
Publicado: (2023) -
Revisiting Stochastic Approximation and Stochastic Gradient Descent
por: Karandikar, Rajeeva Laxman, et al.
Publicado: (2025) -
Gaussian Approximation and Multiplier Bootstrap for Stochastic Gradient Descent
por: Sheshukova, Marina, et al.
Publicado: (2025) -
Learning to reflect: A unifying approach for data-driven stochastic control strategies
por: Christensen, Sören, et al.
Publicado: (2021) -
Gaussian Approximation and Multiplier Bootstrap for Polyak-Ruppert Averaged Linear Stochastic Approximation with Applications to TD Learning
por: Samsonov, Sergey, et al.
Publicado: (2024)