On Gaussian approximation for entropy-regularized Q-learning with function approximation
Fuente:
arXiv
Saved in:
| Main Authors: | Rubtsov, Artemy, Singh, Rahul, Moulines, Eric, Naumov, Alexey, Samsonov, Sergey |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Gaussian Approximation for Asynchronous Q-learning
by: Rubtsov, Artemy, et al.
Published: (2026)
by: Rubtsov, Artemy, et al.
Published: (2026)
Gaussian Approximation for Two-Timescale Linear Stochastic Approximation
by: Butyrin, Bogdan, et al.
Published: (2025)
by: Butyrin, Bogdan, et al.
Published: (2025)
A note on concentration inequalities for the overlapped batch mean variance estimators for Markov chains
by: Moulines, Eric, et al.
Published: (2025)
by: Moulines, Eric, et al.
Published: (2025)
Improved High-Probability Bounds for the Temporal Difference Learning Algorithm via Exponential Stability
by: Samsonov, Sergey, et al.
Published: (2023)
by: Samsonov, Sergey, et al.
Published: (2023)
Statistical inference for Linear Stochastic Approximation with Markovian Noise
by: Samsonov, Sergey, et al.
Published: (2025)
by: Samsonov, Sergey, et al.
Published: (2025)
Rates of convergence for density estimation with generative adversarial networks
by: Puchkin, Nikita, et al.
Published: (2021)
by: Puchkin, Nikita, et al.
Published: (2021)
Statistical analysis of Inverse Entropy-regularized Reinforcement Learning
by: Belomestny, Denis, et al.
Published: (2025)
by: Belomestny, Denis, et al.
Published: (2025)
First Order Methods with Markovian Noise: from Acceleration to Variational Inequalities
by: Beznosikov, Aleksandr, et al.
Published: (2023)
by: Beznosikov, Aleksandr, et al.
Published: (2023)
Rosenthal-type inequalities for linear statistics of Markov chains
by: Durmus, Alain, et al.
Published: (2023)
by: Durmus, Alain, et al.
Published: (2023)
Gaussian Approximation and Multiplier Bootstrap for Polyak-Ruppert Averaged Linear Stochastic Approximation with Applications to TD Learning
by: Samsonov, Sergey, et al.
Published: (2024)
by: Samsonov, Sergey, et al.
Published: (2024)
Nonasymptotic Analysis of Stochastic Gradient Descent with the Richardson-Romberg Extrapolation
by: Sheshukova, Marina, et al.
Published: (2024)
by: Sheshukova, Marina, et al.
Published: (2024)
SCAFFLSA: Taming Heterogeneity in Federated Linear Stochastic Approximation and TD Learning
by: Mangold, Paul, et al.
Published: (2024)
by: Mangold, Paul, et al.
Published: (2024)
Gaussian Approximation and Multiplier Bootstrap for Stochastic Gradient Descent
by: Sheshukova, Marina, et al.
Published: (2025)
by: Sheshukova, Marina, et al.
Published: (2025)
High-Order Error Bounds for Markovian LSA with Richardson-Romberg Extrapolation
by: Levin, Ilya, et al.
Published: (2025)
by: Levin, Ilya, et al.
Published: (2025)
Gaussian Approximation and Multiplier Bootstrap for Federated Linear Stochastic Approximation
by: Levin, Ilya, et al.
Published: (2026)
by: Levin, Ilya, et al.
Published: (2026)
UVIP: Model-Free Approach to Evaluate Reinforcement Learning Algorithms
by: Belomestny, Denis, et al.
Published: (2021)
by: Belomestny, Denis, et al.
Published: (2021)
Improved Central Limit Theorem and Bootstrap Approximations for Linear Stochastic Approximation
by: Butyrin, Bogdan, et al.
Published: (2025)
by: Butyrin, Bogdan, et al.
Published: (2025)
Queuing dynamics of asynchronous Federated Learning
by: Leconte, Louis, et al.
Published: (2024)
by: Leconte, Louis, et al.
Published: (2024)
On the Upper Bounds for the Matrix Spectral Norm
by: Naumov, Alexey, et al.
Published: (2025)
by: Naumov, Alexey, et al.
Published: (2025)
Improving GFlowNets with Monte Carlo Tree Search
by: Morozov, Nikita, et al.
Published: (2024)
by: Morozov, Nikita, et al.
Published: (2024)
Theoretical guarantees for neural control variates in MCMC
by: Belomestny, Denis, et al.
Published: (2023)
by: Belomestny, Denis, et al.
Published: (2023)
Refined Analysis of Federated Averaging and Federated Richardson-Romberg
by: Mangold, Paul, et al.
Published: (2024)
by: Mangold, Paul, et al.
Published: (2024)
Regret and Sample Complexity of Online Q-Learning via Concentration of Stochastic Approximation with Time-Inhomogeneous Markov Chains
by: Singh, Rahul, et al.
Published: (2026)
by: Singh, Rahul, et al.
Published: (2026)
Likelihood approximations via Gaussian approximate inference
by: Bui, Thang D.
Published: (2024)
by: Bui, Thang D.
Published: (2024)
Adaptive Destruction Processes for Diffusion Samplers
by: Gritsaev, Timofei, et al.
Published: (2025)
by: Gritsaev, Timofei, et al.
Published: (2025)
Proximal Point Nash Learning from Human Feedback
by: Tiapkin, Daniil, et al.
Published: (2025)
by: Tiapkin, Daniil, et al.
Published: (2025)
Demonstration-Regularized RL
by: Tiapkin, Daniil, et al.
Published: (2023)
by: Tiapkin, Daniil, et al.
Published: (2023)
Machine learning for accuracy in density functional approximations
by: Voss, Johannes
Published: (2023)
by: Voss, Johannes
Published: (2023)
Model-free Posterior Sampling via Learning Rate Randomization
by: Tiapkin, Daniil, et al.
Published: (2023)
by: Tiapkin, Daniil, et al.
Published: (2023)
Theoretical Guarantees for Variational Inference with Fixed-Variance Mixture of Gaussians
by: Huix, Tom, et al.
Published: (2024)
by: Huix, Tom, et al.
Published: (2024)
Inferring entropy production in many-body systems using nonequilibrium maximum entropy
by: Aguilera, Miguel, et al.
Published: (2025)
by: Aguilera, Miguel, et al.
Published: (2025)
Partial information decomposition: redundancy as information bottleneck
by: Kolchinsky, Artemy
Published: (2024)
by: Kolchinsky, Artemy
Published: (2024)
Multi-agent imitation learning with function approximation: Linear Markov games and beyond
by: Viano, Luca, et al.
Published: (2026)
by: Viano, Luca, et al.
Published: (2026)
Inexact subgradient methods for semialgebraic functions
by: Bolte, Jérôme, et al.
Published: (2024)
by: Bolte, Jérôme, et al.
Published: (2024)
Understanding the theoretical properties of projected Bellman equation, linear Q-learning, and approximate value iteration
by: Lim, Han-Dong, et al.
Published: (2025)
by: Lim, Han-Dong, et al.
Published: (2025)
Temporal-difference learning with nonlinear function approximation: lazy training and mean field regimes
by: Agazzi, Andrea, et al.
Published: (2019)
by: Agazzi, Andrea, et al.
Published: (2019)
Maximum entropy GFlowNets with soft Q-learning
by: Mohammadpour, Sobhan, et al.
Published: (2023)
by: Mohammadpour, Sobhan, et al.
Published: (2023)
Distribution learning via neural differential equations: minimal energy regularization and approximation theory
by: Marzouk, Youssef, et al.
Published: (2025)
by: Marzouk, Youssef, et al.
Published: (2025)
Sharp Gaussian approximations for Decentralized Federated Learning
by: Bonnerjee, Soham, et al.
Published: (2025)
by: Bonnerjee, Soham, et al.
Published: (2025)
Deep learning and the rate of approximation by flows
by: Cheng, Jingpu, et al.
Published: (2026)
by: Cheng, Jingpu, et al.
Published: (2026)
Similar Items
-
Gaussian Approximation for Asynchronous Q-learning
by: Rubtsov, Artemy, et al.
Published: (2026) -
Gaussian Approximation for Two-Timescale Linear Stochastic Approximation
by: Butyrin, Bogdan, et al.
Published: (2025) -
A note on concentration inequalities for the overlapped batch mean variance estimators for Markov chains
by: Moulines, Eric, et al.
Published: (2025) -
Improved High-Probability Bounds for the Temporal Difference Learning Algorithm via Exponential Stability
by: Samsonov, Sergey, et al.
Published: (2023) -
Statistical inference for Linear Stochastic Approximation with Markovian Noise
by: Samsonov, Sergey, et al.
Published: (2025)