Stabilizing Temporal Difference Learning via Implicit Stochastic Recursion
Fuente:
arXiv
Saved in:
| Main Authors: | Kim, Hwanwoo, Toulis, Panos, Laber, Eric |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Implicit Updates for Average-Reward Temporal Difference Learning
by: Kim, Hwanwoo, et al.
Published: (2025)
by: Kim, Hwanwoo, et al.
Published: (2025)
Implicit Q-Learning and SARSA: Liberating Policy Control from Step-Size Calibration
by: Kim, Hwanwoo, et al.
Published: (2026)
by: Kim, Hwanwoo, et al.
Published: (2026)
Asymptotic Validity and Finite-Sample Properties of Approximate Randomization Tests
by: Toulis, Panos
Published: (2019)
by: Toulis, Panos
Published: (2019)
Generalized Gaussian Temporal Difference Error for Uncertainty-aware Reinforcement Learning
by: Kim, Seyeon, et al.
Published: (2024)
by: Kim, Seyeon, et al.
Published: (2024)
Characterizing Dynamical Stability of Stochastic Gradient Descent in Overparameterized Learning
by: Chemnitz, Dennis, et al.
Published: (2024)
by: Chemnitz, Dennis, et al.
Published: (2024)
One-Shot Generative Flows: Existence and Obstructions
by: Tsimpos, Panos, et al.
Published: (2026)
by: Tsimpos, Panos, et al.
Published: (2026)
Stochastic Operator Network: A Stochastic Maximum Principle Based Approach to Operator Learning
by: Bausback, Ryan, et al.
Published: (2025)
by: Bausback, Ryan, et al.
Published: (2025)
Diffusion Processes on Implicit Manifolds
by: Kawasaki-Borruat, Victor, et al.
Published: (2026)
by: Kawasaki-Borruat, Victor, et al.
Published: (2026)
Adaptive Learning via Off-Model Training and Importance Sampling for Fully Non-Markovian Optimal Stochastic Control. Complete version
by: Leão, Dorival, et al.
Published: (2026)
by: Leão, Dorival, et al.
Published: (2026)
Controlling the Flow: Stability and Convergence for Stochastic Gradient Descent with Decaying Regularization
by: Kassing, Sebastian, et al.
Published: (2025)
by: Kassing, Sebastian, et al.
Published: (2025)
Stochastic Differential Equations models for Least-Squares Stochastic Gradient Descent
by: Schertzer, Adrien, et al.
Published: (2024)
by: Schertzer, Adrien, et al.
Published: (2024)
Algorithmic Stability of Stochastic Gradient Descent with Momentum under Heavy-Tailed Noise
by: Dang, Thanh, et al.
Published: (2025)
by: Dang, Thanh, et al.
Published: (2025)
Implicit Compressibility of Overparametrized Neural Networks Trained with Heavy-Tailed SGD
by: Wan, Yijun, et al.
Published: (2023)
by: Wan, Yijun, et al.
Published: (2023)
ML-assisted Randomization Tests for Detecting Treatment Effects in A/B Experiments
by: Guo, Wenxuan, et al.
Published: (2025)
by: Guo, Wenxuan, et al.
Published: (2025)
Recursive Maximum Likelihood Estimation for Interacting Particle Systems using Virtual Particles
by: Sharrock, Louis, et al.
Published: (2026)
by: Sharrock, Louis, et al.
Published: (2026)
Dynamic Treatment on Networks
by: Nar, Bengusu, et al.
Published: (2026)
by: Nar, Bengusu, et al.
Published: (2026)
Enhancing Gaussian Process Surrogates for Optimization and Posterior Approximation via Random Exploration
by: Kim, Hwanwoo, et al.
Published: (2024)
by: Kim, Hwanwoo, et al.
Published: (2024)
Random-Bridges as Stochastic Transports for Generative Models
by: Goria, Stefano, et al.
Published: (2025)
by: Goria, Stefano, et al.
Published: (2025)
Decentralized Proximal Stochastic Gradient Langevin Dynamics
by: Islam, Mohammad Rafiqul, et al.
Published: (2026)
by: Islam, Mohammad Rafiqul, et al.
Published: (2026)
Limit Theorems for Stochastic Gradient Descent with Infinite Variance
by: Blanchet, Jose, et al.
Published: (2024)
by: Blanchet, Jose, et al.
Published: (2024)
Partially Stochastic Infinitely Deep Bayesian Neural Networks
by: Calvo-Ordonez, Sergio, et al.
Published: (2024)
by: Calvo-Ordonez, Sergio, et al.
Published: (2024)
Flow Matching: Markov Kernels, Stochastic Processes and Transport Plans
by: Wald, Christian, et al.
Published: (2025)
by: Wald, Christian, et al.
Published: (2025)
Approximation to Deep Q-Network by Stochastic Delay Differential Equations
by: Lu, Jianya, et al.
Published: (2025)
by: Lu, Jianya, et al.
Published: (2025)
Stochastic Scaling Limits and Synchronization by Noise in Deep Transformer Models
by: Agazzi, Andrea, et al.
Published: (2026)
by: Agazzi, Andrea, et al.
Published: (2026)
Efficiency of Parallel and Restart Exploration Strategies in Model Free Stochastic Simulations
by: Garcia, Ernesto, et al.
Published: (2025)
by: Garcia, Ernesto, et al.
Published: (2025)
Convergence, Sticking and Escape: Stochastic Dynamics Near Critical Points in SGD
by: Dudukalov, Dmitry, et al.
Published: (2025)
by: Dudukalov, Dmitry, et al.
Published: (2025)
Towards Continuous-Time Approximations for Stochastic Gradient Descent without Replacement
by: Perko, Stefan
Published: (2025)
by: Perko, Stefan
Published: (2025)
Stochastic Port-Hamiltonian Neural Networks: Universal Approximation with Passivity Guarantees
by: Di Persio, Luca, et al.
Published: (2026)
by: Di Persio, Luca, et al.
Published: (2026)
Exact Gradients for Stochastic Spiking Neural Networks Driven by Rough Signals
by: Holberg, Christian, et al.
Published: (2024)
by: Holberg, Christian, et al.
Published: (2024)
Non-Asymptotic Convergence of Stochastic Iterative Algorithms: A Lyapunov Framework
by: Chen, Zaiwei, et al.
Published: (2026)
by: Chen, Zaiwei, et al.
Published: (2026)
The Proximal Robbins-Monro Method
by: Toulis, Panos, et al.
Published: (2015)
by: Toulis, Panos, et al.
Published: (2015)
Accelerating Distributed Stochastic Optimization via Self-Repellent Random Walks
by: Hu, Jie, et al.
Published: (2024)
by: Hu, Jie, et al.
Published: (2024)
Permutation-Invariant Spectral Learning via Dyson Diffusion
by: Schwarz, Tassilo, et al.
Published: (2025)
by: Schwarz, Tassilo, et al.
Published: (2025)
A General-Purpose Theorem for High-Probability Bounds of Stochastic Approximation with Polyak Averaging
by: Khodadadian, Sajad, et al.
Published: (2025)
by: Khodadadian, Sajad, et al.
Published: (2025)
Steady-State Behavior of Constant-Stepsize Stochastic Approximation: Gaussian Approximation and Tail Bounds
by: Wang, Zedong, et al.
Published: (2026)
by: Wang, Zedong, et al.
Published: (2026)
Learning the Infinitesimal Generator of Stochastic Diffusion Processes
by: Kostic, Vladimir R., et al.
Published: (2024)
by: Kostic, Vladimir R., et al.
Published: (2024)
Stochastic Interpolants: A Unifying Framework for Flows and Diffusions
by: Albergo, Michael S., et al.
Published: (2023)
by: Albergo, Michael S., et al.
Published: (2023)
Sampling via Stochastic Interpolants by Langevin-based Velocity and Initialization Estimation in Flow ODEs
by: Duan, Chenguang, et al.
Published: (2026)
by: Duan, Chenguang, et al.
Published: (2026)
Type-II Saddles and Probabilistic Stability of Stochastic Gradient Descent
by: Ziyin, Liu, et al.
Published: (2023)
by: Ziyin, Liu, et al.
Published: (2023)
Multivariate Gaussian Approximation for Random Forest via Region-based Stabilization
by: Shi, Zhaoyang, et al.
Published: (2024)
by: Shi, Zhaoyang, et al.
Published: (2024)
Similar Items
-
Implicit Updates for Average-Reward Temporal Difference Learning
by: Kim, Hwanwoo, et al.
Published: (2025) -
Implicit Q-Learning and SARSA: Liberating Policy Control from Step-Size Calibration
by: Kim, Hwanwoo, et al.
Published: (2026) -
Asymptotic Validity and Finite-Sample Properties of Approximate Randomization Tests
by: Toulis, Panos
Published: (2019) -
Generalized Gaussian Temporal Difference Error for Uncertainty-aware Reinforcement Learning
by: Kim, Seyeon, et al.
Published: (2024) -
Characterizing Dynamical Stability of Stochastic Gradient Descent in Overparameterized Learning
by: Chemnitz, Dennis, et al.
Published: (2024)