Gauss-Newton Temporal Difference Learning with Nonlinear Function Approximation
Fuente:
arXiv
Salvato in:
| Autori principali: | Ke, Zhifa, Zhang, Junyu, Wen, Zaiwen |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2023
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
An Improved Finite-time Analysis of Temporal Difference Learning with Deep Neural Networks
di: Ke, Zhifa, et al.
Pubblicazione: (2024)
di: Ke, Zhifa, et al.
Pubblicazione: (2024)
Non-Asymptotic Global Convergence of PPO-Clip
di: Liu, Yin, et al.
Pubblicazione: (2025)
di: Liu, Yin, et al.
Pubblicazione: (2025)
Incremental Gauss-Newton Descent for Machine Learning
di: Korbit, Mikalai, et al.
Pubblicazione: (2024)
di: Korbit, Mikalai, et al.
Pubblicazione: (2024)
Error whitening: Why Gauss-Newton outperforms Newton
di: McKay, Maricela Best, et al.
Pubblicazione: (2026)
di: McKay, Maricela Best, et al.
Pubblicazione: (2026)
Gauss-Newton Natural Gradient Descent for Shape Learning
di: King, James, et al.
Pubblicazione: (2026)
di: King, James, et al.
Pubblicazione: (2026)
Federated Temporal Difference Learning with Linear Function Approximation under Environmental Heterogeneity
di: Wang, Han, et al.
Pubblicazione: (2023)
di: Wang, Han, et al.
Pubblicazione: (2023)
Regularized Gauss-Newton for Optimizing Overparameterized Neural Networks
di: Adeoye, Adeyemi D., et al.
Pubblicazione: (2024)
di: Adeoye, Adeyemi D., et al.
Pubblicazione: (2024)
Incremental Gauss--Newton Methods with Superlinear Convergence Rates
di: Zhou, Zhiling, et al.
Pubblicazione: (2024)
di: Zhou, Zhiling, et al.
Pubblicazione: (2024)
An Adaptive Stochastic Gradient Method with Non-negative Gauss-Newton Stepsizes
di: Orvieto, Antonio, et al.
Pubblicazione: (2024)
di: Orvieto, Antonio, et al.
Pubblicazione: (2024)
Non-Asymptotic Optimization and Generalization Bounds for Stochastic Gauss-Newton in Overparameterized Models
di: Cayci, Semih
Pubblicazione: (2025)
di: Cayci, Semih
Pubblicazione: (2025)
Convergence Analysis of Stochastic Gradient Descent with MCMC Estimators
di: Li, Tianyou, et al.
Pubblicazione: (2023)
di: Li, Tianyou, et al.
Pubblicazione: (2023)
Accelerating Optimization via Differentiable Stopping Time
di: Xie, Zhonglin, et al.
Pubblicazione: (2025)
di: Xie, Zhonglin, et al.
Pubblicazione: (2025)
On the Convergence of Policy Mirror Descent with Temporal Difference Evaluation
di: Liu, Jiacai, et al.
Pubblicazione: (2025)
di: Liu, Jiacai, et al.
Pubblicazione: (2025)
Policy Mirror Descent with Temporal Difference Learning: Sample Complexity under Online Markov Data
di: Li, Wenye, et al.
Pubblicazione: (2025)
di: Li, Wenye, et al.
Pubblicazione: (2025)
A Riemannian Optimization Perspective of the Gauss-Newton Method for Feedforward Neural Networks
di: Cayci, Semih
Pubblicazione: (2024)
di: Cayci, Semih
Pubblicazione: (2024)
Can Temporal-Difference and Q-Learning Learn Representation? A Mean-Field Theory
di: Zhang, Yufeng, et al.
Pubblicazione: (2020)
di: Zhang, Yufeng, et al.
Pubblicazione: (2020)
A Gauss-Newton Approach for Min-Max Optimization in Generative Adversarial Networks
di: Mishra, Neel, et al.
Pubblicazione: (2024)
di: Mishra, Neel, et al.
Pubblicazione: (2024)
RGNMR: A Gauss-Newton method for robust matrix completion with theoretical guarantees
di: Laufer, Eilon Vaknin, et al.
Pubblicazione: (2025)
di: Laufer, Eilon Vaknin, et al.
Pubblicazione: (2025)
Reinforcement Learning with Function Approximation for Non-Markov Processes
di: Kara, Ali Devran
Pubblicazione: (2026)
di: Kara, Ali Devran
Pubblicazione: (2026)
Finite-Time Decoupled Convergence in Nonlinear Two-Time-Scale Stochastic Approximation
di: Han, Yuze, et al.
Pubblicazione: (2024)
di: Han, Yuze, et al.
Pubblicazione: (2024)
Infinite-Horizon Reinforcement Learning with Multinomial Logistic Function Approximation
di: Park, Jaehyun, et al.
Pubblicazione: (2024)
di: Park, Jaehyun, et al.
Pubblicazione: (2024)
A Finite-Iteration Theory for Asynchronous Categorical Distributional Temporal-Difference Learning
di: Kaya, Ege C., et al.
Pubblicazione: (2026)
di: Kaya, Ege C., et al.
Pubblicazione: (2026)
Convergence Rate in Nonlinear Two-Time-Scale Stochastic Approximation with State (Time)-Dependence
di: Chen, Zixi, et al.
Pubblicazione: (2025)
di: Chen, Zixi, et al.
Pubblicazione: (2025)
Online Learning for Approximately-Convex Functions with Long-term Adversarial Constraints
di: Sarkar, Dhruv, et al.
Pubblicazione: (2025)
di: Sarkar, Dhruv, et al.
Pubblicazione: (2025)
A Simple Finite-Time Analysis of TD Learning with Linear Function Approximation
di: Mitra, Aritra
Pubblicazione: (2024)
di: Mitra, Aritra
Pubblicazione: (2024)
A Nearly Optimal and Low-Switching Algorithm for Reinforcement Learning with General Function Approximation
di: Zhao, Heyang, et al.
Pubblicazione: (2023)
di: Zhao, Heyang, et al.
Pubblicazione: (2023)
Reinforcement Learning from Partial Observation: Linear Function Approximation with Provable Sample Efficiency
di: Cai, Qi, et al.
Pubblicazione: (2022)
di: Cai, Qi, et al.
Pubblicazione: (2022)
An efficient primal dual semismooth Newton method for semidefinite programming
di: Deng, Zhanwang, et al.
Pubblicazione: (2025)
di: Deng, Zhanwang, et al.
Pubblicazione: (2025)
Convergence of Muon with Newton-Schulz
di: Kim, Gyu Yeol, et al.
Pubblicazione: (2026)
di: Kim, Gyu Yeol, et al.
Pubblicazione: (2026)
qNBO: quasi-Newton Meets Bilevel Optimization
di: Fang, Sheng, et al.
Pubblicazione: (2025)
di: Fang, Sheng, et al.
Pubblicazione: (2025)
Online Learning Guided Quasi-Newton Methods with Global Non-Asymptotic Convergence
di: Jiang, Ruichen, et al.
Pubblicazione: (2024)
di: Jiang, Ruichen, et al.
Pubblicazione: (2024)
Higher-Order Newton Methods with Polynomial Work per Iteration
di: Ahmadi, Amir Ali, et al.
Pubblicazione: (2023)
di: Ahmadi, Amir Ali, et al.
Pubblicazione: (2023)
The Collusion of Memory and Nonlinearity in Stochastic Approximation With Constant Stepsize
di: Huo, Dongyan, et al.
Pubblicazione: (2024)
di: Huo, Dongyan, et al.
Pubblicazione: (2024)
Universal Approximation Power of Deep Residual Neural Networks via Nonlinear Control Theory
di: Tabuada, Paulo, et al.
Pubblicazione: (2020)
di: Tabuada, Paulo, et al.
Pubblicazione: (2020)
A Finite-Time Analysis of TD Learning with Linear Function Approximation without Projections or Strong Convexity
di: Lee, Wei-Cheng, et al.
Pubblicazione: (2025)
di: Lee, Wei-Cheng, et al.
Pubblicazione: (2025)
Stochastic Newton Proximal Extragradient Method
di: Jiang, Ruichen, et al.
Pubblicazione: (2024)
di: Jiang, Ruichen, et al.
Pubblicazione: (2024)
Improving Stochastic Cubic Newton with Momentum
di: Chayti, El Mahdi, et al.
Pubblicazione: (2024)
di: Chayti, El Mahdi, et al.
Pubblicazione: (2024)
General Loss Functions Lead to (Approximate) Interpolation in High Dimensions
di: Lai, Kuo-Wei, et al.
Pubblicazione: (2023)
di: Lai, Kuo-Wei, et al.
Pubblicazione: (2023)
An Augmented Lagrangian Primal-Dual Semismooth Newton Method for Multi-Block Composite Optimization
di: Deng, Zhanwang, et al.
Pubblicazione: (2023)
di: Deng, Zhanwang, et al.
Pubblicazione: (2023)
Convergence of Policy Mirror Descent Beyond Compatible Function Approximation
di: Sherman, Uri, et al.
Pubblicazione: (2025)
di: Sherman, Uri, et al.
Pubblicazione: (2025)
Documenti analoghi
-
An Improved Finite-time Analysis of Temporal Difference Learning with Deep Neural Networks
di: Ke, Zhifa, et al.
Pubblicazione: (2024) -
Non-Asymptotic Global Convergence of PPO-Clip
di: Liu, Yin, et al.
Pubblicazione: (2025) -
Incremental Gauss-Newton Descent for Machine Learning
di: Korbit, Mikalai, et al.
Pubblicazione: (2024) -
Error whitening: Why Gauss-Newton outperforms Newton
di: McKay, Maricela Best, et al.
Pubblicazione: (2026) -
Gauss-Newton Natural Gradient Descent for Shape Learning
di: King, James, et al.
Pubblicazione: (2026)