A primal-dual perspective for distributed TD-learning
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Lim, Han-Dong, Lee, Donghwan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2023
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Auto-conditioned primal-dual hybrid gradient method and alternating direction method of multipliers
von: Lan, Guanghui, et al.
Veröffentlicht: (2024)
von: Lan, Guanghui, et al.
Veröffentlicht: (2024)
Zeroth-Order primal-dual Alternating Projection Gradient Algorithms for Nonconvex Minimax Problems with Coupled linear Constraints
von: Zhang, Huiling, et al.
Veröffentlicht: (2024)
von: Zhang, Huiling, et al.
Veröffentlicht: (2024)
Convergence and sample complexity of natural policy gradient primal-dual methods for constrained MDPs
von: Ding, Dongsheng, et al.
Veröffentlicht: (2022)
von: Ding, Dongsheng, et al.
Veröffentlicht: (2022)
A Finite-Time Analysis of TD Learning with Linear Function Approximation without Projections or Strong Convexity
von: Lee, Wei-Cheng, et al.
Veröffentlicht: (2025)
von: Lee, Wei-Cheng, et al.
Veröffentlicht: (2025)
Analysis of Off-Policy Multi-Step TD-Learning with Linear Function Approximation
von: Lee, Donghwan
Veröffentlicht: (2024)
von: Lee, Donghwan
Veröffentlicht: (2024)
Two-timescale Extragradient for Finding Local Minimax Points
von: Chae, Jiseok, et al.
Veröffentlicht: (2023)
von: Chae, Jiseok, et al.
Veröffentlicht: (2023)
Stochastic Extragradient with Flip-Flop Shuffling & Anchoring: Provable Improvements
von: Chae, Jiseok, et al.
Veröffentlicht: (2024)
von: Chae, Jiseok, et al.
Veröffentlicht: (2024)
Revisiting Inexact Fixed-Point Iterations for Min-Max Problems: Stochasticity and Structured Nonconvexity
von: Alacaoglu, Ahmet, et al.
Veröffentlicht: (2024)
von: Alacaoglu, Ahmet, et al.
Veröffentlicht: (2024)
A Simple Finite-Time Analysis of TD Learning with Linear Function Approximation
von: Mitra, Aritra
Veröffentlicht: (2024)
von: Mitra, Aritra
Veröffentlicht: (2024)
SCAFFLSA: Taming Heterogeneity in Federated Linear Stochastic Approximation and TD Learning
von: Mangold, Paul, et al.
Veröffentlicht: (2024)
von: Mangold, Paul, et al.
Veröffentlicht: (2024)
A brief note on learning problem with global perspectives
von: Befekadu, Getachew K.
Veröffentlicht: (2026)
von: Befekadu, Getachew K.
Veröffentlicht: (2026)
Adversarially-Robust TD Learning with Markovian Data: Finite-Time Rates and Fundamental Limits
von: Maity, Sreejeet, et al.
Veröffentlicht: (2025)
von: Maity, Sreejeet, et al.
Veröffentlicht: (2025)
A continuous perspective on the inertial corrected primal-dual proximal splitting
von: Luo, Hao
Veröffentlicht: (2024)
von: Luo, Hao
Veröffentlicht: (2024)
Rates of Convergence in the Central Limit Theorem for Markov Chains, with an Application to TD Learning
von: Srikant, R.
Veröffentlicht: (2024)
von: Srikant, R.
Veröffentlicht: (2024)
Controllable Machine Unlearning via Gradient Pivoting
von: Hwang, Youngsik, et al.
Veröffentlicht: (2025)
von: Hwang, Youngsik, et al.
Veröffentlicht: (2025)
Analysis of Off-Policy $n$-Step TD-Learning with Linear Function Approximation
von: Lim, Han-Dong, et al.
Veröffentlicht: (2025)
von: Lim, Han-Dong, et al.
Veröffentlicht: (2025)
DGSAM: Domain Generalization via Individual Sharpness-Aware Minimization
von: Song, Youngjun, et al.
Veröffentlicht: (2025)
von: Song, Youngjun, et al.
Veröffentlicht: (2025)
Meta-reinforcement learning with minimum attention
von: Gupta, Shashank, et al.
Veröffentlicht: (2025)
von: Gupta, Shashank, et al.
Veröffentlicht: (2025)
Beating level-set methods for 3D seismic data interpolation: a primal-dual alternating approach
von: Kumar, Rajiv, et al.
Veröffentlicht: (2016)
von: Kumar, Rajiv, et al.
Veröffentlicht: (2016)
Polygonal Unadjusted Langevin Algorithms: Creating stable and efficient adaptive algorithms for neural networks
von: Lim, Dong-Young, et al.
Veröffentlicht: (2021)
von: Lim, Dong-Young, et al.
Veröffentlicht: (2021)
Primal-dual algorithm for contextual stochastic combinatorial optimization
von: Bouvier, Louis, et al.
Veröffentlicht: (2025)
von: Bouvier, Louis, et al.
Veröffentlicht: (2025)
A new perspective on low-rank optimization
von: Bertsimas, Dimitris, et al.
Veröffentlicht: (2021)
von: Bertsimas, Dimitris, et al.
Veröffentlicht: (2021)
An efficient primal dual semismooth Newton method for semidefinite programming
von: Deng, Zhanwang, et al.
Veröffentlicht: (2025)
von: Deng, Zhanwang, et al.
Veröffentlicht: (2025)
On Some Geometric Behavior of Value Iteration on the Orthant: Switching System Perspective
von: Lee, Donghwan
Veröffentlicht: (2023)
von: Lee, Donghwan
Veröffentlicht: (2023)
Novel clustered federated learning based on local loss
von: Gu, Endong, et al.
Veröffentlicht: (2024)
von: Gu, Endong, et al.
Veröffentlicht: (2024)
Lossless Convexification and Duality
von: Lee, Donghwan
Veröffentlicht: (2021)
von: Lee, Donghwan
Veröffentlicht: (2021)
A finite time analysis of distributed Q-learning
von: Lim, Han-Dong, et al.
Veröffentlicht: (2024)
von: Lim, Han-Dong, et al.
Veröffentlicht: (2024)
A unified perspective on fine-tuning and sampling with diffusion and flow models
von: Domingo-Enrich, Carles, et al.
Veröffentlicht: (2026)
von: Domingo-Enrich, Carles, et al.
Veröffentlicht: (2026)
A universal convergence theorem for primal-dual penalty and augmented Lagrangian methods
von: Dolgopolik, M. V.
Veröffentlicht: (2024)
von: Dolgopolik, M. V.
Veröffentlicht: (2024)
Regret Analysis: a control perspective
von: Gibson, Travis E., et al.
Veröffentlicht: (2025)
von: Gibson, Travis E., et al.
Veröffentlicht: (2025)
Convergence analysis of primal-dual augmented Lagrangian methods and duality theory
von: Dolgopolik, M. V.
Veröffentlicht: (2024)
von: Dolgopolik, M. V.
Veröffentlicht: (2024)
Sporadic Gradient Tracking over Directed Graphs: A Theoretical Perspective on Decentralized Federated Learning
von: Zehtabi, Shahryar, et al.
Veröffentlicht: (2026)
von: Zehtabi, Shahryar, et al.
Veröffentlicht: (2026)
An optimal control perspective on diffusion-based generative modeling
von: Berner, Julius, et al.
Veröffentlicht: (2022)
von: Berner, Julius, et al.
Veröffentlicht: (2022)
On characterizing optimal learning trajectories in a class of learning problems
von: Befekadu, Getachew K
Veröffentlicht: (2025)
von: Befekadu, Getachew K
Veröffentlicht: (2025)
A survey on secure decentralized optimization and learning
von: Liu, Changxin, et al.
Veröffentlicht: (2024)
von: Liu, Changxin, et al.
Veröffentlicht: (2024)
A primal-dual backward reflected forward splitting algorithm for structured monotone inclusions
von: Bang, Vu Cong, et al.
Veröffentlicht: (2024)
von: Bang, Vu Cong, et al.
Veröffentlicht: (2024)
SSNCVX: A primal-dual semismooth Newton method for convex composite optimization problem
von: Deng, Zhanwang, et al.
Veröffentlicht: (2025)
von: Deng, Zhanwang, et al.
Veröffentlicht: (2025)
Meta-learning for sample-efficient Bayesian optimisation of fed-batch processes
von: Langdon, Becky, et al.
Veröffentlicht: (2026)
von: Langdon, Becky, et al.
Veröffentlicht: (2026)
On the influence of dependent features in classification problems: a game-theoretic perspective
von: Davila-Pena, Laura, et al.
Veröffentlicht: (2024)
von: Davila-Pena, Laura, et al.
Veröffentlicht: (2024)
Flatness-Aware Stochastic Gradient Langevin Dynamics
von: Bruno, Stefano, et al.
Veröffentlicht: (2025)
von: Bruno, Stefano, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Auto-conditioned primal-dual hybrid gradient method and alternating direction method of multipliers
von: Lan, Guanghui, et al.
Veröffentlicht: (2024) -
Zeroth-Order primal-dual Alternating Projection Gradient Algorithms for Nonconvex Minimax Problems with Coupled linear Constraints
von: Zhang, Huiling, et al.
Veröffentlicht: (2024) -
Convergence and sample complexity of natural policy gradient primal-dual methods for constrained MDPs
von: Ding, Dongsheng, et al.
Veröffentlicht: (2022) -
A Finite-Time Analysis of TD Learning with Linear Function Approximation without Projections or Strong Convexity
von: Lee, Wei-Cheng, et al.
Veröffentlicht: (2025) -
Analysis of Off-Policy Multi-Step TD-Learning with Linear Function Approximation
von: Lee, Donghwan
Veröffentlicht: (2024)