Deflated Dynamics Value Iteration
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Lee, Jongmin, Rakhsha, Amin, Ryu, Ernest K., Farahmand, Amir-massoud |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
PID Accelerated Temporal Difference Algorithms
von: Bedaywi, Mark, et al.
Veröffentlicht: (2024)
von: Bedaywi, Mark, et al.
Veröffentlicht: (2024)
Optimal Non-Asymptotic Rates of Value Iteration for Average-Reward Markov Decision Processes
von: Lee, Jongmin, et al.
Veröffentlicht: (2025)
von: Lee, Jongmin, et al.
Veröffentlicht: (2025)
A Truncated Newton Method for Optimal Transport
von: Kemertas, Mete, et al.
Veröffentlicht: (2025)
von: Kemertas, Mete, et al.
Veröffentlicht: (2025)
Policy Gradient Algorithms in Average-Reward Multichain MDPs
von: Lee, Jongmin, et al.
Veröffentlicht: (2026)
von: Lee, Jongmin, et al.
Veröffentlicht: (2026)
LoRA Training in the NTK Regime has No Spurious Local Minima
von: Jang, Uijeong, et al.
Veröffentlicht: (2024)
von: Jang, Uijeong, et al.
Veröffentlicht: (2024)
Rank-One Modified Value Iteration
von: Kolarijani, Arman Sharifi, et al.
Veröffentlicht: (2025)
von: Kolarijani, Arman Sharifi, et al.
Veröffentlicht: (2025)
Deflation-Free Optimal Scoring
von: Afroz, Sharmin, et al.
Veröffentlicht: (2026)
von: Afroz, Sharmin, et al.
Veröffentlicht: (2026)
Convergence Analyses of Davis-Yin Splitting via Scaled Relative Graphs
von: Lee, Jongmin, et al.
Veröffentlicht: (2022)
von: Lee, Jongmin, et al.
Veröffentlicht: (2022)
Majority of the Bests: Improving Best-of-N via Bootstrapping
von: Rakhsha, Amin, et al.
Veröffentlicht: (2025)
von: Rakhsha, Amin, et al.
Veröffentlicht: (2025)
Switching-Geometry Analysis of Deflated Q-Value Iteration
von: Lee, Donghwan
Veröffentlicht: (2026)
von: Lee, Donghwan
Veröffentlicht: (2026)
Sharpness-Aware Minimization Can Hallucinate Minimizers
von: Park, Chanwoong, et al.
Veröffentlicht: (2025)
von: Park, Chanwoong, et al.
Veröffentlicht: (2025)
On the Error-Propagation of Inexact Hotelling's Deflation for Principal Component Analysis
von: Liao, Fangshuo, et al.
Veröffentlicht: (2023)
von: Liao, Fangshuo, et al.
Veröffentlicht: (2023)
From Optimization to Control: Quasi Policy Iteration
von: Kolarijani, Mohammad Amin Sharifi, et al.
Veröffentlicht: (2023)
von: Kolarijani, Mohammad Amin Sharifi, et al.
Veröffentlicht: (2023)
Higher-Order Newton Methods with Polynomial Work per Iteration
von: Ahmadi, Amir Ali, et al.
Veröffentlicht: (2023)
von: Ahmadi, Amir Ali, et al.
Veröffentlicht: (2023)
Finite-Time Bounds for Average-Reward Fitted Q-Iteration
von: Lee, Jongmin, et al.
Veröffentlicht: (2025)
von: Lee, Jongmin, et al.
Veröffentlicht: (2025)
Pessimistic Nonlinear Least-Squares Value Iteration for Offline Reinforcement Learning
von: Di, Qiwei, et al.
Veröffentlicht: (2023)
von: Di, Qiwei, et al.
Veröffentlicht: (2023)
Differentially Private Clipped-SGD: High-Probability Convergence with Arbitrary Clipping Level
von: Khah, Saleh Vatan, et al.
Veröffentlicht: (2025)
von: Khah, Saleh Vatan, et al.
Veröffentlicht: (2025)
Truncated Variance Reduced Value Iteration
von: Jin, Yujia, et al.
Veröffentlicht: (2024)
von: Jin, Yujia, et al.
Veröffentlicht: (2024)
Provable Model-Parallel Distributed Principal Component Analysis with Parallel Deflation
von: Liao, Fangshuo, et al.
Veröffentlicht: (2025)
von: Liao, Fangshuo, et al.
Veröffentlicht: (2025)
Nesterov Flow May Travel Infinitely Long to Converge to a Minimizer
von: Ryu, Ernest K.
Veröffentlicht: (2026)
von: Ryu, Ernest K.
Veröffentlicht: (2026)
Uniqueness of DRS as the 2 Operator Resolvent-Splitting and Impossibility of 3 Operator Resolvent-Splitting
von: Ryu, Ernest K.
Veröffentlicht: (2018)
von: Ryu, Ernest K.
Veröffentlicht: (2018)
Probabilistic Iterative Hard Thresholding for Sparse Learning
von: Bergamaschi, Matteo, et al.
Veröffentlicht: (2024)
von: Bergamaschi, Matteo, et al.
Veröffentlicht: (2024)
On the Last-Iterate Convergence of Shuffling Gradient Methods
von: Liu, Zijian, et al.
Veröffentlicht: (2024)
von: Liu, Zijian, et al.
Veröffentlicht: (2024)
Accelerating Sinkhorn Algorithm with Sparse Newton Iterations
von: Tang, Xun, et al.
Veröffentlicht: (2024)
von: Tang, Xun, et al.
Veröffentlicht: (2024)
Iterative Minimax Games with Coupled Linear Constraints
von: Zhang, Huiling, et al.
Veröffentlicht: (2022)
von: Zhang, Huiling, et al.
Veröffentlicht: (2022)
Self-Supervised Learning of Iterative Solvers for Constrained Optimization
von: Lüken, Lukas, et al.
Veröffentlicht: (2024)
von: Lüken, Lukas, et al.
Veröffentlicht: (2024)
Robust Sublinear Convergence Rates for Iterative Bregman Projections
von: Peyré, Gabriel
Veröffentlicht: (2026)
von: Peyré, Gabriel
Veröffentlicht: (2026)
Gradient Descent's Last Iterate is Often (slightly) Suboptimal
von: Kornowski, Guy, et al.
Veröffentlicht: (2026)
von: Kornowski, Guy, et al.
Veröffentlicht: (2026)
Convergence Rate of the Last Iterate of Stochastic Proximal Algorithms
von: Vaidyan, Kevin Kurian Thomas, et al.
Veröffentlicht: (2026)
von: Vaidyan, Kevin Kurian Thomas, et al.
Veröffentlicht: (2026)
Revisiting the Last-Iterate Convergence of Stochastic Gradient Methods
von: Liu, Zijian, et al.
Veröffentlicht: (2023)
von: Liu, Zijian, et al.
Veröffentlicht: (2023)
Criteria and Bias of Parameterized Linear Regression under Edge of Stability Regime
von: Zhang, Peiyuan, et al.
Veröffentlicht: (2024)
von: Zhang, Peiyuan, et al.
Veröffentlicht: (2024)
Last Iterate Convergence of Incremental Methods and Applications in Continual Learning
von: Cai, Xufeng, et al.
Veröffentlicht: (2024)
von: Cai, Xufeng, et al.
Veröffentlicht: (2024)
Iterative regularization in classification via hinge loss diagonal descent
von: Apidopoulos, Vassilis, et al.
Veröffentlicht: (2022)
von: Apidopoulos, Vassilis, et al.
Veröffentlicht: (2022)
Fast Last-Iterate Convergence of SGD in the Smooth Interpolation Regime
von: Attia, Amit, et al.
Veröffentlicht: (2025)
von: Attia, Amit, et al.
Veröffentlicht: (2025)
PINS: Proximal Iterations with Sparse Newton and Sinkhorn for Optimal Transport
von: Wu, Di, et al.
Veröffentlicht: (2025)
von: Wu, Di, et al.
Veröffentlicht: (2025)
$ψ$DAG: Projected Stochastic Approximation Iteration for DAG Structure Learning
von: Ziu, Klea, et al.
Veröffentlicht: (2024)
von: Ziu, Klea, et al.
Veröffentlicht: (2024)
Distributionally Robust Optimization via Iterative Algorithms in Continuous Probability Spaces
von: Zhu, Linglingzhi, et al.
Veröffentlicht: (2024)
von: Zhu, Linglingzhi, et al.
Veröffentlicht: (2024)
Decentralized Bilevel Optimization: A Perspective from Transient Iteration Complexity
von: Kong, Boao, et al.
Veröffentlicht: (2024)
von: Kong, Boao, et al.
Veröffentlicht: (2024)
An Iteratively Reweighted Method for Sparse Optimization on Nonconvex $\ell_{p}$ Ball
von: Wang, Hao, et al.
Veröffentlicht: (2021)
von: Wang, Hao, et al.
Veröffentlicht: (2021)
Global Convergence of Iteratively Reweighted Least Squares for Robust Subspace Recovery
von: Lerman, Gilad, et al.
Veröffentlicht: (2025)
von: Lerman, Gilad, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
PID Accelerated Temporal Difference Algorithms
von: Bedaywi, Mark, et al.
Veröffentlicht: (2024) -
Optimal Non-Asymptotic Rates of Value Iteration for Average-Reward Markov Decision Processes
von: Lee, Jongmin, et al.
Veröffentlicht: (2025) -
A Truncated Newton Method for Optimal Transport
von: Kemertas, Mete, et al.
Veröffentlicht: (2025) -
Policy Gradient Algorithms in Average-Reward Multichain MDPs
von: Lee, Jongmin, et al.
Veröffentlicht: (2026) -
LoRA Training in the NTK Regime has No Spurious Local Minima
von: Jang, Uijeong, et al.
Veröffentlicht: (2024)