Understanding the Curse of Unrolling
Fuente:
arXiv
Guardado en:
| Autores principales: | Mehmood, Sheheryar, Knoll, Florian, Ochs, Peter |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Fixed-Point Automatic Differentiation of Forward--Backward Splitting Algorithms for Partly Smooth Functions
por: Mehmood, Sheheryar, et al.
Publicado: (2022)
por: Mehmood, Sheheryar, et al.
Publicado: (2022)
Automatic Differentiation of Optimization Algorithms with Time-Varying Updates
por: Mehmood, Sheheryar, et al.
Publicado: (2024)
por: Mehmood, Sheheryar, et al.
Publicado: (2024)
From Learning to Optimize to Learning Optimization Algorithms
por: Castera, Camille, et al.
Publicado: (2024)
por: Castera, Camille, et al.
Publicado: (2024)
A Generalization Result for Convergence in Learning-to-Optimize
por: Sucker, Michael, et al.
Publicado: (2024)
por: Sucker, Michael, et al.
Publicado: (2024)
Learning-to-Optimize with PAC-Bayesian Guarantees: Theoretical Considerations and Practical Implementation
por: Sucker, Michael, et al.
Publicado: (2024)
por: Sucker, Michael, et al.
Publicado: (2024)
Curse of Dimensionality in Neural Network Optimization
por: Na, Sanghoon, et al.
Publicado: (2025)
por: Na, Sanghoon, et al.
Publicado: (2025)
Analyzing and Enhancing the Backward-Pass Convergence of Unrolled Optimization
por: Kotary, James, et al.
Publicado: (2023)
por: Kotary, James, et al.
Publicado: (2023)
HUANet: Hard-Constrained Unrolled ADMM for Constrained Convex Optimization
por: Tran, Trinh, et al.
Publicado: (2026)
por: Tran, Trinh, et al.
Publicado: (2026)
PDHG-Unrolled Learning-to-Optimize Method for Large-Scale Linear Programming
por: Li, Bingheng, et al.
Publicado: (2024)
por: Li, Bingheng, et al.
Publicado: (2024)
An Efficient Unsupervised Framework for Convex Quadratic Programs via Deep Unrolling
por: Yang, Linxin, et al.
Publicado: (2024)
por: Yang, Linxin, et al.
Publicado: (2024)
Near-optimal Closed-loop Method via Lyapunov Damping for Convex Optimization
por: Maier, Severin, et al.
Publicado: (2023)
por: Maier, Severin, et al.
Publicado: (2023)
Blessings and Curses of Covariate Shifts: Adversarial Learning Dynamics, Directional Convergence, and Equilibria
por: Liang, Tengyuan
Publicado: (2022)
por: Liang, Tengyuan
Publicado: (2022)
From Cursed to Competitive: Closing the ZO-FO Gap via Input-to-State Stability
por: Farzin, Amir Ali, et al.
Publicado: (2026)
por: Farzin, Amir Ali, et al.
Publicado: (2026)
Understanding Lookahead Dynamics Through Laplace Transform
por: Sanyal, Aniket, et al.
Publicado: (2025)
por: Sanyal, Aniket, et al.
Publicado: (2025)
From Gradient Clipping to Normalization for Heavy Tailed SGD
por: Hübler, Florian, et al.
Publicado: (2024)
por: Hübler, Florian, et al.
Publicado: (2024)
Can SGD Handle Heavy-Tailed Noise?
por: Fatkhullin, Ilyas, et al.
Publicado: (2025)
por: Fatkhullin, Ilyas, et al.
Publicado: (2025)
Fragility-aware Classification for Understanding Risk and Improving Generalization
por: Yang, Chen, et al.
Publicado: (2025)
por: Yang, Chen, et al.
Publicado: (2025)
Gradient descent in matrix factorization: Understanding large initialization
por: Chen, Hengchao, et al.
Publicado: (2023)
por: Chen, Hengchao, et al.
Publicado: (2023)
Geometric design of the tangent term in landing algorithms for orthogonality constraints
por: Goyens, Florentin, et al.
Publicado: (2025)
por: Goyens, Florentin, et al.
Publicado: (2025)
Solving Boltzmann Optimization Problems with Deep Learning
por: Knoll, Fiona, et al.
Publicado: (2024)
por: Knoll, Fiona, et al.
Publicado: (2024)
Understanding the Implicit Regularization of Gradient Descent in Over-parameterized Models
por: Ma, Jianhao, et al.
Publicado: (2025)
por: Ma, Jianhao, et al.
Publicado: (2025)
Understanding Outer Optimizers in Local SGD: Learning Rates, Momentum, and Acceleration
por: Khaled, Ahmed, et al.
Publicado: (2025)
por: Khaled, Ahmed, et al.
Publicado: (2025)
Understanding the training of infinitely deep and wide ResNets with Conditional Optimal Transport
por: Barboni, Raphaël, et al.
Publicado: (2024)
por: Barboni, Raphaël, et al.
Publicado: (2024)
Leveraging Continuous Time to Understand Momentum When Training Diagonal Linear Networks
por: Papazov, Hristo, et al.
Publicado: (2024)
por: Papazov, Hristo, et al.
Publicado: (2024)
Understanding Adam Optimizer via Online Learning of Updates: Adam is FTRL in Disguise
por: Ahn, Kwangjun, et al.
Publicado: (2024)
por: Ahn, Kwangjun, et al.
Publicado: (2024)
Understanding SGD with Exponential Moving Average: A Case Study in Linear Regression
por: Li, Xuheng, et al.
Publicado: (2025)
por: Li, Xuheng, et al.
Publicado: (2025)
Towards Understanding Gradient Flow Dynamics of Homogeneous Neural Networks Beyond the Origin
por: Kumar, Akshay, et al.
Publicado: (2025)
por: Kumar, Akshay, et al.
Publicado: (2025)
Towards Understanding Generalization and Stability Gaps between Centralized and Decentralized Federated Learning
por: Sun, Yan, et al.
Publicado: (2023)
por: Sun, Yan, et al.
Publicado: (2023)
Physics-Informed Graph Neural Network for Dynamic Reconfiguration of Power Systems
por: Authier, Jules, et al.
Publicado: (2023)
por: Authier, Jules, et al.
Publicado: (2023)
Decision-Dependent Stochastic Optimization: The Role of Distribution Dynamics
por: He, Zhiyu, et al.
Publicado: (2025)
por: He, Zhiyu, et al.
Publicado: (2025)
Understanding Gradient Orthogonalization for Deep Learning via Non-Euclidean Trust-Region Optimization
por: Kovalev, Dmitry
Publicado: (2025)
por: Kovalev, Dmitry
Publicado: (2025)
Contractivity and linear convergence in bilinear saddle-point problems: An operator-theoretic approach
por: Dirren, Colin, et al.
Publicado: (2024)
por: Dirren, Colin, et al.
Publicado: (2024)
Landing with the Score: Riemannian Optimization through Denoising
por: Kharitenko, Andrey, et al.
Publicado: (2025)
por: Kharitenko, Andrey, et al.
Publicado: (2025)
Semi-on-Demand Transit Feeders with Shared Autonomous Vehicles and Reinforcement-Learning-Based Zonal Dispatching Control
por: Ng, Max T. M., et al.
Publicado: (2025)
por: Ng, Max T. M., et al.
Publicado: (2025)
Towards a Systems Theory of Algorithms
por: Dörfler, Florian, et al.
Publicado: (2024)
por: Dörfler, Florian, et al.
Publicado: (2024)
Optimistic Online LQR via Intrinsic Rewards
por: Bartos, Marcell, et al.
Publicado: (2026)
por: Bartos, Marcell, et al.
Publicado: (2026)
A Computation and Communication Efficient Method for Distributed Nonconvex Problems in the Partial Participation Setting
por: Tyurin, Alexander, et al.
Publicado: (2022)
por: Tyurin, Alexander, et al.
Publicado: (2022)
Non-Euclidean Broximal Point Method: A Blueprint for Geometry-Aware Optimization
por: Gruntkowska, Kaja, et al.
Publicado: (2025)
por: Gruntkowska, Kaja, et al.
Publicado: (2025)
MARINA-P: Superior Performance in Non-smooth Federated Optimization with Adaptive Stepsizes
por: Sokolov, Igor, et al.
Publicado: (2024)
por: Sokolov, Igor, et al.
Publicado: (2024)
Convergence Analysis of the PAGE Stochastic Algorithm for Weakly Convex Finite-Sum Optimization
por: Condat, Laurent, et al.
Publicado: (2025)
por: Condat, Laurent, et al.
Publicado: (2025)
Ejemplares similares
-
Fixed-Point Automatic Differentiation of Forward--Backward Splitting Algorithms for Partly Smooth Functions
por: Mehmood, Sheheryar, et al.
Publicado: (2022) -
Automatic Differentiation of Optimization Algorithms with Time-Varying Updates
por: Mehmood, Sheheryar, et al.
Publicado: (2024) -
From Learning to Optimize to Learning Optimization Algorithms
por: Castera, Camille, et al.
Publicado: (2024) -
A Generalization Result for Convergence in Learning-to-Optimize
por: Sucker, Michael, et al.
Publicado: (2024) -
Learning-to-Optimize with PAC-Bayesian Guarantees: Theoretical Considerations and Practical Implementation
por: Sucker, Michael, et al.
Publicado: (2024)