A Gentle Introduction to Gradient-Based Optimization and Variational Inequalities for Machine Learning
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wadia, Neha S., Dandi, Yatin, Jordan, Michael I. |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2023
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
The Computational Advantage of Depth: Learning High-Dimensional Hierarchical Functions with Gradient Descent
von: Dandi, Yatin, et al.
Veröffentlicht: (2025)
von: Dandi, Yatin, et al.
Veröffentlicht: (2025)
A mixing time bound for Gibbs sampling from log-smooth log-concave distributions
von: Wadia, Neha S.
Veröffentlicht: (2024)
von: Wadia, Neha S.
Veröffentlicht: (2024)
Deep Learning of Compositional Targets with Hierarchical Spectral Methods
von: Tabanelli, Hugo, et al.
Veröffentlicht: (2026)
von: Tabanelli, Hugo, et al.
Veröffentlicht: (2026)
Provable Learning of Random Hierarchy Models and Hierarchical Shallow-to-Deep Chaining
von: Ren, Yunwei, et al.
Veröffentlicht: (2026)
von: Ren, Yunwei, et al.
Veröffentlicht: (2026)
The Benefits of Reusing Batches for Gradient Descent in Two-Layer Networks: Breaking the Curse of Information and Leap Exponents
von: Dandi, Yatin, et al.
Veröffentlicht: (2024)
von: Dandi, Yatin, et al.
Veröffentlicht: (2024)
How Two-Layer Neural Networks Learn, One (Giant) Step at a Time
von: Dandi, Yatin, et al.
Veröffentlicht: (2023)
von: Dandi, Yatin, et al.
Veröffentlicht: (2023)
Repetita Iuvant: Data Repetition Allows SGD to Learn High-Dimensional Multi-Index Functions
von: Arnaboldi, Luca, et al.
Veröffentlicht: (2024)
von: Arnaboldi, Luca, et al.
Veröffentlicht: (2024)
Deep Learning as Neural Low-Degree Filtering: A Spectral Theory of Hierarchical Feature Learning
von: Dandi, Yatin, et al.
Veröffentlicht: (2026)
von: Dandi, Yatin, et al.
Veröffentlicht: (2026)
Sampling with flows, diffusion and autoregressive neural networks: A spin-glass perspective
von: Ghio, Davide, et al.
Veröffentlicht: (2023)
von: Ghio, Davide, et al.
Veröffentlicht: (2023)
Online Learning and Information Exponents: On The Importance of Batch size, and Time/Complexity Tradeoffs
von: Arnaboldi, Luca, et al.
Veröffentlicht: (2024)
von: Arnaboldi, Luca, et al.
Veröffentlicht: (2024)
Rigorous Asymptotics for First-Order Algorithms Through the Dynamical Cavity Method
von: Dandi, Yatin, et al.
Veröffentlicht: (2026)
von: Dandi, Yatin, et al.
Veröffentlicht: (2026)
Asymptotics of Non-Convex Generalized Linear Models in High-Dimensions: A proof of the replica formula
von: Vilucchio, Matteo, et al.
Veröffentlicht: (2025)
von: Vilucchio, Matteo, et al.
Veröffentlicht: (2025)
Perseus: A Simple and Optimal High-Order Method for Variational Inequalities
von: Lin, Tianyi, et al.
Veröffentlicht: (2022)
von: Lin, Tianyi, et al.
Veröffentlicht: (2022)
Universality laws for Gaussian mixtures in generalized linear models
von: Dandi, Yatin, et al.
Veröffentlicht: (2023)
von: Dandi, Yatin, et al.
Veröffentlicht: (2023)
A Random Matrix Theory Perspective on the Spectrum of Learned Features and Asymptotic Generalization Capabilities
von: Dandi, Yatin, et al.
Veröffentlicht: (2024)
von: Dandi, Yatin, et al.
Veröffentlicht: (2024)
Scaling Laws from Sequential Feature Recovery: A Solvable Hierarchical Model
von: Wortsman-Zurich, Arie, et al.
Veröffentlicht: (2026)
von: Wortsman-Zurich, Arie, et al.
Veröffentlicht: (2026)
A Primal-Dual Approach to Solving Variational Inequalities with General Constraints
von: Chavdarova, Tatjana, et al.
Veröffentlicht: (2022)
von: Chavdarova, Tatjana, et al.
Veröffentlicht: (2022)
Fundamental limits of learning in sequence multi-index models and deep attention networks: High-dimensional asymptotics and sharp thresholds
von: Troiani, Emanuele, et al.
Veröffentlicht: (2025)
von: Troiani, Emanuele, et al.
Veröffentlicht: (2025)
Optimal Spectral Transitions in High-Dimensional Multi-Index Models
von: Defilippis, Leonardo, et al.
Veröffentlicht: (2025)
von: Defilippis, Leonardo, et al.
Veröffentlicht: (2025)
A Gentle Introduction to Conformal Time Series Forecasting
von: Stocker, M., et al.
Veröffentlicht: (2025)
von: Stocker, M., et al.
Veröffentlicht: (2025)
A Gentle Introduction and Tutorial on Deep Generative Models in Transportation Research
von: Choi, Seongjin, et al.
Veröffentlicht: (2024)
von: Choi, Seongjin, et al.
Veröffentlicht: (2024)
Learning Variational Inequalities from Data: Fast Generalization Rates under Strong Monotonicity
von: Zhao, Eric, et al.
Veröffentlicht: (2024)
von: Zhao, Eric, et al.
Veröffentlicht: (2024)
Fundamental computational limits of weak learnability in high-dimensional multi-index models
von: Troiani, Emanuele, et al.
Veröffentlicht: (2024)
von: Troiani, Emanuele, et al.
Veröffentlicht: (2024)
Symbol Grounding in Neuro-Symbolic AI: A Gentle Introduction to Reasoning Shortcuts
von: Marconato, Emanuele, et al.
Veröffentlicht: (2025)
von: Marconato, Emanuele, et al.
Veröffentlicht: (2025)
Two-Timescale Gradient Descent Ascent Algorithms for Nonconvex Minimax Optimization
von: Lin, Tianyi, et al.
Veröffentlicht: (2024)
von: Lin, Tianyi, et al.
Veröffentlicht: (2024)
Asymptotics of feature learning in two-layer networks after one gradient-step
von: Cui, Hugo, et al.
Veröffentlicht: (2024)
von: Cui, Hugo, et al.
Veröffentlicht: (2024)
Introduction to Machine Learning
von: Younes, Laurent
Veröffentlicht: (2024)
von: Younes, Laurent
Veröffentlicht: (2024)
Delegating Data Collection in Decentralized Machine Learning
von: Ananthakrishnan, Nivasini, et al.
Veröffentlicht: (2023)
von: Ananthakrishnan, Nivasini, et al.
Veröffentlicht: (2023)
Reduced-Rank Multi-objective Policy Learning and Optimization
von: Nwankwo, Ezinne, et al.
Veröffentlicht: (2024)
von: Nwankwo, Ezinne, et al.
Veröffentlicht: (2024)
Gentle Local Robustness implies Generalization
von: Than, Khoat, et al.
Veröffentlicht: (2024)
von: Than, Khoat, et al.
Veröffentlicht: (2024)
An Optimistic Algorithm for Online Convex Optimization with Adversarial Constraints
von: Lekeufack, Jordan, et al.
Veröffentlicht: (2024)
von: Lekeufack, Jordan, et al.
Veröffentlicht: (2024)
An Integrative Genome-Scale Metabolic Modeling and Machine Learning Framework for Predicting and Optimizing Single-Cell Protein Production in Saccharomyces cerevisiae
von: Nair, Neha K., et al.
Veröffentlicht: (2026)
von: Nair, Neha K., et al.
Veröffentlicht: (2026)
Gradient Equilibrium in Online Learning: Theory and Applications
von: Angelopoulos, Anastasios N., et al.
Veröffentlicht: (2025)
von: Angelopoulos, Anastasios N., et al.
Veröffentlicht: (2025)
Learning Gentle Grasping from Human-Free Force Control Demonstration
von: Li, Mingxuan, et al.
Veröffentlicht: (2024)
von: Li, Mingxuan, et al.
Veröffentlicht: (2024)
A Variational Estimator for $L_p$ Calibration Errors
von: Berta, Eugène, et al.
Veröffentlicht: (2026)
von: Berta, Eugène, et al.
Veröffentlicht: (2026)
On Gradient Descent Ascent for Nonconvex-Concave Minimax Problems
von: Lin, Tianyi, et al.
Veröffentlicht: (2019)
von: Lin, Tianyi, et al.
Veröffentlicht: (2019)
An Introduction to Discrete Variational Autoencoders
von: Jeffares, Alan, et al.
Veröffentlicht: (2025)
von: Jeffares, Alan, et al.
Veröffentlicht: (2025)
A Brief Introduction to Causal Inference in Machine Learning
von: Cho, Kyunghyun
Veröffentlicht: (2024)
von: Cho, Kyunghyun
Veröffentlicht: (2024)
Faster Rates For Federated Variational Inequalities
von: Wang, Guanghui, et al.
Veröffentlicht: (2026)
von: Wang, Guanghui, et al.
Veröffentlicht: (2026)
Improved Dimension Dependence for Bandit Convex Optimization with Gradient Variations
von: Yu, Hang, et al.
Veröffentlicht: (2026)
von: Yu, Hang, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
The Computational Advantage of Depth: Learning High-Dimensional Hierarchical Functions with Gradient Descent
von: Dandi, Yatin, et al.
Veröffentlicht: (2025) -
A mixing time bound for Gibbs sampling from log-smooth log-concave distributions
von: Wadia, Neha S.
Veröffentlicht: (2024) -
Deep Learning of Compositional Targets with Hierarchical Spectral Methods
von: Tabanelli, Hugo, et al.
Veröffentlicht: (2026) -
Provable Learning of Random Hierarchy Models and Hierarchical Shallow-to-Deep Chaining
von: Ren, Yunwei, et al.
Veröffentlicht: (2026) -
The Benefits of Reusing Batches for Gradient Descent in Two-Layer Networks: Breaking the Curse of Information and Leap Exponents
von: Dandi, Yatin, et al.
Veröffentlicht: (2024)