Learning Multi-Index Models with Neural Networks via Mean-Field Langevin Dynamics
Fuente:
arXiv
Saved in:
| Main Authors: | Mousavi-Hosseini, Alireza, Wu, Denny, Erdogdu, Murat A. |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Robust Feature Learning for Multi-Index Models in High Dimensions
by: Mousavi-Hosseini, Alireza, et al.
Published: (2024)
by: Mousavi-Hosseini, Alireza, et al.
Published: (2024)
When Do Transformers Outperform Feedforward and Recurrent Networks? A Statistical Perspective
by: Mousavi-Hosseini, Alireza, et al.
Published: (2025)
by: Mousavi-Hosseini, Alireza, et al.
Published: (2025)
Post-Training with Policy Gradients: Optimality and the Base Model Barrier
by: Mousavi-Hosseini, Alireza, et al.
Published: (2026)
by: Mousavi-Hosseini, Alireza, et al.
Published: (2026)
Mean-Field Langevin Dynamics for Signed Measures via a Bilevel Approach
by: Wang, Guillaume, et al.
Published: (2024)
by: Wang, Guillaume, et al.
Published: (2024)
From Information to Generative Exponent: Learning Rate Induces Phase Transitions in SGD
by: Tsiolis, Konstantinos Christopher, et al.
Published: (2025)
by: Tsiolis, Konstantinos Christopher, et al.
Published: (2025)
A Separation in Heavy-Tailed Sampling: Gaussian vs. Stable Oracles for Proximal Samplers
by: He, Ye, et al.
Published: (2024)
by: He, Ye, et al.
Published: (2024)
Learning quadratic neural networks in high dimensions: SGD dynamics and scaling laws
by: Arous, Gérard Ben, et al.
Published: (2025)
by: Arous, Gérard Ben, et al.
Published: (2025)
Sampling from the Mean-Field Stationary Distribution
by: Kook, Yunbum, et al.
Published: (2024)
by: Kook, Yunbum, et al.
Published: (2024)
Pruning is Optimal for Learning Sparse Features in High-Dimensions
by: Vural, Nuri Mert, et al.
Published: (2024)
by: Vural, Nuri Mert, et al.
Published: (2024)
Analysis of Langevin Monte Carlo from Poincaré to Log-Sobolev
by: Chewi, Sinho, et al.
Published: (2021)
by: Chewi, Sinho, et al.
Published: (2021)
Thinned Mean Field Langevin Dynamics
by: Chen, Zonghao, et al.
Published: (2026)
by: Chen, Zonghao, et al.
Published: (2026)
Mirror Mean-Field Langevin Dynamics
by: Gu, Anming, et al.
Published: (2025)
by: Gu, Anming, et al.
Published: (2025)
Optimal Excess Risk Bounds for Empirical Risk Minimization on $p$-Norm Linear Regression
by: Hanchi, Ayoub El, et al.
Published: (2023)
by: Hanchi, Ayoub El, et al.
Published: (2023)
Propagation of Chaos for Mean-Field Langevin Dynamics and its Application to Model Ensemble
by: Nitanda, Atsushi, et al.
Published: (2025)
by: Nitanda, Atsushi, et al.
Published: (2025)
Private Continuous-Time Synthetic Trajectory Generation via Mean-Field Langevin Dynamics
by: Gu, Anming, et al.
Published: (2025)
by: Gu, Anming, et al.
Published: (2025)
On the Efficiency of ERM in Feature Learning
by: Hanchi, Ayoub El, et al.
Published: (2024)
by: Hanchi, Ayoub El, et al.
Published: (2024)
On Fitting Flow Models with Large Sinkhorn Couplings
by: Zhang, Stephen, et al.
Published: (2025)
by: Zhang, Stephen, et al.
Published: (2025)
A Geometric Analysis of PCA
by: Hanchi, Ayoub El, et al.
Published: (2025)
by: Hanchi, Ayoub El, et al.
Published: (2025)
Minimax Linear Regression under the Quantile Risk
by: Hanchi, Ayoub El, et al.
Published: (2024)
by: Hanchi, Ayoub El, et al.
Published: (2024)
Beyond Labeling Oracles: What does it mean to steal ML models?
by: Shafran, Avital, et al.
Published: (2023)
by: Shafran, Avital, et al.
Published: (2023)
Neural Collapse Beyond the Unconstrained Features Model: Landscape, Dynamics, and Generalization in the Mean-Field Regime
by: Wu, Diyuan, et al.
Published: (2025)
by: Wu, Diyuan, et al.
Published: (2025)
Flow Matching with Semidiscrete Couplings
by: Mousavi-Hosseini, Alireza, et al.
Published: (2025)
by: Mousavi-Hosseini, Alireza, et al.
Published: (2025)
Langevin Flows for Modeling Neural Latent Dynamics
by: Song, Yue, et al.
Published: (2025)
by: Song, Yue, et al.
Published: (2025)
Propagation of Chaos in One-hidden-layer Neural Networks beyond Logarithmic Time
by: Glasgow, Margalit, et al.
Published: (2025)
by: Glasgow, Margalit, et al.
Published: (2025)
Adaptive Stepsizing for Stochastic Gradient Langevin Dynamics in Bayesian Neural Networks
by: Rajpal, Rajit, et al.
Published: (2025)
by: Rajpal, Rajit, et al.
Published: (2025)
Dynamical Mean-Field Theory of Self-Attention Neural Networks
by: Poc-López, Ángel, et al.
Published: (2024)
by: Poc-López, Ángel, et al.
Published: (2024)
Symmetric Mean-field Langevin Dynamics for Distributional Minimax Problems
by: Kim, Juno, et al.
Published: (2023)
by: Kim, Juno, et al.
Published: (2023)
Learning Latent Energy-Based Models via Interacting Particle Langevin Dynamics
by: Marks, Joanna, et al.
Published: (2025)
by: Marks, Joanna, et al.
Published: (2025)
Improved Particle Approximation Error for Mean Field Neural Networks
by: Nitanda, Atsushi
Published: (2024)
by: Nitanda, Atsushi
Published: (2024)
BACE: Behavior-Adaptive Connectivity Estimation for Interpretable Graphs of Neural Dynamics
by: Asadi, Mehrnaz, et al.
Published: (2025)
by: Asadi, Mehrnaz, et al.
Published: (2025)
Imposing Boundary Conditions on Neural Operators via Learned Function Extensions
by: Mousavi, Sepehr, et al.
Published: (2026)
by: Mousavi, Sepehr, et al.
Published: (2026)
Controlled Langevin Dynamics for Sampling of Feedforward Neural Networks Trained with Minibatches
by: Zambon, Alessandro, et al.
Published: (2026)
by: Zambon, Alessandro, et al.
Published: (2026)
Full-Batch Gradient Descent Outperforms One-Pass SGD: Sample Complexity Separation in Single-Index Learning
by: Kovačević, Filip, et al.
Published: (2026)
by: Kovačević, Filip, et al.
Published: (2026)
Symmetries in Overparametrized Neural Networks: A Mean-Field View
by: Maass, Javier, et al.
Published: (2024)
by: Maass, Javier, et al.
Published: (2024)
Microcanonical Langevin Ensembles: Advancing the Sampling of Bayesian Neural Networks
by: Sommer, Emanuel, et al.
Published: (2025)
by: Sommer, Emanuel, et al.
Published: (2025)
Efficient Relation-aware Neighborhood Aggregation in Graph Neural Networks via Tensor Decomposition
by: Baghershahi, Peyman, et al.
Published: (2022)
by: Baghershahi, Peyman, et al.
Published: (2022)
On the Robustness of Langevin Dynamics to Score Function Error
by: Cao, Daniel Yiming, et al.
Published: (2026)
by: Cao, Daniel Yiming, et al.
Published: (2026)
Sampling from Bayesian Neural Network Posteriors with Symmetric Minibatch Splitting Langevin Dynamics
by: Paulin, Daniel, et al.
Published: (2024)
by: Paulin, Daniel, et al.
Published: (2024)
Mean-Field Langevin Diffusions with Density-dependent Temperature
by: Huang, Yu-Jui, et al.
Published: (2025)
by: Huang, Yu-Jui, et al.
Published: (2025)
On the Hardness of Sampling from Mixture Distributions via Langevin Dynamics
by: Cheng, Xiwei, et al.
Published: (2024)
by: Cheng, Xiwei, et al.
Published: (2024)
Similar Items
-
Robust Feature Learning for Multi-Index Models in High Dimensions
by: Mousavi-Hosseini, Alireza, et al.
Published: (2024) -
When Do Transformers Outperform Feedforward and Recurrent Networks? A Statistical Perspective
by: Mousavi-Hosseini, Alireza, et al.
Published: (2025) -
Post-Training with Policy Gradients: Optimality and the Base Model Barrier
by: Mousavi-Hosseini, Alireza, et al.
Published: (2026) -
Mean-Field Langevin Dynamics for Signed Measures via a Bilevel Approach
by: Wang, Guillaume, et al.
Published: (2024) -
From Information to Generative Exponent: Learning Rate Induces Phase Transitions in SGD
by: Tsiolis, Konstantinos Christopher, et al.
Published: (2025)