Saved in:
| Main Authors: | Damian, Alex, Pillaud-Vivien, Loucas, Lee, Jason D., Bruna, Joan |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2403.05529 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Joint Learning in the Gaussian Single Index Model
by: Pillaud-Vivien, Loucas, et al.
Published: (2025)
by: Pillaud-Vivien, Loucas, et al.
Published: (2025)
Stochastic Differential Equations models for Least-Squares Stochastic Gradient Descent
by: Schertzer, Adrien, et al.
Published: (2024)
by: Schertzer, Adrien, et al.
Published: (2024)
On learning Gaussian multi‐index models with gradient flow part I: General properties and two‐timescale learning
by: Alberto Bietti, et al.
Published: (2025)
by: Alberto Bietti, et al.
Published: (2025)
Gradient flow dynamics of shallow ReLU networks for square loss and orthogonal inputs
by: Boursier, Etienne, et al.
Published: (2022)
by: Boursier, Etienne, et al.
Published: (2022)
The Generative Leap: Sharp Sample Complexity for Efficiently Learning Gaussian Multi-Index Models
by: Damian, Alex, et al.
Published: (2025)
by: Damian, Alex, et al.
Published: (2025)
Convergence of Shallow ReLU Networks on Weakly Interacting Data
by: Dana, Léo, et al.
Published: (2025)
by: Dana, Léo, et al.
Published: (2025)
Variational Inference for Uncertainty Quantification: an Analysis of Trade-offs
by: Margossian, Charles C., et al.
Published: (2024)
by: Margossian, Charles C., et al.
Published: (2024)
Oscillating solutions to the mean-field Langevin descent-ascent flow
by: Mourrat, Jean-Christophe, et al.
Published: (2026)
by: Mourrat, Jean-Christophe, et al.
Published: (2026)
Batch and match: black-box variational inference with a score-based divergence
by: Cai, Diana, et al.
Published: (2024)
by: Cai, Diana, et al.
Published: (2024)
A Noise Sensitivity Exponent Controls Large Statistical-to-Computational Gaps in Single- and Multi-Index Models
by: Defilippis, Leonardo, et al.
Published: (2026)
by: Defilippis, Leonardo, et al.
Published: (2026)
Improved high-dimensional estimation with Langevin dynamics and stochastic weight averaging
by: Wei, Stanley, et al.
Published: (2026)
by: Wei, Stanley, et al.
Published: (2026)
Provable Guarantees for Nonlinear Feature Learning in Three-Layer Neural Networks
by: Nichani, Eshaan, et al.
Published: (2023)
by: Nichani, Eshaan, et al.
Published: (2023)
Sample and Computationally Efficient Robust Learning of Gaussian Single-Index Models
by: Wang, Puqian, et al.
Published: (2024)
by: Wang, Puqian, et al.
Published: (2024)
How Transformers Learn Causal Structure with Gradient Descent
by: Nichani, Eshaan, et al.
Published: (2024)
by: Nichani, Eshaan, et al.
Published: (2024)
Learning Orthogonal Multi-Index Models: A Fine-Grained Information Exponent Analysis
by: Ren, Yunwei, et al.
Published: (2024)
by: Ren, Yunwei, et al.
Published: (2024)
Uniform-in-Time Weak Propagation-of-Chaos in Shallow Neural Networks
by: Glasgow, Margalit, et al.
Published: (2026)
by: Glasgow, Margalit, et al.
Published: (2026)
Towards Infinitely Long Neural Simulations: Self-Refining Neural Surrogate Models for Dynamical Systems
by: Liu, Qi, et al.
Published: (2026)
by: Liu, Qi, et al.
Published: (2026)
Survey on Algorithms for multi-index models
by: Bruna, Joan, et al.
Published: (2025)
by: Bruna, Joan, et al.
Published: (2025)
Fine-Tuning Language Models with Just Forward Passes
by: Malladi, Sadhika, et al.
Published: (2023)
by: Malladi, Sadhika, et al.
Published: (2023)
Posterior Sampling with Denoising Oracles via Tilted Transport
by: Bruna, Joan, et al.
Published: (2024)
by: Bruna, Joan, et al.
Published: (2024)
LLM Watermarking Using Mixtures and Statistical-to-Computational Gaps
by: Abdalla, Pedro, et al.
Published: (2025)
by: Abdalla, Pedro, et al.
Published: (2025)
Statistical Inference in Tensor Completion: Optimal Uncertainty Quantification and Statistical-to-Computational Gaps
by: Ma, Wanteng, et al.
Published: (2024)
by: Ma, Wanteng, et al.
Published: (2024)
Axial Neural Networks for Dimension-Free Foundation Models
by: Kim, Hyunsu, et al.
Published: (2025)
by: Kim, Hyunsu, et al.
Published: (2025)
Sparse Gaussian Graphical Models with Discrete Optimization: Computational and Statistical Perspectives
by: Behdin, Kayhan, et al.
Published: (2023)
by: Behdin, Kayhan, et al.
Published: (2023)
Learning Compositional Functions with Transformers from Easy-to-Hard Data
by: Wang, Zixuan, et al.
Published: (2025)
by: Wang, Zixuan, et al.
Published: (2025)
On the Benefits of Rank in Attention Layers
by: Amsel, Noah, et al.
Published: (2024)
by: Amsel, Noah, et al.
Published: (2024)
Propagation of Chaos in One-hidden-layer Neural Networks beyond Logarithmic Time
by: Glasgow, Margalit, et al.
Published: (2025)
by: Glasgow, Margalit, et al.
Published: (2025)
Thermalizer: Stable autoregressive neural emulation of spatiotemporal chaos
by: Pedersen, Chris, et al.
Published: (2025)
by: Pedersen, Chris, et al.
Published: (2025)
Statistical-Computational Trade-offs in Learning Multi-Index Models via Harmonic Analysis
by: Latourelle-Vigeant, Hugo, et al.
Published: (2026)
by: Latourelle-Vigeant, Hugo, et al.
Published: (2026)
Understanding Optimization in Deep Learning with Central Flows
by: Cohen, Jeremy M., et al.
Published: (2024)
by: Cohen, Jeremy M., et al.
Published: (2024)
Computational-Statistical Gaps for Improper Learning in Sparse Linear Regression
by: Buhai, Rares-Darius, et al.
Published: (2024)
by: Buhai, Rares-Darius, et al.
Published: (2024)
On the Statistical Query Complexity of Learning Semiautomata: a Random Walk Approach
by: Giapitzakis, George, et al.
Published: (2025)
by: Giapitzakis, George, et al.
Published: (2025)
Limitations of SGD for Multi-Index Models Beyond Statistical Queries
by: Barzilai, Daniel, et al.
Published: (2026)
by: Barzilai, Daniel, et al.
Published: (2026)
Data-driven multiscale modeling for correcting dynamical systems
by: Otness, Karl, et al.
Published: (2023)
by: Otness, Karl, et al.
Published: (2023)
Generative Modeling from Black-box Corruptions via Self-Consistent Stochastic Interpolants
by: Modi, Chirag, et al.
Published: (2025)
by: Modi, Chirag, et al.
Published: (2025)
Omnipredicting Single-Index Models with Multi-Index Models
by: Hu, Lunjia, et al.
Published: (2024)
by: Hu, Lunjia, et al.
Published: (2024)
Neural Network-based Partial-Linear Single-Index Models for Environmental Mixtures Analysis
by: Do, Hyungrok, et al.
Published: (2025)
by: Do, Hyungrok, et al.
Published: (2025)
Are Smaller Open-Weight LLMs Closing the Gap to Proprietary Models for Biomedical Question Answering?
by: Stachura, Damian, et al.
Published: (2025)
by: Stachura, Damian, et al.
Published: (2025)
Statistical Learning Theory in Lean 4: Empirical Processes from Scratch
by: Zhang, Yuanhe, et al.
Published: (2026)
by: Zhang, Yuanhe, et al.
Published: (2026)
Neural Networks Learn Generic Multi-Index Models Near Information-Theoretic Limit
by: Zhang, Bohan, et al.
Published: (2025)
by: Zhang, Bohan, et al.
Published: (2025)
Similar Items
-
Joint Learning in the Gaussian Single Index Model
by: Pillaud-Vivien, Loucas, et al.
Published: (2025) -
Stochastic Differential Equations models for Least-Squares Stochastic Gradient Descent
by: Schertzer, Adrien, et al.
Published: (2024) -
On learning Gaussian multi‐index models with gradient flow part I: General properties and two‐timescale learning
by: Alberto Bietti, et al.
Published: (2025) -
Gradient flow dynamics of shallow ReLU networks for square loss and orthogonal inputs
by: Boursier, Etienne, et al.
Published: (2022) -
The Generative Leap: Sharp Sample Complexity for Efficiently Learning Gaussian Multi-Index Models
by: Damian, Alex, et al.
Published: (2025)