Langevin Monte-Carlo Provably Learns Depth Two Neural Nets at Any Size and Data
Fuente:
arXiv
Saved in:
| Main Authors: | Kumar, Dibyakanti, Jha, Samyak, Mukherjee, Anirbit |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Convergent Stochastic Training of Attention and Understanding LoRA
by: Sun, Zhengkai, et al.
Published: (2026)
by: Sun, Zhengkai, et al.
Published: (2026)
Investigating the Ability of PINNs To Solve Burgers' PDE Near Finite-Time BlowUp
by: Kumar, Dibyakanti, et al.
Published: (2023)
by: Kumar, Dibyakanti, et al.
Published: (2023)
Towards Size-Independent Generalization Bounds for Deep Operator Nets
by: Gopalani, Pulkit, et al.
Published: (2022)
by: Gopalani, Pulkit, et al.
Published: (2022)
Global Convergence of SGD For Logistic Loss on Two Layer Neural Nets
by: Gopalani, Pulkit, et al.
Published: (2023)
by: Gopalani, Pulkit, et al.
Published: (2023)
Generalization Bounds for Physics-Informed Neural Networks for the Incompressible Navier-Stokes Equations
by: Andre-Sloan, Sebastien, et al.
Published: (2026)
by: Andre-Sloan, Sebastien, et al.
Published: (2026)
Bayesian Optimization through Gaussian Cox Process Models for Spatio-temporal Data
by: Mei, Yongsheng, et al.
Published: (2024)
by: Mei, Yongsheng, et al.
Published: (2024)
Convergence of Kinetic Langevin Monte Carlo on Lie groups
by: Kong, Lingkai, et al.
Published: (2024)
by: Kong, Lingkai, et al.
Published: (2024)
Neural Hilbert Ladders: Multi-Layer Neural Networks in Function Space
by: Chen, Zhengdao
Published: (2023)
by: Chen, Zhengdao
Published: (2023)
Does the Barron space really defy the curse of dimensionality?
by: Schavemaker, Olov
Published: (2025)
by: Schavemaker, Olov
Published: (2025)
Distributionally robust approximation property of neural networks
by: Ceylan, Mihriban, et al.
Published: (2025)
by: Ceylan, Mihriban, et al.
Published: (2025)
Smoothed Distance Kernels for MMDs and Applications in Wasserstein Gradient Flows
by: Rux, Nicolaj, et al.
Published: (2025)
by: Rux, Nicolaj, et al.
Published: (2025)
Gaussian Processes with Sample Paths in Reproducing Kernel Banach Spaces
by: Karvonen, Toni, et al.
Published: (2026)
by: Karvonen, Toni, et al.
Published: (2026)
Size Lowerbounds for Deep Operator Networks
by: Mukherjee, Anirbit, et al.
Published: (2023)
by: Mukherjee, Anirbit, et al.
Published: (2023)
Approximating Langevin Monte Carlo with ResNet-like Neural Network architectures
by: Miranda, Charles, et al.
Published: (2023)
by: Miranda, Charles, et al.
Published: (2023)
Monte Carlo Neural PDE Solver for Learning PDEs via Probabilistic Representation
by: Zhang, Rui, et al.
Published: (2023)
by: Zhang, Rui, et al.
Published: (2023)
Analysis of kinetic Langevin Monte Carlo under the stochastic exponential Euler discretization from underdamped all the way to overdamped
by: Kim, Kyurae, et al.
Published: (2025)
by: Kim, Kyurae, et al.
Published: (2025)
Poincaré Inequality for Local Log-Polyak-Lojasiewicz Measures : Non-asymptotic Analysis in Low-temperature Regime
by: Gong, Yun, et al.
Published: (2025)
by: Gong, Yun, et al.
Published: (2025)
An Interpretable Approach to Load Profile Forecasting in Power Grids using Galerkin-Approximated Koopman Pseudospectra
by: Tavasoli, Ali, et al.
Published: (2023)
by: Tavasoli, Ali, et al.
Published: (2023)
Pointwise Generalization in Deep Neural Networks
by: Li, Shaojie, et al.
Published: (2026)
by: Li, Shaojie, et al.
Published: (2026)
High-Order Langevin Monte Carlo Algorithms
by: Dang, Thanh, et al.
Published: (2025)
by: Dang, Thanh, et al.
Published: (2025)
Constructive Approximation under Carleman's Condition, with Applications to Smoothed Analysis
by: Koehler, Frederic, et al.
Published: (2025)
by: Koehler, Frederic, et al.
Published: (2025)
Stochastic approximation in infinite dimensions
by: Karandikar, Rajeeva Laxman, et al.
Published: (2024)
by: Karandikar, Rajeeva Laxman, et al.
Published: (2024)
Sharp Matrix Empirical Bernstein Inequalities
by: Wang, Hongjian, et al.
Published: (2024)
by: Wang, Hongjian, et al.
Published: (2024)
Convergence rate of Tsallis entropic regularized optimal transport
by: Suguro, Takeshi, et al.
Published: (2023)
by: Suguro, Takeshi, et al.
Published: (2023)
Hoeffding decomposition of black-box models with dependent inputs
by: Idrissi, Marouane Il, et al.
Published: (2023)
by: Idrissi, Marouane Il, et al.
Published: (2023)
Global Convergence of SGD On Two Layer Neural Nets
by: Gopalani, Pulkit, et al.
Published: (2022)
by: Gopalani, Pulkit, et al.
Published: (2022)
Accelerating Multilevel Markov Chain Monte Carlo Using Machine Learning Models
by: Reddy, Sohail, et al.
Published: (2024)
by: Reddy, Sohail, et al.
Published: (2024)
Regime-Switching Langevin Monte Carlo Algorithms
by: Wang, Xiaoyu, et al.
Published: (2025)
by: Wang, Xiaoyu, et al.
Published: (2025)
Contraction of Markovian Operators in Orlicz Spaces and Error Bounds for Markov Chain Monte Carlo
by: Esposito, Amedeo Roberto, et al.
Published: (2024)
by: Esposito, Amedeo Roberto, et al.
Published: (2024)
Hypocoercive Langevin dynamics on the Lie group $\mathrm{SE}(2)$
by: Grothaus, Martin, et al.
Published: (2026)
by: Grothaus, Martin, et al.
Published: (2026)
Wasserstein Contraction of Coordinate Ascent Variational Inference
by: Caprio, Rocco, et al.
Published: (2026)
by: Caprio, Rocco, et al.
Published: (2026)
Distribution-Free Stochastic Analysis and Robust Multilevel Vector Field Anomaly Detection
by: Castrillon-Candas, Julio E, et al.
Published: (2022)
by: Castrillon-Candas, Julio E, et al.
Published: (2022)
Accelerating Langevin Monte Carlo Sampling: A Large Deviations Analysis
by: Yao, Nian, et al.
Published: (2025)
by: Yao, Nian, et al.
Published: (2025)
Two random constructions inside lacunary sets
by: Neuwirth, Stefan
Published: (2026)
by: Neuwirth, Stefan
Published: (2026)
Parallelized Midpoint Randomization for Langevin Monte Carlo
by: Yu, Lu, et al.
Published: (2024)
by: Yu, Lu, et al.
Published: (2024)
Minimum Norm Interpolation via The Local Theory of Banach Spaces: The Role of $2$-Uniform Convexity
by: Kur, Gil, et al.
Published: (2026)
by: Kur, Gil, et al.
Published: (2026)
A hierarchical entropy method for the delocalization of bias in high-dimensional Langevin Monte Carlo
by: Lacker, Daniel, et al.
Published: (2025)
by: Lacker, Daniel, et al.
Published: (2025)
Riemannian Langevin Dynamics: Strong Convergence of Geometric Euler-Maruyama Scheme
by: Zhan, Zhiyuan, et al.
Published: (2026)
by: Zhan, Zhiyuan, et al.
Published: (2026)
Sampling from Bayesian Neural Network Posteriors with Symmetric Minibatch Splitting Langevin Dynamics
by: Paulin, Daniel, et al.
Published: (2024)
by: Paulin, Daniel, et al.
Published: (2024)
Non-asymptotic Analysis of Diffusion Annealed Langevin Monte Carlo for Generative Modelling
by: Cordero-Encinar, Paula, et al.
Published: (2025)
by: Cordero-Encinar, Paula, et al.
Published: (2025)
Similar Items
-
Convergent Stochastic Training of Attention and Understanding LoRA
by: Sun, Zhengkai, et al.
Published: (2026) -
Investigating the Ability of PINNs To Solve Burgers' PDE Near Finite-Time BlowUp
by: Kumar, Dibyakanti, et al.
Published: (2023) -
Towards Size-Independent Generalization Bounds for Deep Operator Nets
by: Gopalani, Pulkit, et al.
Published: (2022) -
Global Convergence of SGD For Logistic Loss on Two Layer Neural Nets
by: Gopalani, Pulkit, et al.
Published: (2023) -
Generalization Bounds for Physics-Informed Neural Networks for the Incompressible Navier-Stokes Equations
by: Andre-Sloan, Sebastien, et al.
Published: (2026)