Path Regularization: A Near-Complete and Optimal Nonasymptotic Generalization Theory for Multilayer Neural Networks and Double Descent Phenomenon
Fuente:
arXiv
Gespeichert in:
| 1. Verfasser: | Yu, Hao |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Near-Optimal Regret for KL-Regularized Multi-Armed Bandits
von: Ji, Kaixuan, et al.
Veröffentlicht: (2026)
von: Ji, Kaixuan, et al.
Veröffentlicht: (2026)
Neural Networks Learn Generic Multi-Index Models Near Information-Theoretic Limit
von: Zhang, Bohan, et al.
Veröffentlicht: (2025)
von: Zhang, Bohan, et al.
Veröffentlicht: (2025)
On the Optimal Sample Complexity of Offline Multi-Armed Bandits with KL Regularization
von: Ji, Kaixuan, et al.
Veröffentlicht: (2026)
von: Ji, Kaixuan, et al.
Veröffentlicht: (2026)
Neural Networks Generalize on Low Complexity Data
von: Chatterjee, Sourav, et al.
Veröffentlicht: (2024)
von: Chatterjee, Sourav, et al.
Veröffentlicht: (2024)
A Unified Pair-GRPO Family: From Implicit to Explicit Preference Constraints for Stable and General RL Alignment
von: Yu, Hao
Veröffentlicht: (2026)
von: Yu, Hao
Veröffentlicht: (2026)
Generalizability of Neural Networks Minimizing Empirical Risk Based on Expressive Ability
von: Yu, Lijia, et al.
Veröffentlicht: (2025)
von: Yu, Lijia, et al.
Veröffentlicht: (2025)
Near-Optimal Learning and Planning in Separated Latent MDPs
von: Chen, Fan, et al.
Veröffentlicht: (2024)
von: Chen, Fan, et al.
Veröffentlicht: (2024)
Chemical Reaction Networks Learn Better than Spiking Neural Networks
von: Jaffard, Sophie, et al.
Veröffentlicht: (2026)
von: Jaffard, Sophie, et al.
Veröffentlicht: (2026)
A Theory of the Mechanics of Information: Generalization Through Measurement of Uncertainty (Learning is Measuring)
von: Hazard, Christopher J., et al.
Veröffentlicht: (2025)
von: Hazard, Christopher J., et al.
Veröffentlicht: (2025)
The Geometry of Benchmarks: A New Path Toward AGI
von: Chojecki, Przemyslaw
Veröffentlicht: (2025)
von: Chojecki, Przemyslaw
Veröffentlicht: (2025)
From Spikes to Heavy Tails: Unveiling the Spectral Evolution of Neural Networks
von: Kothapalli, Vignesh, et al.
Veröffentlicht: (2024)
von: Kothapalli, Vignesh, et al.
Veröffentlicht: (2024)
Differentially Private Two-Stage Gradient Descent for Instrumental Variable Regression
von: Liang, Haodong, et al.
Veröffentlicht: (2025)
von: Liang, Haodong, et al.
Veröffentlicht: (2025)
Gradient Descent with Projection Finds Over-Parameterized Neural Networks for Learning Low-Degree Polynomials with Nearly Minimax Optimal Rate
von: Yang, Yingzhen, et al.
Veröffentlicht: (2026)
von: Yang, Yingzhen, et al.
Veröffentlicht: (2026)
Generalization Bounds: Perspectives from Information Theory and PAC-Bayes
von: Hellström, Fredrik, et al.
Veröffentlicht: (2023)
von: Hellström, Fredrik, et al.
Veröffentlicht: (2023)
On the Nonasymptotic Scaling Guarantee of Hyperparameter Estimation in Inhomogeneous, Weakly-Dependent Complex Network Dynamical Systems
von: Yu, Yi, et al.
Veröffentlicht: (2026)
von: Yu, Yi, et al.
Veröffentlicht: (2026)
Towards a Sharp Analysis of Offline Policy Learning for $f$-Divergence-Regularized Contextual Bandits
von: Zhao, Qingyue, et al.
Veröffentlicht: (2025)
von: Zhao, Qingyue, et al.
Veröffentlicht: (2025)
Enjoying Non-linearity in Multinomial Logistic Bandits: A Minimax-Optimal Algorithm
von: Boudart, Pierre, et al.
Veröffentlicht: (2025)
von: Boudart, Pierre, et al.
Veröffentlicht: (2025)
A Computational Theory for Efficient Mini Agent Evaluation with Causal Guarantees
von: Yan, Hedong
Veröffentlicht: (2025)
von: Yan, Hedong
Veröffentlicht: (2025)
Improving Kernel-Based Nonasymptotic Simultaneous Confidence Bands
von: Csáji, Balázs Csanád, et al.
Veröffentlicht: (2024)
von: Csáji, Balázs Csanád, et al.
Veröffentlicht: (2024)
A note on the impossibility of conditional PAC-efficient reasoning in large language models
von: Zeng, Hao
Veröffentlicht: (2025)
von: Zeng, Hao
Veröffentlicht: (2025)
Geometry-induced Regularization in Deep ReLU Neural Networks
von: Bona-Pellissier, Joachim, et al.
Veröffentlicht: (2024)
von: Bona-Pellissier, Joachim, et al.
Veröffentlicht: (2024)
Optimal rates for density and mode estimation with expand-and-sparsify representations
von: Sinha, Kaushik, et al.
Veröffentlicht: (2026)
von: Sinha, Kaushik, et al.
Veröffentlicht: (2026)
Minimax Optimal Variance-Aware Regret Bounds for Multinomial Logistic MDPs
von: Boudart, Pierre, et al.
Veröffentlicht: (2026)
von: Boudart, Pierre, et al.
Veröffentlicht: (2026)
Maximizing the Potential of Synthetic Data: Insights from Random Matrix Theory
von: Firdoussi, Aymane El, et al.
Veröffentlicht: (2024)
von: Firdoussi, Aymane El, et al.
Veröffentlicht: (2024)
O(d/T) Convergence Theory for Diffusion Probabilistic Models under Minimal Assumptions
von: Li, Gen, et al.
Veröffentlicht: (2024)
von: Li, Gen, et al.
Veröffentlicht: (2024)
Fast Rates for Offline Contextual Bandits with Forward-KL Regularization under Single-Policy Concentrability
von: Zhao, Qingyue, et al.
Veröffentlicht: (2026)
von: Zhao, Qingyue, et al.
Veröffentlicht: (2026)
RACER: Risk-Aware Calibrated Efficient Routing for Large Language Models
von: Hao, Sai, et al.
Veröffentlicht: (2026)
von: Hao, Sai, et al.
Veröffentlicht: (2026)
Implicit Regularization Paths of Weighted Neural Representations
von: Du, Jin-Hong, et al.
Veröffentlicht: (2024)
von: Du, Jin-Hong, et al.
Veröffentlicht: (2024)
Dropout Drops Double Descent
von: Yang, Tian-Le, et al.
Veröffentlicht: (2023)
von: Yang, Tian-Le, et al.
Veröffentlicht: (2023)
Compression, Generalization and Learning
von: Campi, Marco C., et al.
Veröffentlicht: (2023)
von: Campi, Marco C., et al.
Veröffentlicht: (2023)
On the Statistical Capacity of Deep Generative Models
von: Tam, Edric, et al.
Veröffentlicht: (2025)
von: Tam, Edric, et al.
Veröffentlicht: (2025)
Convergence of Shallow ReLU Networks on Weakly Interacting Data
von: Dana, Léo, et al.
Veröffentlicht: (2025)
von: Dana, Léo, et al.
Veröffentlicht: (2025)
Statistical Inference for Optimal Transport Maps: Recent Advances and Perspectives
von: Balakrishnan, Sivaraman, et al.
Veröffentlicht: (2025)
von: Balakrishnan, Sivaraman, et al.
Veröffentlicht: (2025)
Learning Hierarchical Polynomials of Multiple Nonlinear Features with Three-Layer Networks
von: Fu, Hengyu, et al.
Veröffentlicht: (2024)
von: Fu, Hengyu, et al.
Veröffentlicht: (2024)
Counterfactual Generative Modeling with Variational Causal Inference
von: Wu, Yulun, et al.
Veröffentlicht: (2024)
von: Wu, Yulun, et al.
Veröffentlicht: (2024)
Generalization and Scaling Laws for Mixture-of-Experts Transformers
von: Mayaki, Mansour Zoubeirou a
Veröffentlicht: (2026)
von: Mayaki, Mansour Zoubeirou a
Veröffentlicht: (2026)
On the Provable Performance Guarantee of Efficient Reasoning Models
von: Zeng, Hao, et al.
Veröffentlicht: (2025)
von: Zeng, Hao, et al.
Veröffentlicht: (2025)
Training Implicit Generative Models via an Invariant Statistical Loss
von: de Frutos, José Manuel, et al.
Veröffentlicht: (2024)
von: de Frutos, José Manuel, et al.
Veröffentlicht: (2024)
On the Statistical Properties of Generative Adversarial Models for Low Intrinsic Data Dimension
von: Chakraborty, Saptarshi, et al.
Veröffentlicht: (2024)
von: Chakraborty, Saptarshi, et al.
Veröffentlicht: (2024)
Generalization Properties of Score-matching Diffusion Models for Intrinsically Low-dimensional Data
von: Chakraborty, Saptarshi, et al.
Veröffentlicht: (2026)
von: Chakraborty, Saptarshi, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Near-Optimal Regret for KL-Regularized Multi-Armed Bandits
von: Ji, Kaixuan, et al.
Veröffentlicht: (2026) -
Neural Networks Learn Generic Multi-Index Models Near Information-Theoretic Limit
von: Zhang, Bohan, et al.
Veröffentlicht: (2025) -
On the Optimal Sample Complexity of Offline Multi-Armed Bandits with KL Regularization
von: Ji, Kaixuan, et al.
Veröffentlicht: (2026) -
Neural Networks Generalize on Low Complexity Data
von: Chatterjee, Sourav, et al.
Veröffentlicht: (2024) -
A Unified Pair-GRPO Family: From Implicit to Explicit Preference Constraints for Stable and General RL Alignment
von: Yu, Hao
Veröffentlicht: (2026)