Generalization of Scaled Deep ResNets in the Mean-Field Regime
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Chen, Yihang, Liu, Fanghui, Lu, Yiping, Chrysos, Grigorios G., Cevher, Volkan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Revisiting Character-level Adversarial Attacks for Language Models
von: Rocamora, Elias Abad, et al.
Veröffentlicht: (2024)
von: Rocamora, Elias Abad, et al.
Veröffentlicht: (2024)
Robust NAS under adversarial training: benchmark, theory, and beyond
von: Wu, Yongtao, et al.
Veröffentlicht: (2024)
von: Wu, Yongtao, et al.
Veröffentlicht: (2024)
Efficient local linearity regularization to overcome catastrophic overfitting
von: Rocamora, Elias Abad, et al.
Veröffentlicht: (2024)
von: Rocamora, Elias Abad, et al.
Veröffentlicht: (2024)
Certified Robustness Under Bounded Levenshtein Distance
von: Rocamora, Elias Abad, et al.
Veröffentlicht: (2025)
von: Rocamora, Elias Abad, et al.
Veröffentlicht: (2025)
High-Dimensional Kernel Methods under Covariate Shift: Data-Dependent Implicit Regularization
von: Chen, Yihang, et al.
Veröffentlicht: (2024)
von: Chen, Yihang, et al.
Veröffentlicht: (2024)
Scaling ResNets in the Large-depth Regime
von: Marion, Pierre, et al.
Veröffentlicht: (2022)
von: Marion, Pierre, et al.
Veröffentlicht: (2022)
REST: Efficient and Accelerated EEG Seizure Analysis through Residual State Updates
von: Afzal, Arshia, et al.
Veröffentlicht: (2024)
von: Afzal, Arshia, et al.
Veröffentlicht: (2024)
Leveraging the Context through Multi-Round Interactions for Jailbreaking Attacks
von: Cheng, Yixin, et al.
Veröffentlicht: (2024)
von: Cheng, Yixin, et al.
Veröffentlicht: (2024)
Multilinear Operator Networks
von: Cheng, Yixin, et al.
Veröffentlicht: (2024)
von: Cheng, Yixin, et al.
Veröffentlicht: (2024)
Learning to Remove Cuts in Integer Linear Programming
von: Puigdemont, Pol, et al.
Veröffentlicht: (2024)
von: Puigdemont, Pol, et al.
Veröffentlicht: (2024)
The Last Mile to Supervised Performance: Semi-Supervised Domain Adaptation for Semantic Segmentation
von: Morales-Brotons, Daniel, et al.
Veröffentlicht: (2024)
von: Morales-Brotons, Daniel, et al.
Veröffentlicht: (2024)
Hadamard product in deep learning: Introduction, Advances and Challenges
von: Chrysos, Grigorios G, et al.
Veröffentlicht: (2025)
von: Chrysos, Grigorios G, et al.
Veröffentlicht: (2025)
Learning with Norm Constrained, Over-parameterized, Two-layer Neural Networks
von: Liu, Fanghui, et al.
Veröffentlicht: (2024)
von: Liu, Fanghui, et al.
Veröffentlicht: (2024)
Single-pass Detection of Jailbreaking Input in Large Language Models
von: Candogan, Leyla Naz, et al.
Veröffentlicht: (2025)
von: Candogan, Leyla Naz, et al.
Veröffentlicht: (2025)
Arithmetic-Mean $μ$P for Modern Architectures: A Unified Learning-Rate Scale for CNNs and ResNets
von: Zhang, Haosong, et al.
Veröffentlicht: (2025)
von: Zhang, Haosong, et al.
Veröffentlicht: (2025)
Field theory for optimal signal propagation in ResNets
von: Fischer, Kirsten, et al.
Veröffentlicht: (2023)
von: Fischer, Kirsten, et al.
Veröffentlicht: (2023)
Branch Scaling Manifests as Implicit Architectural Regularization for Improving Generalization in Overparameterized ResNets
von: Yu, Zixiong, et al.
Veröffentlicht: (2024)
von: Yu, Zixiong, et al.
Veröffentlicht: (2024)
Neural Collapse is Globally Optimal in Deep Regularized ResNets and Transformers
von: Súkeník, Peter, et al.
Veröffentlicht: (2025)
von: Súkeník, Peter, et al.
Veröffentlicht: (2025)
ResNets Are Deeper Than You Think
von: Mehmeti-Göpel, Christian H. X. Ali, et al.
Veröffentlicht: (2025)
von: Mehmeti-Göpel, Christian H. X. Ali, et al.
Veröffentlicht: (2025)
L-Lipschitz Gershgorin ResNet Network
von: Juston, Marius F. R., et al.
Veröffentlicht: (2025)
von: Juston, Marius F. R., et al.
Veröffentlicht: (2025)
Transformative or Conservative? Conservation laws for ResNets and Transformers
von: Marcotte, Sibylle, et al.
Veröffentlicht: (2025)
von: Marcotte, Sibylle, et al.
Veröffentlicht: (2025)
Naturally Computed Scale Invariance in the Residual Stream of ResNet18
von: Longon, André
Veröffentlicht: (2025)
von: Longon, André
Veröffentlicht: (2025)
Interpreting the Residual Stream of ResNet18
von: Longon, André
Veröffentlicht: (2024)
von: Longon, André
Veröffentlicht: (2024)
Approximation theory for 1-Lipschitz ResNets
von: Murari, Davide, et al.
Veröffentlicht: (2025)
von: Murari, Davide, et al.
Veröffentlicht: (2025)
Going beyond Compositions, DDPMs Can Produce Zero-Shot Interpolations
von: Deschenaux, Justin, et al.
Veröffentlicht: (2024)
von: Deschenaux, Justin, et al.
Veröffentlicht: (2024)
Overparameterization of deep ResNet: zero loss and mean-field analysis
von: Ding, Zhiyan, et al.
Veröffentlicht: (2021)
von: Ding, Zhiyan, et al.
Veröffentlicht: (2021)
Deep Learning as a Convex Paradigm of Computation: Minimizing Circuit Size with ResNets
von: Jacot, Arthur
Veröffentlicht: (2025)
von: Jacot, Arthur
Veröffentlicht: (2025)
On Dissipativity of Cross-Entropy Loss in Training ResNets
von: Püttschneider, Jens, et al.
Veröffentlicht: (2024)
von: Püttschneider, Jens, et al.
Veröffentlicht: (2024)
Collective Kernel EFT for Pre-activation ResNets
von: Kawase, Hidetoshi, et al.
Veröffentlicht: (2026)
von: Kawase, Hidetoshi, et al.
Veröffentlicht: (2026)
Towards an Optimal Control Perspective of ResNet Training
von: Püttschneider, Jens, et al.
Veröffentlicht: (2025)
von: Püttschneider, Jens, et al.
Veröffentlicht: (2025)
On the inductive bias of infinite-depth ResNets and the bottleneck rank
von: Boix-Adsera, Enric
Veröffentlicht: (2025)
von: Boix-Adsera, Enric
Veröffentlicht: (2025)
Acute Lymphoblastic Leukemia Diagnosis Employing YOLOv11, YOLOv8, ResNet50, and Inception-ResNet-v2 Deep Learning Models
von: Awad, Alaa, et al.
Veröffentlicht: (2025)
von: Awad, Alaa, et al.
Veröffentlicht: (2025)
Hamiltonian Mechanics of Feature Learning: Bottleneck Structure in Leaky ResNets
von: Jacot, Arthur, et al.
Veröffentlicht: (2024)
von: Jacot, Arthur, et al.
Veröffentlicht: (2024)
Poly-MgNet: Polynomial Building Blocks in Multigrid-Inspired ResNets
von: van Betteray, Antonia, et al.
Veröffentlicht: (2025)
von: van Betteray, Antonia, et al.
Veröffentlicht: (2025)
Activation-Free Backbones for Image Recognition: Polynomial Alternatives within MetaFormer-Style Vision Models
von: Wang, Jeffrey, et al.
Veröffentlicht: (2026)
von: Wang, Jeffrey, et al.
Veröffentlicht: (2026)
The Hidden Width of Deep ResNets: Tight Error Bounds and Phase Diagram
von: Chizat, Lénaïc
Veröffentlicht: (2025)
von: Chizat, Lénaïc
Veröffentlicht: (2025)
Progressive Feedforward Collapse of ResNet Training
von: Wang, Sicong, et al.
Veröffentlicht: (2024)
von: Wang, Sicong, et al.
Veröffentlicht: (2024)
Understanding the training of infinitely deep and wide ResNets with Conditional Optimal Transport
von: Barboni, Raphaël, et al.
Veröffentlicht: (2024)
von: Barboni, Raphaël, et al.
Veröffentlicht: (2024)
ResNets of All Shapes and Sizes: Convergence of Training Dynamics in the Large-scale Limit
von: Chaintron, Louis-Pierre, et al.
Veröffentlicht: (2026)
von: Chaintron, Louis-Pierre, et al.
Veröffentlicht: (2026)
Approximating Langevin Monte Carlo with ResNet-like Neural Network architectures
von: Miranda, Charles, et al.
Veröffentlicht: (2023)
von: Miranda, Charles, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
Revisiting Character-level Adversarial Attacks for Language Models
von: Rocamora, Elias Abad, et al.
Veröffentlicht: (2024) -
Robust NAS under adversarial training: benchmark, theory, and beyond
von: Wu, Yongtao, et al.
Veröffentlicht: (2024) -
Efficient local linearity regularization to overcome catastrophic overfitting
von: Rocamora, Elias Abad, et al.
Veröffentlicht: (2024) -
Certified Robustness Under Bounded Levenshtein Distance
von: Rocamora, Elias Abad, et al.
Veröffentlicht: (2025) -
High-Dimensional Kernel Methods under Covariate Shift: Data-Dependent Implicit Regularization
von: Chen, Yihang, et al.
Veröffentlicht: (2024)