Enregistré dans:
| Auteurs principaux: | Marcotte, Sibylle, Peyré, Gabriel, Gribonval, Rémi |
|---|---|
| Format: | Preprint |
| Publié: |
2025
|
| Sujets: | |
| Accès en ligne: | https://arxiv.org/abs/2508.07370 |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Transformative or Conservative? Conservation laws for ResNets and Transformers
par: Marcotte, Sibylle, et autres
Publié: (2025)
par: Marcotte, Sibylle, et autres
Publié: (2025)
Abide by the Law and Follow the Flow: Conservation Laws for Gradient Flows
par: Marcotte, Sibylle, et autres
Publié: (2023)
par: Marcotte, Sibylle, et autres
Publié: (2023)
Keep the Momentum: Conservation Laws beyond Euclidean Gradient Flows
par: Marcotte, Sibylle, et autres
Publié: (2024)
par: Marcotte, Sibylle, et autres
Publié: (2024)
Path-conditioned training: a principled way to rescale ReLU neural networks
par: Lebeurrier, Arthur, et autres
Publié: (2026)
par: Lebeurrier, Arthur, et autres
Publié: (2026)
Ultra-fast feature learning for the training of two-layer neural networks in the two-timescale regime
par: Barboni, Raphaël, et autres
Publié: (2025)
par: Barboni, Raphaël, et autres
Publié: (2025)
Revisiting RIP guarantees for sketching operators on mixture models
par: Belhadji, Ayoub, et autres
Publié: (2023)
par: Belhadji, Ayoub, et autres
Publié: (2023)
Sketch and shift: a robust decoder for compressive clustering
par: Belhadji, Ayoub, et autres
Publié: (2023)
par: Belhadji, Ayoub, et autres
Publié: (2023)
Understanding the training of infinitely deep and wide ResNets with Conditional Optimal Transport
par: Barboni, Raphaël, et autres
Publié: (2024)
par: Barboni, Raphaël, et autres
Publié: (2024)
A path-norm toolkit for modern networks: consequences, promises and challenges
par: Gonon, Antoine, et autres
Publié: (2023)
par: Gonon, Antoine, et autres
Publié: (2023)
On the Statistical Complexity of Estimation and Testing under Privacy Constraints
par: Lalanne, Clément, et autres
Publié: (2022)
par: Lalanne, Clément, et autres
Publié: (2022)
Robust Sublinear Convergence Rates for Iterative Bregman Projections
par: Peyré, Gabriel
Publié: (2026)
par: Peyré, Gabriel
Publié: (2026)
Effective regions and kernels in continuous sparse regularisation, with application to sketched mixtures
par: De Castro, Yohann, et autres
Publié: (2025)
par: De Castro, Yohann, et autres
Publié: (2025)
Convexity in ReLU Neural Networks: beyond ICNNs?
par: Gagneux, Anne, et autres
Publié: (2025)
par: Gagneux, Anne, et autres
Publié: (2025)
Non-Vacuous Generalization Bounds: Can Rescaling Invariances Help?
par: Rouchouse, Damien, et autres
Publié: (2025)
par: Rouchouse, Damien, et autres
Publié: (2025)
A Rescaling-Invariant Lipschitz Bound Based on Path-Metrics for Modern ReLU Network Parameterizations
par: Gonon, Antoine, et autres
Publié: (2024)
par: Gonon, Antoine, et autres
Publié: (2024)
Optimal and Diffusion Transports in Machine Learning
par: Peyré, Gabriel
Publié: (2025)
par: Peyré, Gabriel
Publié: (2025)
Optimal Transport for Machine Learners
par: Peyré, Gabriel
Publié: (2025)
par: Peyré, Gabriel
Publié: (2025)
Muon Dynamics as a Spectral Wasserstein Flow
par: Peyré, Gabriel
Publié: (2026)
par: Peyré, Gabriel
Publié: (2026)
The Generation Phases of Flow Matching: a Denoising Perspective
par: Gagneux, Anne, et autres
Publié: (2025)
par: Gagneux, Anne, et autres
Publié: (2025)
PASCO (PArallel Structured COarsening): an overlay to speed up graph clustering algorithms
par: Lasalle, Etienne, et autres
Publié: (2024)
par: Lasalle, Etienne, et autres
Publié: (2024)
Towards Understanding the Universality of Transformers for Next-Token Prediction
par: Sander, Michael E., et autres
Publié: (2024)
par: Sander, Michael E., et autres
Publié: (2024)
Asymmetrical estimator for training encapsulated deep photonic neural networks
par: Wang, Yizhi, et autres
Publié: (2024)
par: Wang, Yizhi, et autres
Publié: (2024)
Comparison of neural network training strategies for the simulation of dynamical systems
par: Strasser, Paul, et autres
Publié: (2025)
par: Strasser, Paul, et autres
Publié: (2025)
Wasserstein distributional adversarial training for deep neural networks
par: Bai, Xingjian, et autres
Publié: (2025)
par: Bai, Xingjian, et autres
Publié: (2025)
Can overfitted deep neural networks in adversarial training generalize? -- An approximation viewpoint
par: Shi, Zhongjie, et autres
Publié: (2024)
par: Shi, Zhongjie, et autres
Publié: (2024)
How many samples are needed to train a deep neural network?
par: Golestaneh, Pegah, et autres
Publié: (2024)
par: Golestaneh, Pegah, et autres
Publié: (2024)
How Smooth Is Attention?
par: Castin, Valérie, et autres
Publié: (2023)
par: Castin, Valérie, et autres
Publié: (2023)
Geometry-Aware Discretization Error of Diffusion Models
par: Hurault, Samuel, et autres
Publié: (2026)
par: Hurault, Samuel, et autres
Publié: (2026)
Learning from Samples: Inverse Problems over measures via Sharpened Fenchel-Young Losses
par: Andrade, Francisco, et autres
Publié: (2025)
par: Andrade, Francisco, et autres
Publié: (2025)
On the global convergence of gradient descent for wide shallow models with bounded nonlinearities
par: Petit, Romain, et autres
Publié: (2026)
par: Petit, Romain, et autres
Publié: (2026)
Power side-channel leakage localization through adversarial training of deep neural networks
par: Gammell, Jimmy, et autres
Publié: (2024)
par: Gammell, Jimmy, et autres
Publié: (2024)
A simple theory for training response of deep neural networks
par: Nakazato, Kenichi
Publié: (2024)
par: Nakazato, Kenichi
Publié: (2024)
Efficient and provably convergent end-to-end training of deep neural networks with linear constraints
par: Yang, Zonglin, et autres
Publié: (2026)
par: Yang, Zonglin, et autres
Publié: (2026)
Enhancing Hypergradients Estimation: A Study of Preconditioning and Reparameterization
par: Ye, Zhenzhang, et autres
Publié: (2024)
par: Ye, Zhenzhang, et autres
Publié: (2024)
Towards generalizable deep ptychography neural networks
par: Vong, Albert, et autres
Publié: (2025)
par: Vong, Albert, et autres
Publié: (2025)
An axiomatized PDE model of deep neural networks
par: Wang, Tangjun, et autres
Publié: (2023)
par: Wang, Tangjun, et autres
Publié: (2023)
Mitigating multiple single-event upsets during deep neural network inference using fault-aware training
par: Vinck, Toon, et autres
Publié: (2025)
par: Vinck, Toon, et autres
Publié: (2025)
Approximation and interpolation of deep neural networks
par: Constantinescu, Vlad-Raul, et autres
Publié: (2023)
par: Constantinescu, Vlad-Raul, et autres
Publié: (2023)
On Feynman--Kac training of partial Bayesian neural networks
par: Zhao, Zheng, et autres
Publié: (2023)
par: Zhao, Zheng, et autres
Publié: (2023)
A lift for input-convex neural network training
par: Siahkoohi, Ali, et autres
Publié: (2026)
par: Siahkoohi, Ali, et autres
Publié: (2026)
Documents similaires
-
Transformative or Conservative? Conservation laws for ResNets and Transformers
par: Marcotte, Sibylle, et autres
Publié: (2025) -
Abide by the Law and Follow the Flow: Conservation Laws for Gradient Flows
par: Marcotte, Sibylle, et autres
Publié: (2023) -
Keep the Momentum: Conservation Laws beyond Euclidean Gradient Flows
par: Marcotte, Sibylle, et autres
Publié: (2024) -
Path-conditioned training: a principled way to rescale ReLU neural networks
par: Lebeurrier, Arthur, et autres
Publié: (2026) -
Ultra-fast feature learning for the training of two-layer neural networks in the two-timescale regime
par: Barboni, Raphaël, et autres
Publié: (2025)