Logarithmic Width Suffices for Robust Memorization
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Egosi, Amitsour, Yehudai, Gilad, Shamir, Ohad |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
From Tempered to Benign Overfitting in ReLU Neural Networks
von: Kornowski, Guy, et al.
Veröffentlicht: (2023)
von: Kornowski, Guy, et al.
Veröffentlicht: (2023)
Querying Kernel Methods Suffices for Reconstructing their Training Data
von: Barzilai, Daniel, et al.
Veröffentlicht: (2025)
von: Barzilai, Daniel, et al.
Veröffentlicht: (2025)
MALT Powers Up Adversarial Attacks
von: Melamed, Odelia, et al.
Veröffentlicht: (2024)
von: Melamed, Odelia, et al.
Veröffentlicht: (2024)
Depth Separation in Norm-Bounded Infinite-Width Neural Networks
von: Parkinson, Suzanna, et al.
Veröffentlicht: (2024)
von: Parkinson, Suzanna, et al.
Veröffentlicht: (2024)
Depth-Width tradeoffs in Algorithmic Reasoning of Graph Tasks with Transformers
von: Yehudai, Gilad, et al.
Veröffentlicht: (2025)
von: Yehudai, Gilad, et al.
Veröffentlicht: (2025)
When Models Don't Collapse: On the Consistency of Iterative MLE
von: Barzilai, Daniel, et al.
Veröffentlicht: (2025)
von: Barzilai, Daniel, et al.
Veröffentlicht: (2025)
Hardness of Learning Fixed Parities with Neural Networks
von: Shoshani, Itamar, et al.
Veröffentlicht: (2025)
von: Shoshani, Itamar, et al.
Veröffentlicht: (2025)
Limitations of SGD for Multi-Index Models Beyond Statistical Queries
von: Barzilai, Daniel, et al.
Veröffentlicht: (2026)
von: Barzilai, Daniel, et al.
Veröffentlicht: (2026)
On the Benefits of Rank in Attention Layers
von: Amsel, Noah, et al.
Veröffentlicht: (2024)
von: Amsel, Noah, et al.
Veröffentlicht: (2024)
An Algorithm with Optimal Dimension-Dependence for Zero-Order Nonsmooth Nonconvex Stochastic Optimization
von: Kornowski, Guy, et al.
Veröffentlicht: (2023)
von: Kornowski, Guy, et al.
Veröffentlicht: (2023)
Gradient Descent's Last Iterate is Often (slightly) Suboptimal
von: Kornowski, Guy, et al.
Veröffentlicht: (2026)
von: Kornowski, Guy, et al.
Veröffentlicht: (2026)
Generalization in Kernel Regression Under Realistic Assumptions
von: Barzilai, Daniel, et al.
Veröffentlicht: (2023)
von: Barzilai, Daniel, et al.
Veröffentlicht: (2023)
On the Complexity of Finding Small Subgradients in Nonsmooth Optimization
von: Kornowski, Guy, et al.
Veröffentlicht: (2022)
von: Kornowski, Guy, et al.
Veröffentlicht: (2022)
Open Problem: Anytime Convergence Rate of Gradient Descent
von: Kornowski, Guy, et al.
Veröffentlicht: (2024)
von: Kornowski, Guy, et al.
Veröffentlicht: (2024)
Simple Relative Deviation Bounds for Covariance and Gram Matrices
von: Barzilai, Daniel, et al.
Veröffentlicht: (2024)
von: Barzilai, Daniel, et al.
Veröffentlicht: (2024)
RedEx: Beyond Fixed Representation Methods via Convex Optimization
von: Daniely, Amit, et al.
Veröffentlicht: (2024)
von: Daniely, Amit, et al.
Veröffentlicht: (2024)
Implicit Regularization Towards Rank Minimization in ReLU Networks
von: Timor, Nadav, et al.
Veröffentlicht: (2022)
von: Timor, Nadav, et al.
Veröffentlicht: (2022)
The Oracle Complexity of Simplex-based Matrix Games
von: Kornowski, Guy, et al.
Veröffentlicht: (2024)
von: Kornowski, Guy, et al.
Veröffentlicht: (2024)
Compositional Reasoning with Transformers, RNNs, and Chain of Thought
von: Yehudai, Gilad, et al.
Veröffentlicht: (2025)
von: Yehudai, Gilad, et al.
Veröffentlicht: (2025)
On the Reconstruction of Training Data from Group Invariant Networks
von: Elbaz, Ran, et al.
Veröffentlicht: (2024)
von: Elbaz, Ran, et al.
Veröffentlicht: (2024)
Provable Unlearning with Gradient Ascent on Two-Layer ReLU Neural Networks
von: Melamed, Odelia, et al.
Veröffentlicht: (2025)
von: Melamed, Odelia, et al.
Veröffentlicht: (2025)
Beyond Benign Overfitting in Nadaraya-Watson Interpolators
von: Barzilai, Daniel, et al.
Veröffentlicht: (2025)
von: Barzilai, Daniel, et al.
Veröffentlicht: (2025)
Are Convex Optimization Curves Convex?
von: Barzilai, Guy, et al.
Veröffentlicht: (2025)
von: Barzilai, Guy, et al.
Veröffentlicht: (2025)
On the Hardness of Meaningful Local Guarantees in Nonsmooth Nonconvex Optimization
von: Kornowski, Guy, et al.
Veröffentlicht: (2024)
von: Kornowski, Guy, et al.
Veröffentlicht: (2024)
REED-VAE: RE-Encode Decode Training for Iterative Image Editing with Diffusion Models
von: Almog, Gal, et al.
Veröffentlicht: (2025)
von: Almog, Gal, et al.
Veröffentlicht: (2025)
Immediate Derivatives Suffice for Online Recurrent Adaptation
von: Merin, Aur Shalev
Veröffentlicht: (2026)
von: Merin, Aur Shalev
Veröffentlicht: (2026)
Lost in Tokenization: Fundamental Trade-offs in Graph Tokenization for Transformers
von: Bechler-Speicher, Maya, et al.
Veröffentlicht: (2026)
von: Bechler-Speicher, Maya, et al.
Veröffentlicht: (2026)
When Does Confidence-Based Cascade Deferral Suffice?
von: Jitkrittum, Wittawat, et al.
Veröffentlicht: (2023)
von: Jitkrittum, Wittawat, et al.
Veröffentlicht: (2023)
Deterministic Nonsmooth Nonconvex Optimization
von: Jordan, Michael I., et al.
Veröffentlicht: (2023)
von: Jordan, Michael I., et al.
Veröffentlicht: (2023)
Adaptive Regret for Bandits Made Possible: Two Queries Suffice
von: Lu, Zhou, et al.
Veröffentlicht: (2024)
von: Lu, Zhou, et al.
Veröffentlicht: (2024)
Feature Augmentation of GNNs for ILPs: Local Uniqueness Suffices
von: Han, Qingyu, et al.
Veröffentlicht: (2025)
von: Han, Qingyu, et al.
Veröffentlicht: (2025)
Subsampling Suffices for Adaptive Data Analysis
von: Blanc, Guy
Veröffentlicht: (2023)
von: Blanc, Guy
Veröffentlicht: (2023)
Transformers on Markov Data: Constant Depth Suffices
von: Rajaraman, Nived, et al.
Veröffentlicht: (2024)
von: Rajaraman, Nived, et al.
Veröffentlicht: (2024)
On the Over-Memorization During Natural, Robust and Catastrophic Overfitting
von: Lin, Runqi, et al.
Veröffentlicht: (2023)
von: Lin, Runqi, et al.
Veröffentlicht: (2023)
Rethinking Multi-Label Node Classification: Do Tuned Classic GNNs Suffice?
von: Xiao, Yuxuan, et al.
Veröffentlicht: (2026)
von: Xiao, Yuxuan, et al.
Veröffentlicht: (2026)
The Multiple Ticket Hypothesis: Random Sparse Subnetworks Suffice for RLVR
von: Adewuyi, Israel, et al.
Veröffentlicht: (2026)
von: Adewuyi, Israel, et al.
Veröffentlicht: (2026)
When Can Transformers Count to n?
von: Yehudai, Gilad, et al.
Veröffentlicht: (2024)
von: Yehudai, Gilad, et al.
Veröffentlicht: (2024)
LLM-Generated Explanations Do Not Suffice for Ultra-Strong Machine Learning
von: Ai, Lun, et al.
Veröffentlicht: (2025)
von: Ai, Lun, et al.
Veröffentlicht: (2025)
Federated Offline Reinforcement Learning: Collaborative Single-Policy Coverage Suffices
von: Woo, Jiin, et al.
Veröffentlicht: (2024)
von: Woo, Jiin, et al.
Veröffentlicht: (2024)
Reconstructing Training Data From Real World Models Trained with Transfer Learning
von: Oz, Yakir, et al.
Veröffentlicht: (2024)
von: Oz, Yakir, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
From Tempered to Benign Overfitting in ReLU Neural Networks
von: Kornowski, Guy, et al.
Veröffentlicht: (2023) -
Querying Kernel Methods Suffices for Reconstructing their Training Data
von: Barzilai, Daniel, et al.
Veröffentlicht: (2025) -
MALT Powers Up Adversarial Attacks
von: Melamed, Odelia, et al.
Veröffentlicht: (2024) -
Depth Separation in Norm-Bounded Infinite-Width Neural Networks
von: Parkinson, Suzanna, et al.
Veröffentlicht: (2024) -
Depth-Width tradeoffs in Algorithmic Reasoning of Graph Tasks with Transformers
von: Yehudai, Gilad, et al.
Veröffentlicht: (2025)