Saved in:
| Main Authors: | Li, Binghui, Li, Yuanzhi |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2410.08503 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
On the Clean Generalization and Robust Overfitting in Adversarial Training from Two Theoretical Views: Representation Complexity and Training Dynamics
by: Li, Binghui, et al.
Published: (2023)
by: Li, Binghui, et al.
Published: (2023)
Provably learning a multi-head attention layer
by: Chen, Sitan, et al.
Published: (2024)
by: Chen, Sitan, et al.
Published: (2024)
Larger Datasets Can Be Repeated More: A Theoretical Analysis of Multi-Epoch Scaling in Linear Regression
by: Yan, Tingkai, et al.
Published: (2025)
by: Yan, Tingkai, et al.
Published: (2025)
Feature Averaging: An Implicit Bias of Gradient Descent Leading to Non-Robustness in Neural Networks
by: Li, Binghui, et al.
Published: (2024)
by: Li, Binghui, et al.
Published: (2024)
Certifiable Black-Box Attacks with Randomized Adversarial Examples: Breaking Defenses with Provable Confidence
by: Hong, Hanbin, et al.
Published: (2023)
by: Hong, Hanbin, et al.
Published: (2023)
Measure-Theoretic Anti-Causal Representation Learning
by: Behnam, Arman, et al.
Published: (2025)
by: Behnam, Arman, et al.
Published: (2025)
Inf2Guard: An Information-Theoretic Framework for Learning Privacy-Preserving Representations against Inference Attacks
by: Noorbakhsh, Sayedeh Leila, et al.
Published: (2024)
by: Noorbakhsh, Sayedeh Leila, et al.
Published: (2024)
Provable Unrestricted Adversarial Training without Compromise with Generalizability
by: Zhang, Lilin, et al.
Published: (2023)
by: Zhang, Lilin, et al.
Published: (2023)
Disentangling Feature Structure: A Mathematically Provable Two-Stage Training Dynamics in Transformers
by: Gong, Zixuan, et al.
Published: (2025)
by: Gong, Zixuan, et al.
Published: (2025)
Learning Robust and Privacy-Preserving Representations via Information Theory
by: Zhang, Binghui, et al.
Published: (2024)
by: Zhang, Binghui, et al.
Published: (2024)
Unlabeled Data Can Provably Enhance In-Context Learning of Transformers
by: Liu, Renpu, et al.
Published: (2026)
by: Liu, Renpu, et al.
Published: (2026)
How Does Overparameterization Affect Features?
by: Duzgun, Ahmet Cagri, et al.
Published: (2024)
by: Duzgun, Ahmet Cagri, et al.
Published: (2024)
You Can't Ignore Either: Unifying Structure and Feature Denoising for Robust Graph Learning
by: Yang, Tianmeng, et al.
Published: (2024)
by: Yang, Tianmeng, et al.
Published: (2024)
Adversarial Training Improves Generalization Under Distribution Shifts in Bioacoustics
by: Heinrich, René, et al.
Published: (2025)
by: Heinrich, René, et al.
Published: (2025)
Physics of Language Models: Part 1, Learning Hierarchical Language Structures
by: Allen-Zhu, Zeyuan, et al.
Published: (2023)
by: Allen-Zhu, Zeyuan, et al.
Published: (2023)
FedTilt: Towards Multi-Level Fairness-Preserving and Robust Federated Learning
by: Zhang, Binghui, et al.
Published: (2025)
by: Zhang, Binghui, et al.
Published: (2025)
Detecting Adversarial Data via Provable Adversarial Noise Amplification
by: Mumcu, Furkan, et al.
Published: (2026)
by: Mumcu, Furkan, et al.
Published: (2026)
MEAT: Median-Ensemble Adversarial Training for Improving Robustness and Generalization
by: Hu, Zhaozhe, et al.
Published: (2024)
by: Hu, Zhaozhe, et al.
Published: (2024)
Muon in Associative Memory Learning: Training Dynamics and Scaling Laws
by: Li, Binghui, et al.
Published: (2026)
by: Li, Binghui, et al.
Published: (2026)
Non-Adversarial Imitation Learning Provably Free of Compounding Errors: The Role of Bellman Constraints
by: Xu, Tian, et al.
Published: (2026)
by: Xu, Tian, et al.
Published: (2026)
Provably Invincible Adversarial Attacks on Reinforcement Learning Systems: A Rate-Distortion Information-Theoretic Approach
by: Lu, Ziqing, et al.
Published: (2025)
by: Lu, Ziqing, et al.
Published: (2025)
Provable Robustness Against a Union of $\ell_0$ Adversarial Attacks
by: Hammoudeh, Zayd, et al.
Published: (2023)
by: Hammoudeh, Zayd, et al.
Published: (2023)
Robust Satisficing Gaussian Process Bandits Under Adversarial Attacks
by: Saday, Artun, et al.
Published: (2025)
by: Saday, Artun, et al.
Published: (2025)
Feature Statistics with Uncertainty Help Adversarial Robustness
by: Wang, Ran, et al.
Published: (2025)
by: Wang, Ran, et al.
Published: (2025)
Freeze then Train: Towards Provable Representation Learning under Spurious Correlations and Feature Noise
by: Ye, Haotian, et al.
Published: (2022)
by: Ye, Haotian, et al.
Published: (2022)
Provable Adversarial Robustness in In-Context Learning
by: Zhang, Di
Published: (2026)
by: Zhang, Di
Published: (2026)
Can Mamba Learn In Context with Outliers? A Theoretical Generalization Analysis
by: Li, Hongkang, et al.
Published: (2025)
by: Li, Hongkang, et al.
Published: (2025)
Transformers Trained via Gradient Descent Can Provably Learn a Class of Teacher Models
by: Zhang, Chenyang, et al.
Published: (2026)
by: Zhang, Chenyang, et al.
Published: (2026)
Provable Training for Graph Contrastive Learning
by: Yu, Yue, et al.
Published: (2023)
by: Yu, Yue, et al.
Published: (2023)
When and How Unlabeled Data Provably Improve In-Context Learning
by: Li, Yingcong, et al.
Published: (2025)
by: Li, Yingcong, et al.
Published: (2025)
TCRL: Temporal-Coupled Adversarial Training for Robust Constrained Reinforcement Learning in Worst-Case Scenarios
by: Xu, Wentao, et al.
Published: (2026)
by: Xu, Wentao, et al.
Published: (2026)
Test-Time Training Provably Improves Transformers as In-context Learners
by: Gozeten, Halil Alperen, et al.
Published: (2025)
by: Gozeten, Halil Alperen, et al.
Published: (2025)
Leveraging Local Structure for Improving Model Explanations: An Information Propagation Approach
by: Yang, Ruo, et al.
Published: (2024)
by: Yang, Ruo, et al.
Published: (2024)
Data-Dependent Stability Analysis of Adversarial Training
by: Wang, Yihan, et al.
Published: (2024)
by: Wang, Yihan, et al.
Published: (2024)
Provable Effects of Data Replay in Continual Learning: A Feature Learning Perspective
by: Ding, Meng, et al.
Published: (2026)
by: Ding, Meng, et al.
Published: (2026)
A Theoretical Analysis of Mamba's Training Dynamics: Filtering Relevant Features for Generalization in State Space Models
by: Shandirasegaran, Mugunthan, et al.
Published: (2026)
by: Shandirasegaran, Mugunthan, et al.
Published: (2026)
Robustness Feature Adapter for Efficient Adversarial Training
by: Wu, Quanwei, et al.
Published: (2025)
by: Wu, Quanwei, et al.
Published: (2025)
ProFeAT: Projected Feature Adversarial Training for Self-Supervised Learning of Robust Representations
by: Addepalli, Sravanti, et al.
Published: (2024)
by: Addepalli, Sravanti, et al.
Published: (2024)
Unveiling Induction Heads: Provable Training Dynamics and Feature Learning in Transformers
by: Chen, Siyu, et al.
Published: (2024)
by: Chen, Siyu, et al.
Published: (2024)
Theoretical limitations of multi-layer Transformer
by: Chen, Lijie, et al.
Published: (2024)
by: Chen, Lijie, et al.
Published: (2024)
Similar Items
-
On the Clean Generalization and Robust Overfitting in Adversarial Training from Two Theoretical Views: Representation Complexity and Training Dynamics
by: Li, Binghui, et al.
Published: (2023) -
Provably learning a multi-head attention layer
by: Chen, Sitan, et al.
Published: (2024) -
Larger Datasets Can Be Repeated More: A Theoretical Analysis of Multi-Epoch Scaling in Linear Regression
by: Yan, Tingkai, et al.
Published: (2025) -
Feature Averaging: An Implicit Bias of Gradient Descent Leading to Non-Robustness in Neural Networks
by: Li, Binghui, et al.
Published: (2024) -
Certifiable Black-Box Attacks with Randomized Adversarial Examples: Breaking Defenses with Provable Confidence
by: Hong, Hanbin, et al.
Published: (2023)