On the Benefits of Over-parameterization for Out-of-Distribution Generalization
Fuente:
arXiv
Saved in:
| Main Authors: | Hao, Yifan, Lin, Yong, Zou, Difan, Zhang, Tong |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Spurious Feature Diversification Improves Out-of-distribution Generalization
by: Lin, Yong, et al.
Published: (2023)
by: Lin, Yong, et al.
Published: (2023)
Stochastic Trust-Region Methods for Over-parameterized Models
by: Yang, Aike, et al.
Published: (2026)
by: Yang, Aike, et al.
Published: (2026)
STRAP: Spatio-Temporal Pattern Retrieval for Out-of-Distribution Generalization
by: Zhang, Haoyu, et al.
Published: (2025)
by: Zhang, Haoyu, et al.
Published: (2025)
DivIL: Unveiling and Addressing Over-Invariance for Out-of- Distribution Generalization
by: Wang, Jiaqi, et al.
Published: (2025)
by: Wang, Jiaqi, et al.
Published: (2025)
Diagonal Over-parameterization in Reproducing Kernel Hilbert Spaces as an Adaptive Feature Model: Generalization and Adaptivity
by: Li, Yicheng, et al.
Published: (2025)
by: Li, Yicheng, et al.
Published: (2025)
Depth-induced NTK: Bridging Over-parameterized Neural Networks and Deep Neural Kernels
by: Tian, Yong-Ming, et al.
Published: (2025)
by: Tian, Yong-Ming, et al.
Published: (2025)
Hierarchical Koopman Diffusion: Fast Generation with Interpretable Diffusion Trajectory
by: Bai, Hanru, et al.
Published: (2025)
by: Bai, Hanru, et al.
Published: (2025)
Physics-Informed Distillation of Diffusion Models for PDE-Constrained Generation
by: Zhang, Yi, et al.
Published: (2025)
by: Zhang, Yi, et al.
Published: (2025)
What Can Transformer Learn with Varying Depth? Case Studies on Sequence Learning Tasks
by: Chen, Xingwu, et al.
Published: (2024)
by: Chen, Xingwu, et al.
Published: (2024)
Improving Group Robustness on Spurious Correlation Requires Preciser Group Inference
by: Han, Yujin, et al.
Published: (2024)
by: Han, Yujin, et al.
Published: (2024)
Faster Sampling via Stochastic Gradient Proximal Sampler
by: Huang, Xunpeng, et al.
Published: (2024)
by: Huang, Xunpeng, et al.
Published: (2024)
Capturing Conditional Dependence via Auto-regressive Diffusion Models
by: Huang, Xunpeng, et al.
Published: (2025)
by: Huang, Xunpeng, et al.
Published: (2025)
The Implicit Bias of Adam on Separable Data
by: Zhang, Chenyang, et al.
Published: (2024)
by: Zhang, Chenyang, et al.
Published: (2024)
Sparsity and Out-of-Distribution Generalization
by: Aaronson, Scott, et al.
Published: (2026)
by: Aaronson, Scott, et al.
Published: (2026)
An Improved Analysis of Langevin Algorithms with Prior Diffusion for Non-Log-Concave Sampling
by: Huang, Xunpeng, et al.
Published: (2024)
by: Huang, Xunpeng, et al.
Published: (2024)
Structured Role-Aware Policy Optimization for Multimodal Reasoning
by: Jiang, Bingqing, et al.
Published: (2026)
by: Jiang, Bingqing, et al.
Published: (2026)
On the Memorization of Consistency Distillation for Diffusion Models
by: Jiang, Bingqing, et al.
Published: (2026)
by: Jiang, Bingqing, et al.
Published: (2026)
Understanding the Generalization of Stochastic Gradient Adam in Learning Neural Networks
by: Tang, Xuan, et al.
Published: (2025)
by: Tang, Xuan, et al.
Published: (2025)
Towards Better Generalization via Distributional Input Projection Network
by: Hao, Yifan, et al.
Published: (2025)
by: Hao, Yifan, et al.
Published: (2025)
Mild Over-Parameterization Benefits Asymmetric Tensor PCA
by: Ding, Shihong, et al.
Published: (2026)
by: Ding, Shihong, et al.
Published: (2026)
Reverse Transition Kernel: A Flexible Framework to Accelerate Diffusion Inference
by: Huang, Xunpeng, et al.
Published: (2024)
by: Huang, Xunpeng, et al.
Published: (2024)
On the $ε$-Free Inference Complexity of Absorbing Discrete Diffusion
by: Huang, Xunpeng, et al.
Published: (2025)
by: Huang, Xunpeng, et al.
Published: (2025)
Learning under Quantization for High-Dimensional Linear Regression
by: Zhang, Dechen, et al.
Published: (2025)
by: Zhang, Dechen, et al.
Published: (2025)
How Transformers Utilize Multi-Head Attention in In-Context Learning? A Case Study on Sparse Linear Regression
by: Chen, Xingwu, et al.
Published: (2024)
by: Chen, Xingwu, et al.
Published: (2024)
Improving Implicit Regularization of SGD with Preconditioning for Least Square Problems
by: Su, Junwei, et al.
Published: (2024)
by: Su, Junwei, et al.
Published: (2024)
An In-depth Investigation of Sparse Rate Reduction in Transformer-like Models
by: Hu, Yunzhe, et al.
Published: (2024)
by: Hu, Yunzhe, et al.
Published: (2024)
PRES: Toward Scalable Memory-Based Dynamic Graph Neural Networks
by: Su, Junwei, et al.
Published: (2024)
by: Su, Junwei, et al.
Published: (2024)
Physics-Informed Neural PDE Solvers via Spatio-Temporal MeanFlow
by: Bai, Hanru, et al.
Published: (2026)
by: Bai, Hanru, et al.
Published: (2026)
On the Limitation and Experience Replay for GNNs in Continual Learning
by: Su, Junwei, et al.
Published: (2023)
by: Su, Junwei, et al.
Published: (2023)
The Implicit Bias of Steepest Descent with Mini-batch Stochastic Gradient
by: Li, Jichu, et al.
Published: (2026)
by: Li, Jichu, et al.
Published: (2026)
Hyper-SET: Designing Transformers via Hyperspherical Energy Minimization
by: Hu, Yunzhe, et al.
Published: (2025)
by: Hu, Yunzhe, et al.
Published: (2025)
Learning with Norm Constrained, Over-parameterized, Two-layer Neural Networks
by: Liu, Fanghui, et al.
Published: (2024)
by: Liu, Fanghui, et al.
Published: (2024)
A Human-Like Reasoning Framework for Multi-Phases Planning Task with Large Language Models
by: Xie, Chengxing, et al.
Published: (2024)
by: Xie, Chengxing, et al.
Published: (2024)
Towards Robust Out-of-Distribution Generalization Bounds via Sharpness
by: Zou, Yingtian, et al.
Published: (2024)
by: Zou, Yingtian, et al.
Published: (2024)
Coverage-Guaranteed Prediction Sets for Out-of-Distribution Data
by: Zou, Xin, et al.
Published: (2024)
by: Zou, Xin, et al.
Published: (2024)
Generative Pre-trained Ranking Model with Over-parameterization at Web-Scale (Extended Abstract)
by: Li, Yuchen, et al.
Published: (2024)
by: Li, Yuchen, et al.
Published: (2024)
Faster Sampling without Isoperimetry via Diffusion-based Monte Carlo
by: Huang, Xunpeng, et al.
Published: (2024)
by: Huang, Xunpeng, et al.
Published: (2024)
Almost Linear Convergence under Minimal Score Assumptions: Quantized Transition Diffusion
by: Huang, Xunpeng, et al.
Published: (2025)
by: Huang, Xunpeng, et al.
Published: (2025)
Towards Robust Graph Incremental Learning on Evolving Graphs
by: Su, Junwei, et al.
Published: (2024)
by: Su, Junwei, et al.
Published: (2024)
F-Adapter: Frequency-Adaptive Parameter-Efficient Fine-Tuning in Scientific Machine Learning
by: Zhang, Hangwei, et al.
Published: (2025)
by: Zhang, Hangwei, et al.
Published: (2025)
Similar Items
-
Spurious Feature Diversification Improves Out-of-distribution Generalization
by: Lin, Yong, et al.
Published: (2023) -
Stochastic Trust-Region Methods for Over-parameterized Models
by: Yang, Aike, et al.
Published: (2026) -
STRAP: Spatio-Temporal Pattern Retrieval for Out-of-Distribution Generalization
by: Zhang, Haoyu, et al.
Published: (2025) -
DivIL: Unveiling and Addressing Over-Invariance for Out-of- Distribution Generalization
by: Wang, Jiaqi, et al.
Published: (2025) -
Diagonal Over-parameterization in Reproducing Kernel Hilbert Spaces as an Adaptive Feature Model: Generalization and Adaptivity
by: Li, Yicheng, et al.
Published: (2025)