Disentangle Sample Size and Initialization Effect on Perfect Generalization for Single-Neuron Target
Fuente:
arXiv
Saved in:
| Main Authors: | Zhao, Jiajie, Bai, Zhiwei, Zhang, Yaoyu |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Connectivity Shapes Implicit Regularization in Matrix Factorization Models for Matrix Completion
by: Bai, Zhiwei, et al.
Published: (2024)
by: Bai, Zhiwei, et al.
Published: (2024)
Local Linear Recovery Guarantee of Deep Neural Networks at Overparameterization
by: Zhang, Yaoyu, et al.
Published: (2024)
by: Zhang, Yaoyu, et al.
Published: (2024)
Initialization is Critical to Whether Transformers Fit Composite Functions by Reasoning or Memorizing
by: Zhang, Zhongwang, et al.
Published: (2024)
by: Zhang, Zhongwang, et al.
Published: (2024)
Embedding Principle in Depth for the Loss Landscape Analysis of Deep Neural Networks
by: Bai, Zhiwei, et al.
Published: (2022)
by: Bai, Zhiwei, et al.
Published: (2022)
Adaptive Preconditioners Trigger Loss Spikes in Adam
by: Bai, Zhiwei, et al.
Published: (2025)
by: Bai, Zhiwei, et al.
Published: (2025)
Latent Prototype Routing: Achieving Near-Perfect Load Balancing in Mixture-of-Experts
by: Yang, Jiajie
Published: (2025)
by: Yang, Jiajie
Published: (2025)
Uncovering Critical Sets of Deep Neural Networks via Sample-Independent Critical Lifting
by: Zhang, Leyang, et al.
Published: (2025)
by: Zhang, Leyang, et al.
Published: (2025)
MeGU: Machine-Guided Unlearning with Target Feature Disentanglement
by: Wang, Haoyu, et al.
Published: (2026)
by: Wang, Haoyu, et al.
Published: (2026)
Complexity Control Facilitates Reasoning-Based Compositional Generalization in Transformers
by: Zhang, Zhongwang, et al.
Published: (2025)
by: Zhang, Zhongwang, et al.
Published: (2025)
Enhancing Size Generalization in Graph Neural Networks through Disentangled Representation Learning
by: Huang, Zheng, et al.
Published: (2024)
by: Huang, Zheng, et al.
Published: (2024)
Mechanistic Independence: A Principle for Identifiable Disentangled Representations
by: Matthes, Stefan, et al.
Published: (2025)
by: Matthes, Stefan, et al.
Published: (2025)
Differentiable Annealed Importance Sampling Minimizes The Symmetrized Kullback-Leibler Divergence Between Initial and Target Distribution
by: Zenn, Johannes, et al.
Published: (2024)
by: Zenn, Johannes, et al.
Published: (2024)
Exploring the Impact of Dataset Statistical Effect Size on Model Performance and Data Sample Size Sufficiency
by: Hatamian, Arya, et al.
Published: (2025)
by: Hatamian, Arya, et al.
Published: (2025)
Structural Disentanglement of Causal and Correlated Concepts
by: Zhao, Qilong, et al.
Published: (2024)
by: Zhao, Qilong, et al.
Published: (2024)
Embedding principle of homogeneous neural network for classification problem
by: Zhang, Jiahan, et al.
Published: (2025)
by: Zhang, Jiahan, et al.
Published: (2025)
Early Neuron Alignment in Two-layer ReLU Networks with Small Initialization
by: Min, Hancheng, et al.
Published: (2023)
by: Min, Hancheng, et al.
Published: (2023)
Targeted Neuron Modulation via Contrastive Pair Search
by: Herring, Sam, et al.
Published: (2026)
by: Herring, Sam, et al.
Published: (2026)
Linear Independence of Generalized Neurons and Related Functions
by: Zhang, Leyang
Published: (2024)
by: Zhang, Leyang
Published: (2024)
Disentangled Hyperbolic Representation Learning for Heterogeneous Graphs
by: Bai, Qijie, et al.
Published: (2024)
by: Bai, Qijie, et al.
Published: (2024)
Effective Sample Size and Generalization Bounds for Temporal Networks
by: Gahtan, Barak, et al.
Published: (2025)
by: Gahtan, Barak, et al.
Published: (2025)
Geometry of Critical Sets and Existence of Saddle Branches for Two-layer Neural Networks
by: Zhang, Leyang, et al.
Published: (2024)
by: Zhang, Leyang, et al.
Published: (2024)
Geometry and Local Recovery of Global Minima of Two-layer Neural Networks at Overparameterization
by: Zhang, Leyang, et al.
Published: (2023)
by: Zhang, Leyang, et al.
Published: (2023)
NeuronSeek: On Stability and Expressivity of Task-driven Neurons
by: Pei, Hanyu, et al.
Published: (2025)
by: Pei, Hanyu, et al.
Published: (2025)
No One-Size-Fits-All Neurons: Task-based Neurons for Artificial Neural Networks
by: Fan, Feng-Lei, et al.
Published: (2024)
by: Fan, Feng-Lei, et al.
Published: (2024)
Neural Force Field: Few-shot Learning of Generalized Physical Reasoning
by: Li, Shiqian, et al.
Published: (2025)
by: Li, Shiqian, et al.
Published: (2025)
Overview frequency principle/spectral bias in deep learning
by: Xu, Zhi-Qin John, et al.
Published: (2022)
by: Xu, Zhi-Qin John, et al.
Published: (2022)
An overview of condensation phenomenon in deep learning
by: Xu, Zhi-Qin John, et al.
Published: (2025)
by: Xu, Zhi-Qin John, et al.
Published: (2025)
Determinism in the Undetermined: Deterministic Output in Charge-Conserving Continuous-Time Neuromorphic Systems with Temporal Stochasticity
by: Yan, Jing, et al.
Published: (2026)
by: Yan, Jing, et al.
Published: (2026)
To See a World in a Spark of Neuron: Disentangling Multi-task Interference for Training-free Model Merging
by: Fang, Zitao, et al.
Published: (2025)
by: Fang, Zitao, et al.
Published: (2025)
Disentanglement of Variations with Multimodal Generative Modeling
by: Zhang, Yijie, et al.
Published: (2025)
by: Zhang, Yijie, et al.
Published: (2025)
Improving Neuron-level Interpretability with White-box Language Models
by: Bai, Hao, et al.
Published: (2024)
by: Bai, Hao, et al.
Published: (2024)
On the Effects of Irrelevant Variables in Treatment Effect Estimation with Deep Disentanglement
by: Khan, Ahmad Saeed, et al.
Published: (2024)
by: Khan, Ahmad Saeed, et al.
Published: (2024)
IT-OSE: Exploring Optimal Sample Size for Industrial Data Augmentation
by: Sun, Mingchun, et al.
Published: (2026)
by: Sun, Mingchun, et al.
Published: (2026)
Robust Thompson Sampling Algorithms Against Reward Poisoning Attacks
by: Xu, Yinglun, et al.
Published: (2024)
by: Xu, Yinglun, et al.
Published: (2024)
SafeNeuron: Neuron-Level Safety Alignment for Large Language Models
by: Wang, Zhaoxin, et al.
Published: (2026)
by: Wang, Zhaoxin, et al.
Published: (2026)
Light Alignment Improves LLM Safety via Model Self-Reflection with a Single Neuron
by: Shen, Sicheng, et al.
Published: (2026)
by: Shen, Sicheng, et al.
Published: (2026)
A Bayesian Model for Online Activity Sample Sizes
by: Richardson, Thomas, et al.
Published: (2021)
by: Richardson, Thomas, et al.
Published: (2021)
On Size-Independent Sample Complexity of ReLU Networks
by: Sellke, Mark
Published: (2023)
by: Sellke, Mark
Published: (2023)
Disentangling Polysemantic Neurons with a Null-Calibrated Polysemanticity Index and Causal Patch Interventions
by: Gupta, Manan, et al.
Published: (2025)
by: Gupta, Manan, et al.
Published: (2025)
A Critical Look at Targeted Instruction Selection: Disentangling What Matters (and What Doesn't)
by: Nayak, Nihal V., et al.
Published: (2026)
by: Nayak, Nihal V., et al.
Published: (2026)
Similar Items
-
Connectivity Shapes Implicit Regularization in Matrix Factorization Models for Matrix Completion
by: Bai, Zhiwei, et al.
Published: (2024) -
Local Linear Recovery Guarantee of Deep Neural Networks at Overparameterization
by: Zhang, Yaoyu, et al.
Published: (2024) -
Initialization is Critical to Whether Transformers Fit Composite Functions by Reasoning or Memorizing
by: Zhang, Zhongwang, et al.
Published: (2024) -
Embedding Principle in Depth for the Loss Landscape Analysis of Deep Neural Networks
by: Bai, Zhiwei, et al.
Published: (2022) -
Adaptive Preconditioners Trigger Loss Spikes in Adam
by: Bai, Zhiwei, et al.
Published: (2025)