Salvato in:
| Autori principali: | Lyu, Shen-Huan, Wu, Jin-Hui, Zheng, Qin-Cheng, Ye, Baoliu |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | https://arxiv.org/abs/2407.05108 |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Enhance and Reuse: A Dual-Mechanism Approach to Boost Deep Forest for Label Distribution Learning
di: Xu, Jia-Le, et al.
Pubblicazione: (2026)
di: Xu, Jia-Le, et al.
Pubblicazione: (2026)
Mask-Encoded Sparsification: Mitigating Biased Gradients in Communication-Efficient Split Learning
di: Zhou, Wenxuan, et al.
Pubblicazione: (2024)
di: Zhou, Wenxuan, et al.
Pubblicazione: (2024)
Compressing Model with Few Class-Imbalance Samples: An Out-of-Distribution Expedition
di: Wu, Tian-Shuang, et al.
Pubblicazione: (2025)
di: Wu, Tian-Shuang, et al.
Pubblicazione: (2025)
Enhance Learning Efficiency of Oblique Decision Tree via Feature Concatenation
di: Lyu, Shen-Huan, et al.
Pubblicazione: (2025)
di: Lyu, Shen-Huan, et al.
Pubblicazione: (2025)
Theoretical Investigation on Inductive Bias of Isolation Forest
di: Zheng, Qin-Cheng, et al.
Pubblicazione: (2025)
di: Zheng, Qin-Cheng, et al.
Pubblicazione: (2025)
Breaking the Prototype Bias Loop: Confidence-Aware Federated Contrastive Learning for Highly Imbalanced Clients
di: Wu, Tian-Shuang, et al.
Pubblicazione: (2026)
di: Wu, Tian-Shuang, et al.
Pubblicazione: (2026)
On the Role of Depth in the Expressivity of RNNs
di: Lizaire, Maude, et al.
Pubblicazione: (2026)
di: Lizaire, Maude, et al.
Pubblicazione: (2026)
The WidthWall: A Strict Expressivity Hierarchy for Hypergraph Neural Networks
di: Jiang, Fengqing, et al.
Pubblicazione: (2026)
di: Jiang, Fengqing, et al.
Pubblicazione: (2026)
Spectral Condition for $μ$P under Width-Depth Scaling
di: Zheng, Chenyu, et al.
Pubblicazione: (2026)
di: Zheng, Chenyu, et al.
Pubblicazione: (2026)
The Effect of Depth on the Expressivity of Deep Linear State-Space Models
di: Bao, Zeyu, et al.
Pubblicazione: (2025)
di: Bao, Zeyu, et al.
Pubblicazione: (2025)
Mixture of Universal Experts: Scaling Virtual Width via Depth-Width Transformation
di: Chen, Yilong, et al.
Pubblicazione: (2026)
di: Chen, Yilong, et al.
Pubblicazione: (2026)
Sparsity is Combinatorial Depth: Quantifying MoE Expressivity via Tropical Geometry
di: Su, Ye, et al.
Pubblicazione: (2026)
di: Su, Ye, et al.
Pubblicazione: (2026)
Improving Generalization of Deep Neural Networks by Leveraging Margin Distribution
di: Lyu, Shen-Huan, et al.
Pubblicazione: (2018)
di: Lyu, Shen-Huan, et al.
Pubblicazione: (2018)
Depth Separation in Norm-Bounded Infinite-Width Neural Networks
di: Parkinson, Suzanna, et al.
Pubblicazione: (2024)
di: Parkinson, Suzanna, et al.
Pubblicazione: (2024)
On Expressivity of Height in Neural Networks
di: Fan, Feng-Lei, et al.
Pubblicazione: (2023)
di: Fan, Feng-Lei, et al.
Pubblicazione: (2023)
AdaPerceiver: Transformers with Adaptive Width, Depth, and Tokens
di: Jajal, Purvish, et al.
Pubblicazione: (2025)
di: Jajal, Purvish, et al.
Pubblicazione: (2025)
Virtual Width Networks
di: Seed, et al.
Pubblicazione: (2025)
di: Seed, et al.
Pubblicazione: (2025)
Deep Linear Network Training Dynamics from Random Initialization: Data, Width, Depth, and Hyperparameter Transfer
di: Bordelon, Blake, et al.
Pubblicazione: (2025)
di: Bordelon, Blake, et al.
Pubblicazione: (2025)
Completed Hyperparameter Transfer across Modules, Width, Depth, Batch and Duration
di: Mlodozeniec, Bruno, et al.
Pubblicazione: (2025)
di: Mlodozeniec, Bruno, et al.
Pubblicazione: (2025)
Depth-Width tradeoffs in Algorithmic Reasoning of Graph Tasks with Transformers
di: Yehudai, Gilad, et al.
Pubblicazione: (2025)
di: Yehudai, Gilad, et al.
Pubblicazione: (2025)
Minimum Width of Deep Narrow Networks for Universal Approximation
di: Yang, Xiao-Song, et al.
Pubblicazione: (2025)
di: Yang, Xiao-Song, et al.
Pubblicazione: (2025)
Pause Tokens Strictly Increase the Expressivity of Constant-Depth Transformers
di: London, Charles, et al.
Pubblicazione: (2025)
di: London, Charles, et al.
Pubblicazione: (2025)
On the Expressiveness of Rational ReLU Neural Networks With Bounded Depth
di: Averkov, Gennadiy, et al.
Pubblicazione: (2025)
di: Averkov, Gennadiy, et al.
Pubblicazione: (2025)
A Little Depth Goes a Long Way: The Expressive Power of Log-Depth Transformers
di: Merrill, William, et al.
Pubblicazione: (2025)
di: Merrill, William, et al.
Pubblicazione: (2025)
Expressivity of Transformers: A Tropical Geometry Perspective
di: Su, Ye, et al.
Pubblicazione: (2026)
di: Su, Ye, et al.
Pubblicazione: (2026)
Theory of Scaling Laws for In-Context Regression: Depth, Width, Context and Time
di: Bordelon, Blake, et al.
Pubblicazione: (2025)
di: Bordelon, Blake, et al.
Pubblicazione: (2025)
On the Infinite Width and Depth Limits of Predictive Coding Networks
di: Innocenti, Francesco, et al.
Pubblicazione: (2026)
di: Innocenti, Francesco, et al.
Pubblicazione: (2026)
DeepMartingale: Duality of the Optimal Stopping Problem with Expressivity and High-Dimensional Hedging
di: Ye, Junyan, et al.
Pubblicazione: (2025)
di: Ye, Junyan, et al.
Pubblicazione: (2025)
Network Threat Detection: Addressing Class Imbalanced Data with Deep Forest
di: Chen, Jiaqi, et al.
Pubblicazione: (2025)
di: Chen, Jiaqi, et al.
Pubblicazione: (2025)
Transformers are Expressive, But Are They Expressive Enough for Regression?
di: Nath, Swaroop, et al.
Pubblicazione: (2024)
di: Nath, Swaroop, et al.
Pubblicazione: (2024)
Embedding Principle in Depth for the Loss Landscape Analysis of Deep Neural Networks
di: Bai, Zhiwei, et al.
Pubblicazione: (2022)
di: Bai, Zhiwei, et al.
Pubblicazione: (2022)
GEVRM: Goal-Expressive Video Generation Model For Robust Visual Manipulation
di: Zhang, Hongyin, et al.
Pubblicazione: (2025)
di: Zhang, Hongyin, et al.
Pubblicazione: (2025)
The Median is Easier than it Looks: Approximation with a Constant-Depth, Linear-Width ReLU Network
di: Dutta, Abhigyan, et al.
Pubblicazione: (2026)
di: Dutta, Abhigyan, et al.
Pubblicazione: (2026)
Riemann-Lebesgue Forest for Regression
di: Qin, Tian, et al.
Pubblicazione: (2024)
di: Qin, Tian, et al.
Pubblicazione: (2024)
Expressive Power of Deep Networks on Manifolds: Simultaneous Approximation
di: Zhou, Hanfei, et al.
Pubblicazione: (2025)
di: Zhou, Hanfei, et al.
Pubblicazione: (2025)
A Provable Expressiveness Hierarchy in Hybrid Linear-Full Attention
di: Ye, Xiaowei, et al.
Pubblicazione: (2026)
di: Ye, Xiaowei, et al.
Pubblicazione: (2026)
Identifiable Bayesian Deep Generative Copulas with Unknown Layer Widths for Data with Arbitrary Marginal Distributions
di: Feldman, Joseph, et al.
Pubblicazione: (2026)
di: Feldman, Joseph, et al.
Pubblicazione: (2026)
Constructing Enhanced Mutual Information for Online Class-Incremental Learning
di: Zhang, Huan, et al.
Pubblicazione: (2024)
di: Zhang, Huan, et al.
Pubblicazione: (2024)
DOFEN: Deep Oblivious Forest ENsemble
di: Chen, Kuan-Yu, et al.
Pubblicazione: (2024)
di: Chen, Kuan-Yu, et al.
Pubblicazione: (2024)
Post-LayerNorm Is Back: Stable, ExpressivE, and Deep
di: Chen, Chen, et al.
Pubblicazione: (2026)
di: Chen, Chen, et al.
Pubblicazione: (2026)
Documenti analoghi
-
Enhance and Reuse: A Dual-Mechanism Approach to Boost Deep Forest for Label Distribution Learning
di: Xu, Jia-Le, et al.
Pubblicazione: (2026) -
Mask-Encoded Sparsification: Mitigating Biased Gradients in Communication-Efficient Split Learning
di: Zhou, Wenxuan, et al.
Pubblicazione: (2024) -
Compressing Model with Few Class-Imbalance Samples: An Out-of-Distribution Expedition
di: Wu, Tian-Shuang, et al.
Pubblicazione: (2025) -
Enhance Learning Efficiency of Oblique Decision Tree via Feature Concatenation
di: Lyu, Shen-Huan, et al.
Pubblicazione: (2025) -
Theoretical Investigation on Inductive Bias of Isolation Forest
di: Zheng, Qin-Cheng, et al.
Pubblicazione: (2025)