Salvato in:
| Autori principali: | Li, Yuqing, Luo, Tao, Zhou, Qixuan |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | https://arxiv.org/abs/2404.04859 |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Quantifying Training Difficulty and Accelerating Convergence in Neural Network-Based PDE Solvers
di: Chen, Chuqi, et al.
Pubblicazione: (2024)
di: Chen, Chuqi, et al.
Pubblicazione: (2024)
Demystifying Distributed Training of Graph Neural Networks for Link Prediction
di: Huang, Xin, et al.
Pubblicazione: (2025)
di: Huang, Xin, et al.
Pubblicazione: (2025)
A priori Estimates for Deep Residual Network in Continuous-time Reinforcement Learning
di: Yin, Shuyu, et al.
Pubblicazione: (2024)
di: Yin, Shuyu, et al.
Pubblicazione: (2024)
LazyDiT: Lazy Learning for the Acceleration of Diffusion Transformers
di: Shen, Xuan, et al.
Pubblicazione: (2024)
di: Shen, Xuan, et al.
Pubblicazione: (2024)
ProPINN: Demystifying Propagation Failures in Physics-Informed Neural Networks
di: Wu, Haixu, et al.
Pubblicazione: (2025)
di: Wu, Haixu, et al.
Pubblicazione: (2025)
Residual Attention Physics-Informed Neural Networks for Robust Multiphysics Simulation of Steady-State Electrothermal Energy Systems
di: Zhou, Yuqing, et al.
Pubblicazione: (2026)
di: Zhou, Yuqing, et al.
Pubblicazione: (2026)
Grokking as the Transition from Lazy to Rich Training Dynamics
di: Kumar, Tanishq, et al.
Pubblicazione: (2023)
di: Kumar, Tanishq, et al.
Pubblicazione: (2023)
Demystifying Higher-Order Graph Neural Networks
di: Besta, Maciej, et al.
Pubblicazione: (2024)
di: Besta, Maciej, et al.
Pubblicazione: (2024)
Demystifying Oversmoothing in Attention-Based Graph Neural Networks
di: Wu, Xinyi, et al.
Pubblicazione: (2023)
di: Wu, Xinyi, et al.
Pubblicazione: (2023)
Convergence of Stochastic Gradient Langevin Dynamics in the Lazy Training Regime
di: Oberweis, Noah, et al.
Pubblicazione: (2025)
di: Oberweis, Noah, et al.
Pubblicazione: (2025)
PRISM: Demystifying Retention and Interaction in Mid-Training
di: Runwal, Bharat, et al.
Pubblicazione: (2026)
di: Runwal, Bharat, et al.
Pubblicazione: (2026)
AutoSGNN: Automatic Propagation Mechanism Discovery for Spectral Graph Neural Networks
di: Mo, Shibing, et al.
Pubblicazione: (2024)
di: Mo, Shibing, et al.
Pubblicazione: (2024)
Has the Deep Neural Network learned the Stochastic Process? An Evaluation Viewpoint
di: Kumar, Harshit, et al.
Pubblicazione: (2024)
di: Kumar, Harshit, et al.
Pubblicazione: (2024)
Phase Diagram of Initial Condensation for Two-layer Neural Networks
di: Chen, Zhengan, et al.
Pubblicazione: (2023)
di: Chen, Zhengan, et al.
Pubblicazione: (2023)
From Lazy to Rich: Exact Learning Dynamics in Deep Linear Networks
di: Dominé, Clémentine C. J., et al.
Pubblicazione: (2024)
di: Dominé, Clémentine C. J., et al.
Pubblicazione: (2024)
Lazy FSCA for Unsupervised Variable Selection
di: Zocco, Federico, et al.
Pubblicazione: (2021)
di: Zocco, Federico, et al.
Pubblicazione: (2021)
Mixed Dynamics In Linear Networks: Unifying the Lazy and Active Regimes
di: Tu, Zhenfeng, et al.
Pubblicazione: (2024)
di: Tu, Zhenfeng, et al.
Pubblicazione: (2024)
Enhancing Trustworthiness of Graph Neural Networks with Rank-Based Conformal Training
di: Wang, Ting, et al.
Pubblicazione: (2025)
di: Wang, Ting, et al.
Pubblicazione: (2025)
Demystifying Network Foundation Models
di: Beltiukov, Sylee, et al.
Pubblicazione: (2025)
di: Beltiukov, Sylee, et al.
Pubblicazione: (2025)
DR-CircuitGNN: Training Acceleration of Heterogeneous Circuit Graph Neural Network on GPUs
di: Luo, Yuebo, et al.
Pubblicazione: (2025)
di: Luo, Yuebo, et al.
Pubblicazione: (2025)
Provable Acceleration of Nesterov's Accelerated Gradient for Rectangular Matrix Factorization and Linear Neural Networks
di: Xu, Zhenghao, et al.
Pubblicazione: (2024)
di: Xu, Zhenghao, et al.
Pubblicazione: (2024)
APEX: Probing Neural Networks via Activation Perturbation
di: Ren, Tao, et al.
Pubblicazione: (2026)
di: Ren, Tao, et al.
Pubblicazione: (2026)
Partially Lazy Gradient Descent for Smoothed Online Learning
di: Mhaisen, Naram, et al.
Pubblicazione: (2026)
di: Mhaisen, Naram, et al.
Pubblicazione: (2026)
John Ellipsoids via Lazy Updates
di: Woodruff, David P., et al.
Pubblicazione: (2025)
di: Woodruff, David P., et al.
Pubblicazione: (2025)
Sharpened Lazy Incremental Quasi-Newton Method
di: Lahoti, Aakash, et al.
Pubblicazione: (2023)
di: Lahoti, Aakash, et al.
Pubblicazione: (2023)
Why Are DMD Students Lazy? Understanding the Copying Behavior in Few-Step Distillation
di: Li, Shucheng, et al.
Pubblicazione: (2026)
di: Li, Shucheng, et al.
Pubblicazione: (2026)
Loss Spike in Training Neural Networks
di: Li, Xiaolong, et al.
Pubblicazione: (2023)
di: Li, Xiaolong, et al.
Pubblicazione: (2023)
LazyDP: Co-Designing Algorithm-Software for Scalable Training of Differentially Private Recommendation Models
di: Lim, Juntaek, et al.
Pubblicazione: (2024)
di: Lim, Juntaek, et al.
Pubblicazione: (2024)
Demystify Protein Generation with Hierarchical Conditional Diffusion Models
di: Ling, Zinan, et al.
Pubblicazione: (2025)
di: Ling, Zinan, et al.
Pubblicazione: (2025)
Lazy Data Practices Harm Fairness Research
di: Simson, Jan, et al.
Pubblicazione: (2024)
di: Simson, Jan, et al.
Pubblicazione: (2024)
Three-Class Text Sentiment Analysis Based on LSTM
di: Qixuan, Yin
Pubblicazione: (2024)
di: Qixuan, Yin
Pubblicazione: (2024)
Demystifying Prediction Powered Inference
di: Song, Yilin, et al.
Pubblicazione: (2026)
di: Song, Yilin, et al.
Pubblicazione: (2026)
Uncertainty in Graph Neural Networks: A Survey
di: Wang, Fangxin, et al.
Pubblicazione: (2024)
di: Wang, Fangxin, et al.
Pubblicazione: (2024)
The Importance of Being Lazy: Scaling Limits of Continual Learning
di: Graldi, Jacopo, et al.
Pubblicazione: (2025)
di: Graldi, Jacopo, et al.
Pubblicazione: (2025)
Flow Matching from Viewpoint of Proximal Operators
di: Fukumizu, Kenji, et al.
Pubblicazione: (2026)
di: Fukumizu, Kenji, et al.
Pubblicazione: (2026)
In-Context Linear Regression Demystified: Training Dynamics and Mechanistic Interpretability of Multi-Head Softmax Attention
di: He, Jianliang, et al.
Pubblicazione: (2025)
di: He, Jianliang, et al.
Pubblicazione: (2025)
Frequency-adaptive Multi-scale Deep Neural Networks
di: Huang, Jizu, et al.
Pubblicazione: (2024)
di: Huang, Jizu, et al.
Pubblicazione: (2024)
BSFA: Leveraging the Subspace Dichotomy to Accelerate Neural Network Training
di: Zhou, Wenjie, et al.
Pubblicazione: (2025)
di: Zhou, Wenjie, et al.
Pubblicazione: (2025)
Uncovering Critical Sets of Deep Neural Networks via Sample-Independent Critical Lifting
di: Zhang, Leyang, et al.
Pubblicazione: (2025)
di: Zhang, Leyang, et al.
Pubblicazione: (2025)
Policy Compatible Skill Incremental Learning via Lazy Learning Interface
di: Lee, Daehee, et al.
Pubblicazione: (2025)
di: Lee, Daehee, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Quantifying Training Difficulty and Accelerating Convergence in Neural Network-Based PDE Solvers
di: Chen, Chuqi, et al.
Pubblicazione: (2024) -
Demystifying Distributed Training of Graph Neural Networks for Link Prediction
di: Huang, Xin, et al.
Pubblicazione: (2025) -
A priori Estimates for Deep Residual Network in Continuous-time Reinforcement Learning
di: Yin, Shuyu, et al.
Pubblicazione: (2024) -
LazyDiT: Lazy Learning for the Acceleration of Diffusion Transformers
di: Shen, Xuan, et al.
Pubblicazione: (2024) -
ProPINN: Demystifying Propagation Failures in Physics-Informed Neural Networks
di: Wu, Haixu, et al.
Pubblicazione: (2025)