Mixed Dynamics In Linear Networks: Unifying the Lazy and Active Regimes
Fuente:
arXiv
Saved in:
| Main Authors: | Tu, Zhenfeng, Aranguri, Santiago, Jacot, Arthur |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Probe-Based Data Attribution: Discovering and Mitigating Undesirable Behaviors in LLM Post-Training
by: Xiao, Frank, et al.
Published: (2026)
by: Xiao, Frank, et al.
Published: (2026)
Bottleneck Structure in Learned Features: Low-Dimension vs Regularity Tradeoff
by: Jacot, Arthur
Published: (2023)
by: Jacot, Arthur
Published: (2023)
Saddle-To-Saddle Dynamics in Deep ReLU Networks: Low-Rank Bias in the First Saddle Escape
by: Bantzis, Ioannis, et al.
Published: (2025)
by: Bantzis, Ioannis, et al.
Published: (2025)
Hamiltonian Mechanics of Feature Learning: Bottleneck Structure in Leaky ResNets
by: Jacot, Arthur, et al.
Published: (2024)
by: Jacot, Arthur, et al.
Published: (2024)
Which Frequencies do CNNs Need? Emergent Bottleneck Structure in Feature Learning
by: Wen, Yuxiao, et al.
Published: (2024)
by: Wen, Yuxiao, et al.
Published: (2024)
Inference-Time Toxicity Mitigation in Protein Language Models
by: Burda, Manuel Fernández, et al.
Published: (2026)
by: Burda, Manuel Fernández, et al.
Published: (2026)
How DNNs break the Curse of Dimensionality: Compositionality and Symmetry Learning
by: Jacot, Arthur, et al.
Published: (2024)
by: Jacot, Arthur, et al.
Published: (2024)
LazyDiT: Lazy Learning for the Acceleration of Diffusion Transformers
by: Shen, Xuan, et al.
Published: (2024)
by: Shen, Xuan, et al.
Published: (2024)
Diffmv: A Unified Diffusion Framework for Healthcare Predictions with Random Missing Views and View Laziness
by: Zhao, Chuang, et al.
Published: (2025)
by: Zhao, Chuang, et al.
Published: (2025)
The Importance of Being Lazy: Scaling Limits of Continual Learning
by: Graldi, Jacopo, et al.
Published: (2025)
by: Graldi, Jacopo, et al.
Published: (2025)
Lazy But Effective: Collaborative Personalized Federated Learning with Heterogeneous Data
by: Rokvic, Ljubomir, et al.
Published: (2025)
by: Rokvic, Ljubomir, et al.
Published: (2025)
LazyLLM: Dynamic Token Pruning for Efficient Long Context LLM Inference
by: Fu, Qichen, et al.
Published: (2024)
by: Fu, Qichen, et al.
Published: (2024)
Fast Sampling for Flows and Diffusions with Lazy and Point Mass Stochastic Interpolants
by: Damsholt, Gabriel, et al.
Published: (2026)
by: Damsholt, Gabriel, et al.
Published: (2026)
Laziness, Barren Plateau, and Noise in Machine Learning
by: Liu, Junyu, et al.
Published: (2022)
by: Liu, Junyu, et al.
Published: (2022)
Phase-aware Training Schedule Simplifies Learning in Flow-Based Generative Models
by: Aranguri, Santiago, et al.
Published: (2024)
by: Aranguri, Santiago, et al.
Published: (2024)
Dynamic Graph Condensation
by: Chen, Dong, et al.
Published: (2025)
by: Chen, Dong, et al.
Published: (2025)
ORLA*: Mobile Manipulator-Based Object Rearrangement with Lazy A Star
by: Gao, Kai, et al.
Published: (2023)
by: Gao, Kai, et al.
Published: (2023)
SLoPe: Double-Pruned Sparse Plus Lazy Low-Rank Adapter Pretraining of LLMs
by: Mozaffari, Mohammad, et al.
Published: (2024)
by: Mozaffari, Mohammad, et al.
Published: (2024)
SiGNN: A Spike-induced Graph Neural Network for Dynamic Graph Representation Learning
by: Chen, Dong, et al.
Published: (2024)
by: Chen, Dong, et al.
Published: (2024)
Reinforcement Learning in Dynamic Treatment Regimes Needs Critical Reexamination
by: Luo, Zhiyao, et al.
Published: (2024)
by: Luo, Zhiyao, et al.
Published: (2024)
On the Learning Dynamics of Two-layer Linear Networks with Label Noise SGD
by: Zhang, Tongcheng, et al.
Published: (2026)
by: Zhang, Tongcheng, et al.
Published: (2026)
Dynamic Meta-Learning for Adaptive XGBoost-Neural Ensembles
by: Sedek, Arthur
Published: (2025)
by: Sedek, Arthur
Published: (2025)
Linear Transformers Implicitly Discover Unified Numerical Algorithms
by: Lutz, Patrick, et al.
Published: (2025)
by: Lutz, Patrick, et al.
Published: (2025)
Dynamical Systems Analysis Reveals Functional Regimes in Large Language Models
by: Ugail, Hassan, et al.
Published: (2026)
by: Ugail, Hassan, et al.
Published: (2026)
Unveiling the Training Dynamics of ReLU Networks through a Linear Lens
by: Ye, Longqing
Published: (2025)
by: Ye, Longqing
Published: (2025)
A Unifying View of Coverage in Linear Off-Policy Evaluation
by: Amortila, Philip, et al.
Published: (2026)
by: Amortila, Philip, et al.
Published: (2026)
Neural Network Optimal Power Flow via Energy Gradient Flow and Unified Dynamics
by: Liu, Xuezhi
Published: (2025)
by: Liu, Xuezhi
Published: (2025)
Graph Mixing Additive Networks
by: Bechler-Speicher, Maya, et al.
Published: (2025)
by: Bechler-Speicher, Maya, et al.
Published: (2025)
MetaLA: Unified Optimal Linear Approximation to Softmax Attention Map
by: Chou, Yuhong, et al.
Published: (2024)
by: Chou, Yuhong, et al.
Published: (2024)
Scalable Production Scheduling: Linear Complexity via Unified Homogeneous Graphs
by: Hoss, Jonathan, et al.
Published: (2026)
by: Hoss, Jonathan, et al.
Published: (2026)
Fast and Interpretable Mixed-Integer Linear Program Solving by Learning Model Reduction
by: Li, Yixuan, et al.
Published: (2024)
by: Li, Yixuan, et al.
Published: (2024)
UniGeM: Unifying Data Mixing and Selection via Geometric Exploration and Mining
by: Wang, Changhao, et al.
Published: (2026)
by: Wang, Changhao, et al.
Published: (2026)
An Active Diffusion Neural Network for Graphs
by: Jiang, Mengying
Published: (2025)
by: Jiang, Mengying
Published: (2025)
DTR-Bench: An in silico Environment and Benchmark Platform for Reinforcement Learning Based Dynamic Treatment Regime
by: Luo, Zhiyao, et al.
Published: (2024)
by: Luo, Zhiyao, et al.
Published: (2024)
RegimeNAS: Regime-Aware Differentiable Architecture Search With Theoretical Guarantees for Financial Trading
by: Devadiga, Prathamesh, et al.
Published: (2025)
by: Devadiga, Prathamesh, et al.
Published: (2025)
Three Quantization Regimes for ReLU Networks
by: Ou, Weigutian, et al.
Published: (2024)
by: Ou, Weigutian, et al.
Published: (2024)
Adversarial Graph Disentanglement
by: Zheng, Shuai, et al.
Published: (2021)
by: Zheng, Shuai, et al.
Published: (2021)
Novel Kernel Models and Exact Representor Theory for Neural Networks Beyond the Over-Parameterized Regime
by: Shilton, Alistair, et al.
Published: (2024)
by: Shilton, Alistair, et al.
Published: (2024)
Polynomial Speedup in Diffusion Models with the Multilevel Euler-Maruyama Method
by: Jacot, Arthur
Published: (2026)
by: Jacot, Arthur
Published: (2026)
Deep Learning as a Convex Paradigm of Computation: Minimizing Circuit Size with ResNets
by: Jacot, Arthur
Published: (2025)
by: Jacot, Arthur
Published: (2025)
Similar Items
-
Probe-Based Data Attribution: Discovering and Mitigating Undesirable Behaviors in LLM Post-Training
by: Xiao, Frank, et al.
Published: (2026) -
Bottleneck Structure in Learned Features: Low-Dimension vs Regularity Tradeoff
by: Jacot, Arthur
Published: (2023) -
Saddle-To-Saddle Dynamics in Deep ReLU Networks: Low-Rank Bias in the First Saddle Escape
by: Bantzis, Ioannis, et al.
Published: (2025) -
Hamiltonian Mechanics of Feature Learning: Bottleneck Structure in Leaky ResNets
by: Jacot, Arthur, et al.
Published: (2024) -
Which Frequencies do CNNs Need? Emergent Bottleneck Structure in Feature Learning
by: Wen, Yuxiao, et al.
Published: (2024)