Mask in the Mirror: Implicit Sparsification
Fuente:
arXiv
Saved in:
| Main Authors: | Jacobs, Tom, Burkholz, Rebekka |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Mirror, Mirror of the Flow: How Does Regularization Shape Implicit Bias?
by: Jacobs, Tom, et al.
Published: (2025)
by: Jacobs, Tom, et al.
Published: (2025)
Never Saddle for Reparameterized Steepest Descent as Mirror Flow
by: Jacobs, Tom, et al.
Published: (2026)
by: Jacobs, Tom, et al.
Published: (2026)
Hyperbolic Aware Minimization: Implicit Bias for Sparsity
by: Jacobs, Tom, et al.
Published: (2025)
by: Jacobs, Tom, et al.
Published: (2025)
HORST: Composing Optimizer Geometries for Sparse Transformer Training
by: Jacobs, Tom, et al.
Published: (2026)
by: Jacobs, Tom, et al.
Published: (2026)
Masks, Signs, And Learning Rate Rewinding
by: Gadhikar, Advait, et al.
Published: (2024)
by: Gadhikar, Advait, et al.
Published: (2024)
Pay Attention to Small Weights
by: Zhou, Chao, et al.
Published: (2025)
by: Zhou, Chao, et al.
Published: (2025)
Sign-In to the Lottery: Reparameterizing Sparse Training From Scratch
by: Gadhikar, Advait, et al.
Published: (2025)
by: Gadhikar, Advait, et al.
Published: (2025)
GATE: How to Keep Out Intrusive Neighbors
by: Mustafa, Nimrah, et al.
Published: (2024)
by: Mustafa, Nimrah, et al.
Published: (2024)
The Graphon Limit Hypothesis: Understanding Neural Network Pruning via Infinite Width Analysis
by: Pham, Hoang, et al.
Published: (2025)
by: Pham, Hoang, et al.
Published: (2025)
Implicit Bias of Mirror Flow in Homogeneous Neural Networks: Sparse and Dense Feature Learning
by: Jacobs, Tom, et al.
Published: (2026)
by: Jacobs, Tom, et al.
Published: (2026)
Fixed Aggregation Features Can Rival GNNs
by: Rubio-Madrigal, Celia, et al.
Published: (2026)
by: Rubio-Madrigal, Celia, et al.
Published: (2026)
Robustness of Mixtures of Experts to Feature Noise
by: Sun, Dong, et al.
Published: (2026)
by: Sun, Dong, et al.
Published: (2026)
SparseOpt: Addressing Normalization-induced Gradient Skew in Sparse Training
by: Adnan, Mohammed, et al.
Published: (2026)
by: Adnan, Mohammed, et al.
Published: (2026)
Spectral Graph Pruning Against Over-Squashing and Over-Smoothing
by: Jamadandi, Adarsh, et al.
Published: (2024)
by: Jamadandi, Adarsh, et al.
Published: (2024)
GNNs Getting ComFy: Community and Feature Similarity Guided Rewiring
by: Rubio-Madrigal, Celia, et al.
Published: (2025)
by: Rubio-Madrigal, Celia, et al.
Published: (2025)
Cyclic Sparse Training: Is it Enough?
by: Gadhikar, Advait, et al.
Published: (2024)
by: Gadhikar, Advait, et al.
Published: (2024)
Multi-Agent Systems are Mixtures of Experts: Who Becomes an Influencer?
by: Bause, Franka, et al.
Published: (2026)
by: Bause, Franka, et al.
Published: (2026)
Pruning neural network models for gene regulatory dynamics using data and domain knowledge
by: Hossain, Intekhab, et al.
Published: (2024)
by: Hossain, Intekhab, et al.
Published: (2024)
When Shift Happens - Confounding Is to Blame
by: Reddy, Abbavaram Gowtham, et al.
Published: (2025)
by: Reddy, Abbavaram Gowtham, et al.
Published: (2025)
Frequency-Based Hyperparameter Selection in Games
by: Sanyal, Aniket, et al.
Published: (2026)
by: Sanyal, Aniket, et al.
Published: (2026)
Bridging Domains through Subspace-Aware Model Merging
by: Chaves, Levy, et al.
Published: (2026)
by: Chaves, Levy, et al.
Published: (2026)
Implicit Bias of Mirror Flow on Separable Data
by: Pesme, Scott, et al.
Published: (2024)
by: Pesme, Scott, et al.
Published: (2024)
SparseForge: Efficient Semi-Structured LLM Sparsification via Annealing of Hessian-Guided Soft-Mask
by: Hanzuo, Liu, et al.
Published: (2026)
by: Hanzuo, Liu, et al.
Published: (2026)
Implicit Bias and Convergence of Matrix Stochastic Mirror Descent
by: Akhtiamov, Danil, et al.
Published: (2026)
by: Akhtiamov, Danil, et al.
Published: (2026)
Mask-Encoded Sparsification: Mitigating Biased Gradients in Communication-Efficient Split Learning
by: Zhou, Wenxuan, et al.
Published: (2024)
by: Zhou, Wenxuan, et al.
Published: (2024)
Implicit Bias of Mirror Flow for Shallow Neural Networks in Univariate Regression
by: Liang, Shuang, et al.
Published: (2024)
by: Liang, Shuang, et al.
Published: (2024)
A Unified Approach to Controlling Implicit Regularization via Mirror Descent
by: Sun, Haoyuan, et al.
Published: (2023)
by: Sun, Haoyuan, et al.
Published: (2023)
Spectral Neural Graph Sparsification
by: Liguori, Angelica, et al.
Published: (2025)
by: Liguori, Angelica, et al.
Published: (2025)
Secure Aggregation Meets Sparsification in Decentralized Learning
by: Biswas, Sayan, et al.
Published: (2024)
by: Biswas, Sayan, et al.
Published: (2024)
MSQ: Memory-Efficient Bit Sparsification Quantization
by: Han, Seokho, et al.
Published: (2025)
by: Han, Seokho, et al.
Published: (2025)
Feather: An Elegant Solution to Effective DNN Sparsification
by: Georgoulakis, Athanasios Glentis, et al.
Published: (2023)
by: Georgoulakis, Athanasios Glentis, et al.
Published: (2023)
Approximation of Log-Partition Function in Policy Mirror Descent Induces Implicit Regularization for LLM Post-Training
by: Xu, Zhenghao, et al.
Published: (2026)
by: Xu, Zhenghao, et al.
Published: (2026)
Efficient Unbiased Sparsification
by: Barnes, Leighton, et al.
Published: (2024)
by: Barnes, Leighton, et al.
Published: (2024)
Joint Model and Data Sparsification via the Marginal Likelihood
by: Timans, Alexander, et al.
Published: (2026)
by: Timans, Alexander, et al.
Published: (2026)
BitSnap: Checkpoint Sparsification and Quantization in LLM Training
by: Peng, Yanxin, et al.
Published: (2025)
by: Peng, Yanxin, et al.
Published: (2025)
Mobility-Aware Asynchronous Federated Learning with Dynamic Sparsification
by: Yan, Jintao, et al.
Published: (2025)
by: Yan, Jintao, et al.
Published: (2025)
Automatic and Structure-Aware Sparsification of Hybrid Neural ODEs
by: Zou, Bob Junyi, et al.
Published: (2025)
by: Zou, Bob Junyi, et al.
Published: (2025)
Graph Sparsification via Mixture of Graphs
by: Zhang, Guibin, et al.
Published: (2024)
by: Zhang, Guibin, et al.
Published: (2024)
Implicit Bias in Noisy-SGD: With Applications to Differentially Private Training
by: Sander, Tom, et al.
Published: (2024)
by: Sander, Tom, et al.
Published: (2024)
Graph Sparsification for Enhanced Conformal Prediction in Graph Neural Networks
by: He, Yuntian, et al.
Published: (2024)
by: He, Yuntian, et al.
Published: (2024)
Similar Items
-
Mirror, Mirror of the Flow: How Does Regularization Shape Implicit Bias?
by: Jacobs, Tom, et al.
Published: (2025) -
Never Saddle for Reparameterized Steepest Descent as Mirror Flow
by: Jacobs, Tom, et al.
Published: (2026) -
Hyperbolic Aware Minimization: Implicit Bias for Sparsity
by: Jacobs, Tom, et al.
Published: (2025) -
HORST: Composing Optimizer Geometries for Sparse Transformer Training
by: Jacobs, Tom, et al.
Published: (2026) -
Masks, Signs, And Learning Rate Rewinding
by: Gadhikar, Advait, et al.
Published: (2024)