Sparse Training from Random Initialization: Aligning Lottery Ticket Masks using Weight Symmetry
Fuente:
arXiv
Salvato in:
| Autori principali: | Adnan, Mohammed, Jain, Rohan, Sharma, Ekansh, Krishnan, Rahul G., Ioannou, Yani |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
SparseOpt: Addressing Normalization-induced Gradient Skew in Sparse Training
di: Adnan, Mohammed, et al.
Pubblicazione: (2026)
di: Adnan, Mohammed, et al.
Pubblicazione: (2026)
When Layers Play the Lottery, all Tickets Win at Initialization
di: Jordao, Artur, et al.
Pubblicazione: (2023)
di: Jordao, Artur, et al.
Pubblicazione: (2023)
Bayesian Lottery Ticket Hypothesis
di: Kuhn, Nicholas, et al.
Pubblicazione: (2026)
di: Kuhn, Nicholas, et al.
Pubblicazione: (2026)
On the Sparsity of the Strong Lottery Ticket Hypothesis
di: Natale, Emanuele, et al.
Pubblicazione: (2024)
di: Natale, Emanuele, et al.
Pubblicazione: (2024)
Partially Frozen Random Networks Contain Compact Strong Lottery Tickets
di: Otsuka, Hikari, et al.
Pubblicazione: (2024)
di: Otsuka, Hikari, et al.
Pubblicazione: (2024)
A Survey of Lottery Ticket Hypothesis
di: Liu, Bohan, et al.
Pubblicazione: (2024)
di: Liu, Bohan, et al.
Pubblicazione: (2024)
Dynamic Sparse Training with Structured Sparsity
di: Lasby, Mike, et al.
Pubblicazione: (2023)
di: Lasby, Mike, et al.
Pubblicazione: (2023)
Insights into the Lottery Ticket Hypothesis and Iterative Magnitude Pruning
di: Saleem, Tausifa Jan, et al.
Pubblicazione: (2024)
di: Saleem, Tausifa Jan, et al.
Pubblicazione: (2024)
KS-Lottery: Finding Certified Lottery Tickets for Multilingual Language Models
di: Yuan, Fei, et al.
Pubblicazione: (2024)
di: Yuan, Fei, et al.
Pubblicazione: (2024)
Towards Scalable Lottery Ticket Networks using Genetic Algorithms
di: Schönberger, Julian, et al.
Pubblicazione: (2025)
di: Schönberger, Julian, et al.
Pubblicazione: (2025)
Quantization vs Pruning: Insights from the Strong Lottery Ticket Hypothesis
di: Kumar, Aakash, et al.
Pubblicazione: (2025)
di: Kumar, Aakash, et al.
Pubblicazione: (2025)
Weisfeiler and Leman Go Gambling: Why Expressive Lottery Tickets Win
di: Kummer, Lorenz, et al.
Pubblicazione: (2025)
di: Kummer, Lorenz, et al.
Pubblicazione: (2025)
Bridging Lottery Ticket and Grokking: Understanding Grokking from Inner Structure of Networks
di: Minegishi, Gouki, et al.
Pubblicazione: (2023)
di: Minegishi, Gouki, et al.
Pubblicazione: (2023)
What is Left After Distillation? How Knowledge Transfer Impacts Fairness and Bias
di: Mohammadshahi, Aida, et al.
Pubblicazione: (2024)
di: Mohammadshahi, Aida, et al.
Pubblicazione: (2024)
Investigating the Lottery Ticket Hypothesis for Variational Quantum Circuits
di: Kölle, Michael, et al.
Pubblicazione: (2025)
di: Kölle, Michael, et al.
Pubblicazione: (2025)
Grokking as Structural Inference: Transformers Need Bayesian Lottery Tickets
di: Hidajat, Kai, et al.
Pubblicazione: (2026)
di: Hidajat, Kai, et al.
Pubblicazione: (2026)
The Strong Lottery Ticket Hypothesis for Multi-Head Attention Mechanisms
di: Otsuka, Hikari, et al.
Pubblicazione: (2025)
di: Otsuka, Hikari, et al.
Pubblicazione: (2025)
Winning the Lottery by Preserving Network Training Dynamics with Concrete Ticket Search
di: Arora, Tanay, et al.
Pubblicazione: (2025)
di: Arora, Tanay, et al.
Pubblicazione: (2025)
Toy Combinatorial Interpretability Models Reveal Lottery Tickets in Early Feature Space
di: Bebchuk, Alon, et al.
Pubblicazione: (2026)
di: Bebchuk, Alon, et al.
Pubblicazione: (2026)
Learning Fine-grained Parameter Sharing via Sparse Tensor Decomposition
di: Üyük, Cem, et al.
Pubblicazione: (2024)
di: Üyük, Cem, et al.
Pubblicazione: (2024)
A Neural Scaling Law from Lottery Ticket Ensembling
di: Liu, Ziming, et al.
Pubblicazione: (2023)
di: Liu, Ziming, et al.
Pubblicazione: (2023)
Early Transformers: A study on Efficient Training of Transformer Models through Early-Bird Lottery Tickets
di: Cheekati, Shravan
Pubblicazione: (2024)
di: Cheekati, Shravan
Pubblicazione: (2024)
The Non-Local Model Merging Problem: Permutation Symmetries and Variance Collapse
di: Sharma, Ekansh, et al.
Pubblicazione: (2024)
di: Sharma, Ekansh, et al.
Pubblicazione: (2024)
Uncovering a Winning Lottery Ticket with Continuously Relaxed Bernoulli Gates
di: Tsayag, Itamar, et al.
Pubblicazione: (2026)
di: Tsayag, Itamar, et al.
Pubblicazione: (2026)
On the Mechanism and Dynamics of Modular Addition: Fourier Features, Lottery Ticket, and Grokking
di: He, Jianliang, et al.
Pubblicazione: (2026)
di: He, Jianliang, et al.
Pubblicazione: (2026)
Meta-GCN: A Dynamically Weighted Loss Minimization Method for Dealing with the Data Imbalance in Graph Neural Networks
di: Mohammadizadeh, Mahdi, et al.
Pubblicazione: (2024)
di: Mohammadizadeh, Mahdi, et al.
Pubblicazione: (2024)
CAWI: Copula-Aligned Weight Initialization for Randomized Neural Networks
di: Akhtar, Mushir, et al.
Pubblicazione: (2026)
di: Akhtar, Mushir, et al.
Pubblicazione: (2026)
SuperTickets: Drawing Task-Agnostic Lottery Tickets from Supernets via Jointly Architecture Searching and Parameter Pruning
di: You, Haoran, et al.
Pubblicazione: (2022)
di: You, Haoran, et al.
Pubblicazione: (2022)
HORST: Composing Optimizer Geometries for Sparse Transformer Training
di: Jacobs, Tom, et al.
Pubblicazione: (2026)
di: Jacobs, Tom, et al.
Pubblicazione: (2026)
The Multiple Ticket Hypothesis: Random Sparse Subnetworks Suffice for RLVR
di: Adewuyi, Israel, et al.
Pubblicazione: (2026)
di: Adewuyi, Israel, et al.
Pubblicazione: (2026)
Winning Lottery Tickets in Neural Networks via a Quantum-Inspired Classical Algorithm
di: Isogai, Natsuto, et al.
Pubblicazione: (2026)
di: Isogai, Natsuto, et al.
Pubblicazione: (2026)
LOTUS: Improving Transformer Efficiency with Sparsity Pruning and Data Lottery Tickets
di: Upadhyay, Ojasw
Pubblicazione: (2024)
di: Upadhyay, Ojasw
Pubblicazione: (2024)
Sign-In to the Lottery: Reparameterizing Sparse Training From Scratch
di: Gadhikar, Advait, et al.
Pubblicazione: (2025)
di: Gadhikar, Advait, et al.
Pubblicazione: (2025)
Polynomially Over-Parameterized Convolutional Neural Networks Contain Structured Strong Winning Lottery Tickets
di: da Cunha, Arthur, et al.
Pubblicazione: (2023)
di: da Cunha, Arthur, et al.
Pubblicazione: (2023)
Random Masking Finds Winning Tickets for Parameter Efficient Fine-tuning
di: Xu, Jing, et al.
Pubblicazione: (2024)
di: Xu, Jing, et al.
Pubblicazione: (2024)
Drawing Robust Scratch Tickets: Subnetworks with Inborn Robustness Are Found within Randomly Initialized Networks
di: Fu, Yonggan, et al.
Pubblicazione: (2021)
di: Fu, Yonggan, et al.
Pubblicazione: (2021)
Lost in Translation: How Language Re-Aligns Vision for Cross-Species Pathology
di: Arora, Ekansh
Pubblicazione: (2026)
di: Arora, Ekansh
Pubblicazione: (2026)
Beyond Masked and Unmasked: Discrete Diffusion Models via Partial Masking
di: Chao, Chen-Hao, et al.
Pubblicazione: (2025)
di: Chao, Chen-Hao, et al.
Pubblicazione: (2025)
You Can Have Better Graph Neural Networks by Not Training Weights at All: Finding Untrained GNNs Tickets
di: Huang, Tianjin, et al.
Pubblicazione: (2022)
di: Huang, Tianjin, et al.
Pubblicazione: (2022)
Playing the Lottery With Concave Regularizers for Sparse Trainable Neural Networks
di: Fracastoro, Giulia, et al.
Pubblicazione: (2025)
di: Fracastoro, Giulia, et al.
Pubblicazione: (2025)
Documenti analoghi
-
SparseOpt: Addressing Normalization-induced Gradient Skew in Sparse Training
di: Adnan, Mohammed, et al.
Pubblicazione: (2026) -
When Layers Play the Lottery, all Tickets Win at Initialization
di: Jordao, Artur, et al.
Pubblicazione: (2023) -
Bayesian Lottery Ticket Hypothesis
di: Kuhn, Nicholas, et al.
Pubblicazione: (2026) -
On the Sparsity of the Strong Lottery Ticket Hypothesis
di: Natale, Emanuele, et al.
Pubblicazione: (2024) -
Partially Frozen Random Networks Contain Compact Strong Lottery Tickets
di: Otsuka, Hikari, et al.
Pubblicazione: (2024)