Finding Stable Subnetworks at Initialization with Dataset Distillation
Fuente:
arXiv
Saved in:
| Main Authors: | McDermott, Luke, Parhi, Rahul |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
LoLA: Low-Rank Linear Attention With Sparse Caching
by: McDermott, Luke, et al.
Published: (2025)
by: McDermott, Luke, et al.
Published: (2025)
Embedding Compression for Efficient Re-Identification
by: McDermott, Luke
Published: (2024)
by: McDermott, Luke
Published: (2024)
Linear Mode Connectivity in Sparse Neural Networks
by: McDermott, Luke, et al.
Published: (2023)
by: McDermott, Luke, et al.
Published: (2023)
Function-Space Optimality of Neural Architectures with Multivariate Nonlinearities
by: Parhi, Rahul, et al.
Published: (2023)
by: Parhi, Rahul, et al.
Published: (2023)
On the Loss Landscape Geometry of Regularized Deep Matrix Factorization: Uniqueness and Sharpness
by: Kamber, Anil, et al.
Published: (2026)
by: Kamber, Anil, et al.
Published: (2026)
Text Conditioned Symbolic Drumbeat Generation using Latent Diffusion Models
by: Jajoria, Pushkar, et al.
Published: (2024)
by: Jajoria, Pushkar, et al.
Published: (2024)
Sharpness of Minima in Deep Matrix Factorization
by: Kamber, Anil, et al.
Published: (2025)
by: Kamber, Anil, et al.
Published: (2025)
Higher-Order Singular-Value Derivatives of Rectangular Real Matrices
by: Luo, Róisín, et al.
Published: (2025)
by: Luo, Róisín, et al.
Published: (2025)
Stable Minima of ReLU Neural Networks Suffer from the Curse of Dimensionality: The Neural Shattering Phenomenon
by: Liang, Tongtong, et al.
Published: (2025)
by: Liang, Tongtong, et al.
Published: (2025)
Nonasymptotic Convergence Rates for Plug-and-Play Methods With MMSE Denoisers
by: Pritchard, Henry, et al.
Published: (2025)
by: Pritchard, Henry, et al.
Published: (2025)
ACES: Automatic Cohort Extraction System for Event-Stream Datasets
by: Xu, Justin, et al.
Published: (2024)
by: Xu, Justin, et al.
Published: (2024)
A Gap Between the Gaussian RKHS and Neural Networks: An Infinite-Center Asymptotic Analysis
by: Kumar, Akash, et al.
Published: (2025)
by: Kumar, Akash, et al.
Published: (2025)
Neural Architecture Codesign for Fast Physics Applications
by: Weitz, Jason, et al.
Published: (2025)
by: Weitz, Jason, et al.
Published: (2025)
Bias In, Bias Out? Finding Unbiased Subnetworks in Vanilla Models
by: Matos, Ivan Luiz De Moura, et al.
Published: (2026)
by: Matos, Ivan Luiz De Moura, et al.
Published: (2026)
Drawing Robust Scratch Tickets: Subnetworks with Inborn Robustness Are Found within Randomly Initialized Networks
by: Fu, Yonggan, et al.
Published: (2021)
by: Fu, Yonggan, et al.
Published: (2021)
Towards Sharp Minimax Risk Bounds for Operator Learning
by: Adcock, Ben, et al.
Published: (2025)
by: Adcock, Ben, et al.
Published: (2025)
Stochastic Subnetwork Annealing: A Regularization Technique for Fine Tuning Pruned Subnetworks
by: Whitaker, Tim, et al.
Published: (2024)
by: Whitaker, Tim, et al.
Published: (2024)
SSFL: Discovering Sparse Unified Subnetworks at Initialization for Efficient Federated Learning
by: Ohib, Riyasat, et al.
Published: (2024)
by: Ohib, Riyasat, et al.
Published: (2024)
Generalization Below the Edge of Stability: The Role of Data Geometry
by: Liang, Tongtong, et al.
Published: (2025)
by: Liang, Tongtong, et al.
Published: (2025)
Variation Spaces for Multi-Output Neural Networks: Insights on Multi-Task Learning and Network Compression
by: Shenouda, Joseph, et al.
Published: (2023)
by: Shenouda, Joseph, et al.
Published: (2023)
Efficient Generative Prediction for EHR Foundation Models: The SCOPE and REACH Estimators
by: Solo, Luke, et al.
Published: (2026)
by: Solo, Luke, et al.
Published: (2026)
Focused Discriminative Training For Streaming CTC-Trained Automatic Speech Recognition Models
by: Haider, Adnan, et al.
Published: (2024)
by: Haider, Adnan, et al.
Published: (2024)
Neural Subnetwork Ensembles
by: Whitaker, Tim
Published: (2023)
by: Whitaker, Tim
Published: (2023)
Optimization-Induced Dynamics of Lipschitz Continuity in Neural Networks
by: Luo, Róisín, et al.
Published: (2025)
by: Luo, Róisín, et al.
Published: (2025)
Instilling Inductive Biases with Subnetworks
by: Zhang, Enyan, et al.
Published: (2023)
by: Zhang, Enyan, et al.
Published: (2023)
Does Sparse Connectivity Improve Generalization? Convolutional Networks Below the Edge of Stability
by: Liang, Tongtong, et al.
Published: (2026)
by: Liang, Tongtong, et al.
Published: (2026)
A Closer Look at AUROC and AUPRC under Class Imbalance
by: McDermott, Matthew B. A., et al.
Published: (2024)
by: McDermott, Matthew B. A., et al.
Published: (2024)
Model Parallelism With Subnetwork Data Parallelism
by: Singh, Vaibhav, et al.
Published: (2025)
by: Singh, Vaibhav, et al.
Published: (2025)
Sequential Bayesian Neural Subnetwork Ensembles
by: Jantre, Sanket, et al.
Published: (2022)
by: Jantre, Sanket, et al.
Published: (2022)
Random ReLU Neural Networks as Non-Gaussian Processes
by: Parhi, Rahul, et al.
Published: (2024)
by: Parhi, Rahul, et al.
Published: (2024)
Towards Stable and Storage-efficient Dataset Distillation: Matching Convexified Trajectory
by: Zhong, Wenliang, et al.
Published: (2024)
by: Zhong, Wenliang, et al.
Published: (2024)
GUIDE: Guided Initialization and Distillation of Embeddings
by: Trinh, Khoa, et al.
Published: (2025)
by: Trinh, Khoa, et al.
Published: (2025)
Adaptive Dual-Teacher Distillation with Subnetwork Rectification for Bridging Semantic Gaps in Black-Box Domain Adaptation
by: Zhang, Zhe, et al.
Published: (2026)
by: Zhang, Zhe, et al.
Published: (2026)
Efficient Stagewise Pretraining via Progressive Subnetworks
by: Panigrahi, Abhishek, et al.
Published: (2024)
by: Panigrahi, Abhishek, et al.
Published: (2024)
Weighted variation spaces and approximation by shallow ReLU networks
by: DeVore, Ronald, et al.
Published: (2023)
by: DeVore, Ronald, et al.
Published: (2023)
Reinforcement Learning Finetunes Small Subnetworks in Large Language Models
by: Mukherjee, Sagnik, et al.
Published: (2025)
by: Mukherjee, Sagnik, et al.
Published: (2025)
REDS: Resource-Efficient Deep Subnetworks for Dynamic Resource Constraints
by: Corti, Francesco, et al.
Published: (2023)
by: Corti, Francesco, et al.
Published: (2023)
Internet Instruction: Spreading the Web.
by: McDermott, Irene E.
Published: (2000)
by: McDermott, Irene E.
Published: (2000)
MEDS-Tab: Automated tabularization and baseline methods for MEDS datasets
by: Oufattole, Nassim, et al.
Published: (2024)
by: Oufattole, Nassim, et al.
Published: (2024)
Memory Constrained Dynamic Subnetwork Update for Transfer Learning
by: Quélennec, Aël, et al.
Published: (2025)
by: Quélennec, Aël, et al.
Published: (2025)
Similar Items
-
LoLA: Low-Rank Linear Attention With Sparse Caching
by: McDermott, Luke, et al.
Published: (2025) -
Embedding Compression for Efficient Re-Identification
by: McDermott, Luke
Published: (2024) -
Linear Mode Connectivity in Sparse Neural Networks
by: McDermott, Luke, et al.
Published: (2023) -
Function-Space Optimality of Neural Architectures with Multivariate Nonlinearities
by: Parhi, Rahul, et al.
Published: (2023) -
On the Loss Landscape Geometry of Regularized Deep Matrix Factorization: Uniqueness and Sharpness
by: Kamber, Anil, et al.
Published: (2026)