Deep Ensembling with No Overhead for either Training or Testing: The All-Round Blessings of Dynamic Sparsity
Fuente:
arXiv
Saved in:
| Main Authors: | Liu, Shiwei, Chen, Tianlong, Atashgahi, Zahra, Chen, Xiaohan, Sokar, Ghada, Mocanu, Elena, Pechenizkiy, Mykola, Wang, Zhangyang, Mocanu, Decebal Constantin |
|---|---|
| Format: | Preprint |
| Published: |
2021
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Adaptive Sparsity Level during Training for Efficient Time Series Forecasting with Transformers
by: Atashgahi, Zahra, et al.
Published: (2023)
by: Atashgahi, Zahra, et al.
Published: (2023)
NeuroTrails: Training with Dynamic Sparse Heads as the Key to Effective Ensembling
by: Grooten, Bram, et al.
Published: (2025)
by: Grooten, Bram, et al.
Published: (2025)
Unveiling the Power of Sparse Neural Networks for Feature Selection
by: Atashgahi, Zahra, et al.
Published: (2024)
by: Atashgahi, Zahra, et al.
Published: (2024)
Addressing the Collaboration Dilemma in Low-Data Federated Learning via Transient Sparsity
by: Xiao, Qiao, et al.
Published: (2025)
by: Xiao, Qiao, et al.
Published: (2025)
Memory-Efficient LLM Training with Dynamic Sparsity: From Stability to Practical Scaling
by: Xiao, Qiao, et al.
Published: (2026)
by: Xiao, Qiao, et al.
Published: (2026)
Leave it to the Specialist: Repair Sparse LLMs with Sparse Fine-Tuning via Sparsity Evolution
by: Xiao, Qiao, et al.
Published: (2025)
by: Xiao, Qiao, et al.
Published: (2025)
You Can Have Better Graph Neural Networks by Not Training Weights at All: Finding Untrained GNNs Tickets
by: Huang, Tianjin, et al.
Published: (2022)
by: Huang, Tianjin, et al.
Published: (2022)
Dynamic Sparse Training versus Dense Training: The Unexpected Winner in Image Corruption Robustness
by: Wu, Boqian, et al.
Published: (2024)
by: Wu, Boqian, et al.
Published: (2024)
E2ENet: Dynamic Sparse Feature Fusion for Accurate and Efficient 3D Medical Image Segmentation
by: Wu, Boqian, et al.
Published: (2023)
by: Wu, Boqian, et al.
Published: (2023)
Boosting Robustness in Preference-Based Reinforcement Learning with Dynamic Sparsity
by: Muslimani, Calarina, et al.
Published: (2024)
by: Muslimani, Calarina, et al.
Published: (2024)
When Data Is Scarce: Scaling Sparse Language Models with Repeated Training
by: Wu, Boqian, et al.
Published: (2026)
by: Wu, Boqian, et al.
Published: (2026)
Sparse-to-Sparse Training of Diffusion Models
by: Oliveira, Inês Cardoso, et al.
Published: (2025)
by: Oliveira, Inês Cardoso, et al.
Published: (2025)
Are Sparse Neural Networks Better Hard Sample Learners?
by: Xiao, Qiao, et al.
Published: (2024)
by: Xiao, Qiao, et al.
Published: (2024)
Dynamic Data Pruning for Automatic Speech Recognition
by: Xiao, Qiao, et al.
Published: (2024)
by: Xiao, Qiao, et al.
Published: (2024)
(PASS) Visual Prompt Locates Good Structure Sparsity through a Recurrent HyperNetwork
by: Huang, Tianjin, et al.
Published: (2024)
by: Huang, Tianjin, et al.
Published: (2024)
Self-Regulated Neurogenesis for Online Data-Incremental Learning
by: Yildirim, Murat Onur, et al.
Published: (2024)
by: Yildirim, Murat Onur, et al.
Published: (2024)
Mind the GAP! The Challenges of Scale in Pixel-based Deep Reinforcement Learning
by: Sokar, Ghada, et al.
Published: (2025)
by: Sokar, Ghada, et al.
Published: (2025)
Batch Matrix-form Equations and Implementation of Multilayer Perceptrons
by: Wesselink, Wieger, et al.
Published: (2025)
by: Wesselink, Wieger, et al.
Published: (2025)
LiMTR: Time Series Motion Prediction for Diverse Road Users through Multimodal Feature Integration
by: Oerlemans, Camiel, et al.
Published: (2024)
by: Oerlemans, Camiel, et al.
Published: (2024)
Dificultăți emoționale și psihosociale ale veteranilor din teatrele de operații
by: Mocanu, Sorin, et al.
Published: (2026)
by: Mocanu, Sorin, et al.
Published: (2026)
Generalizations of four hyperbolic-type metrics and Gromov hyperbolicity
by: Mocanu, Marcelina
Published: (2024)
by: Mocanu, Marcelina
Published: (2024)
Enhancing Adversarial Training via Reweighting Optimization Trajectory
by: Huang, Tianjin, et al.
Published: (2023)
by: Huang, Tianjin, et al.
Published: (2023)
Visual Prompting Upgrades Neural Network Sparsification: A Data-Model Perspective
by: Jin, Can, et al.
Published: (2023)
by: Jin, Can, et al.
Published: (2023)
Outlier Weighed Layerwise Sparsity (OWL): A Missing Secret Sauce for Pruning LLMs to High Sparsity
by: Yin, Lu, et al.
Published: (2023)
by: Yin, Lu, et al.
Published: (2023)
Local points on twists of $X(p)$ with applications
by: Freitas, Nuno, et al.
Published: (2025)
by: Freitas, Nuno, et al.
Published: (2025)
English in Spain: Education, attitudes and native‐speakerism
by: Enric Llurda, et al.
Published: (2024)
by: Enric Llurda, et al.
Published: (2024)
The Stochastic-Dissipative Störmer Problem-Trajectories and Radiation Patterns
by: Harko, Tiberiu, et al.
Published: (2025)
by: Harko, Tiberiu, et al.
Published: (2025)
Fertility preservation in cancer patients: Data collection, analysis, and continuous improvement of cryostorage
by: Edgar Mocanu, et al.
Published: (2025)
by: Edgar Mocanu, et al.
Published: (2025)
The Stochastic‐Dissipative Störmer Problem‐Trajectories and Radiation Patterns
by: Tiberiu Harko, et al.
Published: (2025)
by: Tiberiu Harko, et al.
Published: (2025)
Beyond Discriminant Patterns: On the Robustness of Decision Rule Ensembles
by: Du, Xin, et al.
Published: (2021)
by: Du, Xin, et al.
Published: (2021)
Self-Supervised Learning to Fly using Efficient Semantic Segmentation and Metric Depth Estimation for Low-Cost Autonomous UAVs
by: Mocanu, Sebastian, et al.
Published: (2025)
by: Mocanu, Sebastian, et al.
Published: (2025)
The Ptolemy-Alhazen problem and quadric surface mirror reflection
by: Fujimura, Masayo, et al.
Published: (2021)
by: Fujimura, Masayo, et al.
Published: (2021)
Non-trivial Integer Solutions of $x^r+y^r=Dz^p$
by: Kara, Yasemin, et al.
Published: (2024)
by: Kara, Yasemin, et al.
Published: (2024)
Barrlund's distance function and quasiconformal maps
by: Fujimura, Masayo, et al.
Published: (2019)
by: Fujimura, Masayo, et al.
Published: (2019)
HASARD: A Benchmark for Vision-Based Safe Reinforcement Learning in Embodied Agents
by: Tomilin, Tristan, et al.
Published: (2025)
by: Tomilin, Tristan, et al.
Published: (2025)
Everyone deserves their voice to be heard: Analyzing Predictive Gender Bias in ASR Models Applied to Dutch Speech Data
by: Raes, Rik, et al.
Published: (2024)
by: Raes, Rik, et al.
Published: (2024)
Efficient Exploration in Average-Reward Constrained Reinforcement Learning: Achieving Near-Optimal Regret With Posterior Sampling
by: Provodin, Danil, et al.
Published: (2024)
by: Provodin, Danil, et al.
Published: (2024)
FairSNA: Algorithmic Fairness in Social Network Analysis
by: Saxena, Akrati, et al.
Published: (2022)
by: Saxena, Akrati, et al.
Published: (2022)
Are There Exceptions to Goodhart's Law? On the Moral Justification of Fairness-Aware Machine Learning
by: Weerts, Hilde, et al.
Published: (2022)
by: Weerts, Hilde, et al.
Published: (2022)
Revisiting the equation $x^2+y^3=z^p$
by: Freitas, Nuno, et al.
Published: (2025)
by: Freitas, Nuno, et al.
Published: (2025)
Similar Items
-
Adaptive Sparsity Level during Training for Efficient Time Series Forecasting with Transformers
by: Atashgahi, Zahra, et al.
Published: (2023) -
NeuroTrails: Training with Dynamic Sparse Heads as the Key to Effective Ensembling
by: Grooten, Bram, et al.
Published: (2025) -
Unveiling the Power of Sparse Neural Networks for Feature Selection
by: Atashgahi, Zahra, et al.
Published: (2024) -
Addressing the Collaboration Dilemma in Low-Data Federated Learning via Transient Sparsity
by: Xiao, Qiao, et al.
Published: (2025) -
Memory-Efficient LLM Training with Dynamic Sparsity: From Stability to Practical Scaling
by: Xiao, Qiao, et al.
Published: (2026)