NeuroTrails: Training with Dynamic Sparse Heads as the Key to Effective Ensembling
Fuente:
arXiv
Saved in:
| Main Authors: | Grooten, Bram, Hasanov, Farid, Zhang, Chenxiang, Xiao, Qiao, Wu, Boqian, Atashgahi, Zahra, Sokar, Ghada, Liu, Shiwei, Yin, Lu, Mocanu, Elena, Pechenizkiy, Mykola, Mocanu, Decebal Constantin |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Deep Ensembling with No Overhead for either Training or Testing: The All-Round Blessings of Dynamic Sparsity
by: Liu, Shiwei, et al.
Published: (2021)
by: Liu, Shiwei, et al.
Published: (2021)
Adaptive Sparsity Level during Training for Efficient Time Series Forecasting with Transformers
by: Atashgahi, Zahra, et al.
Published: (2023)
by: Atashgahi, Zahra, et al.
Published: (2023)
Leave it to the Specialist: Repair Sparse LLMs with Sparse Fine-Tuning via Sparsity Evolution
by: Xiao, Qiao, et al.
Published: (2025)
by: Xiao, Qiao, et al.
Published: (2025)
E2ENet: Dynamic Sparse Feature Fusion for Accurate and Efficient 3D Medical Image Segmentation
by: Wu, Boqian, et al.
Published: (2023)
by: Wu, Boqian, et al.
Published: (2023)
Unveiling the Power of Sparse Neural Networks for Feature Selection
by: Atashgahi, Zahra, et al.
Published: (2024)
by: Atashgahi, Zahra, et al.
Published: (2024)
Dynamic Sparse Training versus Dense Training: The Unexpected Winner in Image Corruption Robustness
by: Wu, Boqian, et al.
Published: (2024)
by: Wu, Boqian, et al.
Published: (2024)
Addressing the Collaboration Dilemma in Low-Data Federated Learning via Transient Sparsity
by: Xiao, Qiao, et al.
Published: (2025)
by: Xiao, Qiao, et al.
Published: (2025)
Are Sparse Neural Networks Better Hard Sample Learners?
by: Xiao, Qiao, et al.
Published: (2024)
by: Xiao, Qiao, et al.
Published: (2024)
When Data Is Scarce: Scaling Sparse Language Models with Repeated Training
by: Wu, Boqian, et al.
Published: (2026)
by: Wu, Boqian, et al.
Published: (2026)
Boosting Robustness in Preference-Based Reinforcement Learning with Dynamic Sparsity
by: Muslimani, Calarina, et al.
Published: (2024)
by: Muslimani, Calarina, et al.
Published: (2024)
Memory-Efficient LLM Training with Dynamic Sparsity: From Stability to Practical Scaling
by: Xiao, Qiao, et al.
Published: (2026)
by: Xiao, Qiao, et al.
Published: (2026)
Batch Matrix-form Equations and Implementation of Multilayer Perceptrons
by: Wesselink, Wieger, et al.
Published: (2025)
by: Wesselink, Wieger, et al.
Published: (2025)
LiMTR: Time Series Motion Prediction for Diverse Road Users through Multimodal Feature Integration
by: Oerlemans, Camiel, et al.
Published: (2024)
by: Oerlemans, Camiel, et al.
Published: (2024)
Dynamic Data Pruning for Automatic Speech Recognition
by: Xiao, Qiao, et al.
Published: (2024)
by: Xiao, Qiao, et al.
Published: (2024)
Sparse-to-Sparse Training of Diffusion Models
by: Oliveira, Inês Cardoso, et al.
Published: (2025)
by: Oliveira, Inês Cardoso, et al.
Published: (2025)
Nerva: a Truly Sparse Implementation of Neural Networks
by: Wesselink, Wieger, et al.
Published: (2024)
by: Wesselink, Wieger, et al.
Published: (2024)
You Can Have Better Graph Neural Networks by Not Training Weights at All: Finding Untrained GNNs Tickets
by: Huang, Tianjin, et al.
Published: (2022)
by: Huang, Tianjin, et al.
Published: (2022)
Self-Regulated Neurogenesis for Online Data-Incremental Learning
by: Yildirim, Murat Onur, et al.
Published: (2024)
by: Yildirim, Murat Onur, et al.
Published: (2024)
Dificultăți emoționale și psihosociale ale veteranilor din teatrele de operații
by: Mocanu, Sorin, et al.
Published: (2026)
by: Mocanu, Sorin, et al.
Published: (2026)
Generalizations of four hyperbolic-type metrics and Gromov hyperbolicity
by: Mocanu, Marcelina
Published: (2024)
by: Mocanu, Marcelina
Published: (2024)
MEAL: A Benchmark for Continual Multi-Agent Reinforcement Learning
by: Tomilin, Tristan, et al.
Published: (2025)
by: Tomilin, Tristan, et al.
Published: (2025)
Mind the GAP! The Challenges of Scale in Pixel-based Deep Reinforcement Learning
by: Sokar, Ghada, et al.
Published: (2025)
by: Sokar, Ghada, et al.
Published: (2025)
Local points on twists of $X(p)$ with applications
by: Freitas, Nuno, et al.
Published: (2025)
by: Freitas, Nuno, et al.
Published: (2025)
English in Spain: Education, attitudes and native‐speakerism
by: Enric Llurda, et al.
Published: (2024)
by: Enric Llurda, et al.
Published: (2024)
Efficient Self-Supervised Neuro-Analytic Visual Servoing for Real-time Quadrotor Control
by: Mocanu, Sebastian, et al.
Published: (2025)
by: Mocanu, Sebastian, et al.
Published: (2025)
The Stochastic-Dissipative Störmer Problem-Trajectories and Radiation Patterns
by: Harko, Tiberiu, et al.
Published: (2025)
by: Harko, Tiberiu, et al.
Published: (2025)
Fertility preservation in cancer patients: Data collection, analysis, and continuous improvement of cryostorage
by: Edgar Mocanu, et al.
Published: (2025)
by: Edgar Mocanu, et al.
Published: (2025)
The Stochastic‐Dissipative Störmer Problem‐Trajectories and Radiation Patterns
by: Tiberiu Harko, et al.
Published: (2025)
by: Tiberiu Harko, et al.
Published: (2025)
Beyond Discriminant Patterns: On the Robustness of Decision Rule Ensembles
by: Du, Xin, et al.
Published: (2021)
by: Du, Xin, et al.
Published: (2021)
Self-Supervised Learning to Fly using Efficient Semantic Segmentation and Metric Depth Estimation for Low-Cost Autonomous UAVs
by: Mocanu, Sebastian, et al.
Published: (2025)
by: Mocanu, Sebastian, et al.
Published: (2025)
The Ptolemy-Alhazen problem and quadric surface mirror reflection
by: Fujimura, Masayo, et al.
Published: (2021)
by: Fujimura, Masayo, et al.
Published: (2021)
Non-trivial Integer Solutions of $x^r+y^r=Dz^p$
by: Kara, Yasemin, et al.
Published: (2024)
by: Kara, Yasemin, et al.
Published: (2024)
Barrlund's distance function and quasiconformal maps
by: Fujimura, Masayo, et al.
Published: (2019)
by: Fujimura, Masayo, et al.
Published: (2019)
Rethinking Knowledge Transfer in Learning Using Privileged Information
by: Provodin, Danil, et al.
Published: (2024)
by: Provodin, Danil, et al.
Published: (2024)
HASARD: A Benchmark for Vision-Based Safe Reinforcement Learning in Embodied Agents
by: Tomilin, Tristan, et al.
Published: (2025)
by: Tomilin, Tristan, et al.
Published: (2025)
Everyone deserves their voice to be heard: Analyzing Predictive Gender Bias in ASR Models Applied to Dutch Speech Data
by: Raes, Rik, et al.
Published: (2024)
by: Raes, Rik, et al.
Published: (2024)
Efficient Exploration in Average-Reward Constrained Reinforcement Learning: Achieving Near-Optimal Regret With Posterior Sampling
by: Provodin, Danil, et al.
Published: (2024)
by: Provodin, Danil, et al.
Published: (2024)
FairSNA: Algorithmic Fairness in Social Network Analysis
by: Saxena, Akrati, et al.
Published: (2022)
by: Saxena, Akrati, et al.
Published: (2022)
Are There Exceptions to Goodhart's Law? On the Moral Justification of Fairness-Aware Machine Learning
by: Weerts, Hilde, et al.
Published: (2022)
by: Weerts, Hilde, et al.
Published: (2022)
Revisiting the equation $x^2+y^3=z^p$
by: Freitas, Nuno, et al.
Published: (2025)
by: Freitas, Nuno, et al.
Published: (2025)
Similar Items
-
Deep Ensembling with No Overhead for either Training or Testing: The All-Round Blessings of Dynamic Sparsity
by: Liu, Shiwei, et al.
Published: (2021) -
Adaptive Sparsity Level during Training for Efficient Time Series Forecasting with Transformers
by: Atashgahi, Zahra, et al.
Published: (2023) -
Leave it to the Specialist: Repair Sparse LLMs with Sparse Fine-Tuning via Sparsity Evolution
by: Xiao, Qiao, et al.
Published: (2025) -
E2ENet: Dynamic Sparse Feature Fusion for Accurate and Efficient 3D Medical Image Segmentation
by: Wu, Boqian, et al.
Published: (2023) -
Unveiling the Power of Sparse Neural Networks for Feature Selection
by: Atashgahi, Zahra, et al.
Published: (2024)