Saved in:
| Main Authors: | Xiao, Qiao, Wu, Boqian, Okanovic, Patrik, Sternal, Tomasz, van Keulen, Maurice, Mocanu, Elena, Pechenizkiy, Mykola, Mocanu, Decebal Constantin, Hoefler, Torsten |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2606.00888 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
When Data Is Scarce: Scaling Sparse Language Models with Repeated Training
by: Wu, Boqian, et al.
Published: (2026)
by: Wu, Boqian, et al.
Published: (2026)
Dynamic Sparse Training versus Dense Training: The Unexpected Winner in Image Corruption Robustness
by: Wu, Boqian, et al.
Published: (2024)
by: Wu, Boqian, et al.
Published: (2024)
E2ENet: Dynamic Sparse Feature Fusion for Accurate and Efficient 3D Medical Image Segmentation
by: Wu, Boqian, et al.
Published: (2023)
by: Wu, Boqian, et al.
Published: (2023)
Addressing the Collaboration Dilemma in Low-Data Federated Learning via Transient Sparsity
by: Xiao, Qiao, et al.
Published: (2025)
by: Xiao, Qiao, et al.
Published: (2025)
Adaptive Sparsity Level during Training for Efficient Time Series Forecasting with Transformers
by: Atashgahi, Zahra, et al.
Published: (2023)
by: Atashgahi, Zahra, et al.
Published: (2023)
Leave it to the Specialist: Repair Sparse LLMs with Sparse Fine-Tuning via Sparsity Evolution
by: Xiao, Qiao, et al.
Published: (2025)
by: Xiao, Qiao, et al.
Published: (2025)
Deep Ensembling with No Overhead for either Training or Testing: The All-Round Blessings of Dynamic Sparsity
by: Liu, Shiwei, et al.
Published: (2021)
by: Liu, Shiwei, et al.
Published: (2021)
Are Sparse Neural Networks Better Hard Sample Learners?
by: Xiao, Qiao, et al.
Published: (2024)
by: Xiao, Qiao, et al.
Published: (2024)
NeuroTrails: Training with Dynamic Sparse Heads as the Key to Effective Ensembling
by: Grooten, Bram, et al.
Published: (2025)
by: Grooten, Bram, et al.
Published: (2025)
Boosting Robustness in Preference-Based Reinforcement Learning with Dynamic Sparsity
by: Muslimani, Calarina, et al.
Published: (2024)
by: Muslimani, Calarina, et al.
Published: (2024)
Unveiling the Power of Sparse Neural Networks for Feature Selection
by: Atashgahi, Zahra, et al.
Published: (2024)
by: Atashgahi, Zahra, et al.
Published: (2024)
Dynamic Data Pruning for Automatic Speech Recognition
by: Xiao, Qiao, et al.
Published: (2024)
by: Xiao, Qiao, et al.
Published: (2024)
Sparse-to-Sparse Training of Diffusion Models
by: Oliveira, Inês Cardoso, et al.
Published: (2025)
by: Oliveira, Inês Cardoso, et al.
Published: (2025)
EntryPrune: Neural Network Feature Selection using First Impressions
by: Zimmer, Felix, et al.
Published: (2024)
by: Zimmer, Felix, et al.
Published: (2024)
Batch Matrix-form Equations and Implementation of Multilayer Perceptrons
by: Wesselink, Wieger, et al.
Published: (2025)
by: Wesselink, Wieger, et al.
Published: (2025)
You Can Have Better Graph Neural Networks by Not Training Weights at All: Finding Untrained GNNs Tickets
by: Huang, Tianjin, et al.
Published: (2022)
by: Huang, Tianjin, et al.
Published: (2022)
Confounder Detection via Treatment Intent: A New Observational Study Design
by: Plecko, Drago, et al.
Published: (2026)
by: Plecko, Drago, et al.
Published: (2026)
Epidemiology of Large Language Models: A Benchmark for Observational Distribution Knowledge
by: Plecko, Drago, et al.
Published: (2025)
by: Plecko, Drago, et al.
Published: (2025)
Self-Regulated Neurogenesis for Online Data-Incremental Learning
by: Yildirim, Murat Onur, et al.
Published: (2024)
by: Yildirim, Murat Onur, et al.
Published: (2024)
Process Reward Agents for Steering Knowledge-Intensive Reasoning
by: Sohn, Jiwoong, et al.
Published: (2026)
by: Sohn, Jiwoong, et al.
Published: (2026)
LiMTR: Time Series Motion Prediction for Diverse Road Users through Multimodal Feature Integration
by: Oerlemans, Camiel, et al.
Published: (2024)
by: Oerlemans, Camiel, et al.
Published: (2024)
Large Language Model Selection with Limited Annotations
by: Durmazkeser, Yavuz, et al.
Published: (2026)
by: Durmazkeser, Yavuz, et al.
Published: (2026)
Active Model Selection for Large Language Models
by: Durmazkeser, Yavuz, et al.
Published: (2025)
by: Durmazkeser, Yavuz, et al.
Published: (2025)
Dificultăți emoționale și psihosociale ale veteranilor din teatrele de operații
by: Mocanu, Sorin, et al.
Published: (2026)
by: Mocanu, Sorin, et al.
Published: (2026)
High Performance Unstructured SpMM Computation Using Tensor Cores
by: Okanovic, Patrik, et al.
Published: (2024)
by: Okanovic, Patrik, et al.
Published: (2024)
All models are wrong, some are useful: Model Selection with Limited Labels
by: Okanovic, Patrik, et al.
Published: (2024)
by: Okanovic, Patrik, et al.
Published: (2024)
Generalizations of four hyperbolic-type metrics and Gromov hyperbolicity
by: Mocanu, Marcelina
Published: (2024)
by: Mocanu, Marcelina
Published: (2024)
Chimera: Efficiently Training Large-Scale Neural Networks with Bidirectional Pipelines
by: Li, Shigang, et al.
Published: (2021)
by: Li, Shigang, et al.
Published: (2021)
Local points on twists of $X(p)$ with applications
by: Freitas, Nuno, et al.
Published: (2025)
by: Freitas, Nuno, et al.
Published: (2025)
English in Spain: Education, attitudes and native‐speakerism
by: Enric Llurda, et al.
Published: (2024)
by: Enric Llurda, et al.
Published: (2024)
Self-Supervised Learning to Fly using Efficient Semantic Segmentation and Metric Depth Estimation for Low-Cost Autonomous UAVs
by: Mocanu, Sebastian, et al.
Published: (2025)
by: Mocanu, Sebastian, et al.
Published: (2025)
Efficient Exploration in Average-Reward Constrained Reinforcement Learning: Achieving Near-Optimal Regret With Posterior Sampling
by: Provodin, Danil, et al.
Published: (2024)
by: Provodin, Danil, et al.
Published: (2024)
The Stochastic-Dissipative Störmer Problem-Trajectories and Radiation Patterns
by: Harko, Tiberiu, et al.
Published: (2025)
by: Harko, Tiberiu, et al.
Published: (2025)
Fertility preservation in cancer patients: Data collection, analysis, and continuous improvement of cryostorage
by: Edgar Mocanu, et al.
Published: (2025)
by: Edgar Mocanu, et al.
Published: (2025)
The Stochastic‐Dissipative Störmer Problem‐Trajectories and Radiation Patterns
by: Tiberiu Harko, et al.
Published: (2025)
by: Tiberiu Harko, et al.
Published: (2025)
The Ptolemy-Alhazen problem and quadric surface mirror reflection
by: Fujimura, Masayo, et al.
Published: (2021)
by: Fujimura, Masayo, et al.
Published: (2021)
Non-trivial Integer Solutions of $x^r+y^r=Dz^p$
by: Kara, Yasemin, et al.
Published: (2024)
by: Kara, Yasemin, et al.
Published: (2024)
Barrlund's distance function and quasiconformal maps
by: Fujimura, Masayo, et al.
Published: (2019)
by: Fujimura, Masayo, et al.
Published: (2019)
HASARD: A Benchmark for Vision-Based Safe Reinforcement Learning in Embodied Agents
by: Tomilin, Tristan, et al.
Published: (2025)
by: Tomilin, Tristan, et al.
Published: (2025)
Everyone deserves their voice to be heard: Analyzing Predictive Gender Bias in ASR Models Applied to Dutch Speech Data
by: Raes, Rik, et al.
Published: (2024)
by: Raes, Rik, et al.
Published: (2024)
Similar Items
-
When Data Is Scarce: Scaling Sparse Language Models with Repeated Training
by: Wu, Boqian, et al.
Published: (2026) -
Dynamic Sparse Training versus Dense Training: The Unexpected Winner in Image Corruption Robustness
by: Wu, Boqian, et al.
Published: (2024) -
E2ENet: Dynamic Sparse Feature Fusion for Accurate and Efficient 3D Medical Image Segmentation
by: Wu, Boqian, et al.
Published: (2023) -
Addressing the Collaboration Dilemma in Low-Data Federated Learning via Transient Sparsity
by: Xiao, Qiao, et al.
Published: (2025) -
Adaptive Sparsity Level during Training for Efficient Time Series Forecasting with Transformers
by: Atashgahi, Zahra, et al.
Published: (2023)