Adaptive Sparsity Level during Training for Efficient Time Series Forecasting with Transformers
Fuente:
arXiv
Saved in:
| Main Authors: | Atashgahi, Zahra, Pechenizkiy, Mykola, Veldhuis, Raymond, Mocanu, Decebal Constantin |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Unveiling the Power of Sparse Neural Networks for Feature Selection
by: Atashgahi, Zahra, et al.
Published: (2024)
by: Atashgahi, Zahra, et al.
Published: (2024)
Deep Ensembling with No Overhead for either Training or Testing: The All-Round Blessings of Dynamic Sparsity
by: Liu, Shiwei, et al.
Published: (2021)
by: Liu, Shiwei, et al.
Published: (2021)
Memory-Efficient LLM Training with Dynamic Sparsity: From Stability to Practical Scaling
by: Xiao, Qiao, et al.
Published: (2026)
by: Xiao, Qiao, et al.
Published: (2026)
Addressing the Collaboration Dilemma in Low-Data Federated Learning via Transient Sparsity
by: Xiao, Qiao, et al.
Published: (2025)
by: Xiao, Qiao, et al.
Published: (2025)
NeuroTrails: Training with Dynamic Sparse Heads as the Key to Effective Ensembling
by: Grooten, Bram, et al.
Published: (2025)
by: Grooten, Bram, et al.
Published: (2025)
Boosting Robustness in Preference-Based Reinforcement Learning with Dynamic Sparsity
by: Muslimani, Calarina, et al.
Published: (2024)
by: Muslimani, Calarina, et al.
Published: (2024)
When Data Is Scarce: Scaling Sparse Language Models with Repeated Training
by: Wu, Boqian, et al.
Published: (2026)
by: Wu, Boqian, et al.
Published: (2026)
Are Sparse Neural Networks Better Hard Sample Learners?
by: Xiao, Qiao, et al.
Published: (2024)
by: Xiao, Qiao, et al.
Published: (2024)
Sparse-to-Sparse Training of Diffusion Models
by: Oliveira, Inês Cardoso, et al.
Published: (2025)
by: Oliveira, Inês Cardoso, et al.
Published: (2025)
You Can Have Better Graph Neural Networks by Not Training Weights at All: Finding Untrained GNNs Tickets
by: Huang, Tianjin, et al.
Published: (2022)
by: Huang, Tianjin, et al.
Published: (2022)
Leave it to the Specialist: Repair Sparse LLMs with Sparse Fine-Tuning via Sparsity Evolution
by: Xiao, Qiao, et al.
Published: (2025)
by: Xiao, Qiao, et al.
Published: (2025)
LiMTR: Time Series Motion Prediction for Diverse Road Users through Multimodal Feature Integration
by: Oerlemans, Camiel, et al.
Published: (2024)
by: Oerlemans, Camiel, et al.
Published: (2024)
Dynamic Sparse Training versus Dense Training: The Unexpected Winner in Image Corruption Robustness
by: Wu, Boqian, et al.
Published: (2024)
by: Wu, Boqian, et al.
Published: (2024)
Counterfactual Explanations for Time Series Should be Human-Centered and Temporally Coherent in Interventions
by: Chukwu, Emmanuel C., et al.
Published: (2025)
by: Chukwu, Emmanuel C., et al.
Published: (2025)
Efficient Exploration in Average-Reward Constrained Reinforcement Learning: Achieving Near-Optimal Regret With Posterior Sampling
by: Provodin, Danil, et al.
Published: (2024)
by: Provodin, Danil, et al.
Published: (2024)
E2ENet: Dynamic Sparse Feature Fusion for Accurate and Efficient 3D Medical Image Segmentation
by: Wu, Boqian, et al.
Published: (2023)
by: Wu, Boqian, et al.
Published: (2023)
Self-Regulated Neurogenesis for Online Data-Incremental Learning
by: Yildirim, Murat Onur, et al.
Published: (2024)
by: Yildirim, Murat Onur, et al.
Published: (2024)
Are There Exceptions to Goodhart's Law? On the Moral Justification of Fairness-Aware Machine Learning
by: Weerts, Hilde, et al.
Published: (2022)
by: Weerts, Hilde, et al.
Published: (2022)
Unified Training of Universal Time Series Forecasting Transformers
by: Woo, Gerald, et al.
Published: (2024)
by: Woo, Gerald, et al.
Published: (2024)
Batch Matrix-form Equations and Implementation of Multilayer Perceptrons
by: Wesselink, Wieger, et al.
Published: (2025)
by: Wesselink, Wieger, et al.
Published: (2025)
Conformalized Exceptional Model Mining: Telling Where Your Model Performs (Not) Well
by: Du, Xin, et al.
Published: (2025)
by: Du, Xin, et al.
Published: (2025)
Outlier Weighed Layerwise Sparsity (OWL): A Missing Secret Sauce for Pruning LLMs to High Sparsity
by: Yin, Lu, et al.
Published: (2023)
by: Yin, Lu, et al.
Published: (2023)
Pathformer: Multi-scale Transformers with Adaptive Pathways for Time Series Forecasting
by: Chen, Peng, et al.
Published: (2024)
by: Chen, Peng, et al.
Published: (2024)
Contrastive Time Series Forecasting with Anomalies
by: Ekstrand, Joel, et al.
Published: (2025)
by: Ekstrand, Joel, et al.
Published: (2025)
Large Scale Hierarchical Industrial Demand Time-Series Forecasting incorporating Sparsity
by: Kamarthi, Harshavardhan, et al.
Published: (2024)
by: Kamarthi, Harshavardhan, et al.
Published: (2024)
HASARD: A Benchmark for Vision-Based Safe Reinforcement Learning in Embodied Agents
by: Tomilin, Tristan, et al.
Published: (2025)
by: Tomilin, Tristan, et al.
Published: (2025)
Robust Active Learning (RoAL): Countering Dynamic Adversaries in Active Learning with Elastic Weight Consolidation
by: Fajri, Ricky Maulana, et al.
Published: (2024)
by: Fajri, Ricky Maulana, et al.
Published: (2024)
Ada-MSHyper: Adaptive Multi-Scale Hypergraph Transformer for Time Series Forecasting
by: Shang, Zongjiang, et al.
Published: (2024)
by: Shang, Zongjiang, et al.
Published: (2024)
AWGformer: Adaptive Wavelet-Guided Transformer for Multi-Resolution Time Series Forecasting
by: Li, Wei
Published: (2026)
by: Li, Wei
Published: (2026)
FEATHer: Fourier-Efficient Adaptive Temporal Hierarchy Forecaster for Time-Series Forecasting
by: Lee, Jaehoon, et al.
Published: (2026)
by: Lee, Jaehoon, et al.
Published: (2026)
Forecasting with Guidance: Representation-Level Supervision for Time Series Forecasting
by: Wang, Jiacheng, et al.
Published: (2026)
by: Wang, Jiacheng, et al.
Published: (2026)
TwinFormer: A Dual-Level Transformer for Long-Sequence Time-Series Forecasting
by: Kumavat, Mahima, et al.
Published: (2025)
by: Kumavat, Mahima, et al.
Published: (2025)
Ister: Linear Transformer for Efficient Multivariate Time Series Forecasting
by: Cao, Fanpu, et al.
Published: (2024)
by: Cao, Fanpu, et al.
Published: (2024)
EcoSpa: Efficient Transformer Training with Coupled Sparsity
by: Xiao, Jinqi, et al.
Published: (2025)
by: Xiao, Jinqi, et al.
Published: (2025)
One-Shot Federated Learning with Bayesian Pseudocoresets
by: d'Hondt, Tim, et al.
Published: (2024)
by: d'Hondt, Tim, et al.
Published: (2024)
Investigating Gender Fairness in Machine Learning-driven Personalized Care for Chronic Pain
by: Gajane, Pratik, et al.
Published: (2024)
by: Gajane, Pratik, et al.
Published: (2024)
iTransformer: Inverted Transformers Are Effective for Time Series Forecasting
by: Liu, Yong, et al.
Published: (2023)
by: Liu, Yong, et al.
Published: (2023)
MultiResFormer: Transformer with Adaptive Multi-Resolution Modeling for General Time Series Forecasting
by: Du, Linfeng, et al.
Published: (2023)
by: Du, Linfeng, et al.
Published: (2023)
AdaMixT: Adaptive Weighted Mixture of Multi-Scale Expert Transformers for Time Series Forecasting
by: Zhang, Huanyao, et al.
Published: (2025)
by: Zhang, Huanyao, et al.
Published: (2025)
FaCTR: Factorized Channel-Temporal Representation Transformers for Efficient Time Series Forecasting
by: Vijay, Yash, et al.
Published: (2025)
by: Vijay, Yash, et al.
Published: (2025)
Similar Items
-
Unveiling the Power of Sparse Neural Networks for Feature Selection
by: Atashgahi, Zahra, et al.
Published: (2024) -
Deep Ensembling with No Overhead for either Training or Testing: The All-Round Blessings of Dynamic Sparsity
by: Liu, Shiwei, et al.
Published: (2021) -
Memory-Efficient LLM Training with Dynamic Sparsity: From Stability to Practical Scaling
by: Xiao, Qiao, et al.
Published: (2026) -
Addressing the Collaboration Dilemma in Low-Data Federated Learning via Transient Sparsity
by: Xiao, Qiao, et al.
Published: (2025) -
NeuroTrails: Training with Dynamic Sparse Heads as the Key to Effective Ensembling
by: Grooten, Bram, et al.
Published: (2025)