IKUN: Initialization to Keep snn training and generalization great with sUrrogate-stable variaNce
Fuente:
arXiv
Saved in:
| Main Authors: | Chang, Da, Wang, Deliang, Yang, Xiao |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
DRiVE: Dynamic Recognition in VEhicles using snnTorch
by: Vora, Heerak, et al.
Published: (2025)
by: Vora, Heerak, et al.
Published: (2025)
Towards stable training of parallel continual learning
by: Yuepan, Li, et al.
Published: (2024)
by: Yuepan, Li, et al.
Published: (2024)
Make Interval Bound Propagation great again
by: Krukowski, Patryk, et al.
Published: (2024)
by: Krukowski, Patryk, et al.
Published: (2024)
Why Keep Your Doubts to Yourself? Trading Visual Uncertainties in Multi-Agent Bandit Systems
by: Zhang, Jusheng, et al.
Published: (2026)
by: Zhang, Jusheng, et al.
Published: (2026)
Control Tax: The Price of Keeping AI in Check
by: Terekhov, Mikhail, et al.
Published: (2025)
by: Terekhov, Mikhail, et al.
Published: (2025)
Enhanced Atrial Fibrillation Prediction in ESUS Patients with Hypergraph-based Pre-training
by: Xie, Yuzhang, et al.
Published: (2026)
by: Xie, Yuzhang, et al.
Published: (2026)
Active Learning for Continual Learning: Keeping the Past Alive in the Present
by: Park, Jaehyun, et al.
Published: (2025)
by: Park, Jaehyun, et al.
Published: (2025)
mlx-snn: Spiking Neural Networks on Apple Silicon via MLX
by: Qin, Jiahao
Published: (2026)
by: Qin, Jiahao
Published: (2026)
Causal feature selection framework for stable soft sensor modeling based on time-delayed cross mapping
by: Chen, Shi-Shun, et al.
Published: (2026)
by: Chen, Shi-Shun, et al.
Published: (2026)
A learning-driven automatic planning framework for proton PBS treatments of H&N cancers
by: Wang, Qingqing, et al.
Published: (2025)
by: Wang, Qingqing, et al.
Published: (2025)
OSF: On Pre-training and Scaling of Sleep Foundation Models
by: Shuai, Zitao, et al.
Published: (2026)
by: Shuai, Zitao, et al.
Published: (2026)
Keep Everyone Happy: Online Fair Division of Numerous Items with Few Copies
by: Verma, Arun, et al.
Published: (2024)
by: Verma, Arun, et al.
Published: (2024)
Scaling Smart: Accelerating Large Language Model Pre-training with Small Model Initialization
by: Samragh, Mohammad, et al.
Published: (2024)
by: Samragh, Mohammad, et al.
Published: (2024)
IVP-VAE: Modeling EHR Time Series with Initial Value Problem Solvers
by: Xiao, Jingge, et al.
Published: (2023)
by: Xiao, Jingge, et al.
Published: (2023)
KeepKV: Achieving Periodic Lossless KV Cache Compression for Efficient LLM Inference
by: Tian, Yuxuan, et al.
Published: (2025)
by: Tian, Yuxuan, et al.
Published: (2025)
Beyond Gaussian Initializations: Signal Preserving Weight Initialization for Odd-Sigmoid Activations
by: Lee, Hyunwoo, et al.
Published: (2025)
by: Lee, Hyunwoo, et al.
Published: (2025)
Keep Rehearsing and Refining: Lifelong Learning Vehicle Routing under Continually Drifting Tasks
by: Pei, Jiyuan, et al.
Published: (2026)
by: Pei, Jiyuan, et al.
Published: (2026)
CLoQ: Enhancing Fine-Tuning of Quantized LLMs via Calibrated LoRA Initialization
by: Deng, Yanxia, et al.
Published: (2025)
by: Deng, Yanxia, et al.
Published: (2025)
LoRA-DA: Data-Aware Initialization for Low-Rank Adaptation via Asymptotic Analysis
by: Zhang, Qingyue, et al.
Published: (2025)
by: Zhang, Qingyue, et al.
Published: (2025)
IDInit: A Universal and Stable Initialization Method for Neural Network Training
by: Pan, Yu, et al.
Published: (2025)
by: Pan, Yu, et al.
Published: (2025)
Delayed Bottlenecking: Alleviating Forgetting in Pre-trained Graph Neural Networks
by: Zhao, Zhe, et al.
Published: (2024)
by: Zhao, Zhe, et al.
Published: (2024)
Private-RAG: Answering Multiple Queries with LLMs while Keeping Your Data Private
by: Wu, Ruihan, et al.
Published: (2025)
by: Wu, Ruihan, et al.
Published: (2025)
PoTPTQ: A Two-step Power-of-Two Post-training for LLMs
by: Wang, Xinyu, et al.
Published: (2025)
by: Wang, Xinyu, et al.
Published: (2025)
Principal Components for Neural Network Initialization
by: Phan, Nhan, et al.
Published: (2025)
by: Phan, Nhan, et al.
Published: (2025)
Early-stopping for Transformer model training
by: He, Jing, et al.
Published: (2025)
by: He, Jing, et al.
Published: (2025)
DeCoP: Enhancing Self-Supervised Time Series Representation with Dependency Controlled Pre-training
by: Wu, Yuemin, et al.
Published: (2025)
by: Wu, Yuemin, et al.
Published: (2025)
Efficient Post-training Quantization with FP8 Formats
by: Shen, Haihao, et al.
Published: (2023)
by: Shen, Haihao, et al.
Published: (2023)
Enhancing robustness of data-driven SHM models: adversarial training with circle loss
by: Yang, Xiangli, et al.
Published: (2024)
by: Yang, Xiangli, et al.
Published: (2024)
Data Mixing Agent: Learning to Re-weight Domains for Continual Pre-training
by: Yang, Kailai, et al.
Published: (2025)
by: Yang, Kailai, et al.
Published: (2025)
Task-Agnostic Pre-training and Task-Guided Fine-tuning for Versatile Diffusion Planner
by: Fan, Chenyou, et al.
Published: (2024)
by: Fan, Chenyou, et al.
Published: (2024)
Balanced Edge Pruning for Graph Anomaly Detection with Noisy Labels
by: Wang, Zhu, et al.
Published: (2024)
by: Wang, Zhu, et al.
Published: (2024)
Pre-training with Synthetic Data Helps Offline Reinforcement Learning
by: Wang, Zecheng, et al.
Published: (2023)
by: Wang, Zecheng, et al.
Published: (2023)
Revisiting the Initial Steps in Adaptive Gradient Descent Optimization
by: Abuduweili, Abulikemu, et al.
Published: (2024)
by: Abuduweili, Abulikemu, et al.
Published: (2024)
Efficient Online RL Fine Tuning with Offline Pre-trained Policy Only
by: Xiao, Wei, et al.
Published: (2025)
by: Xiao, Wei, et al.
Published: (2025)
Revisiting Glorot Initialization for Long-Range Linear Recurrences
by: Bar, Noga, et al.
Published: (2025)
by: Bar, Noga, et al.
Published: (2025)
When Bias Meets Trainability: Connecting Theories of Initialization
by: Bassi, Alberto, et al.
Published: (2025)
by: Bassi, Alberto, et al.
Published: (2025)
The Initialization Determines Whether In-Context Learning Is Gradient Descent
by: Xie, Shifeng, et al.
Published: (2025)
by: Xie, Shifeng, et al.
Published: (2025)
Initializing Services in Interactive ML Systems for Diverse Users
by: Bose, Avinandan, et al.
Published: (2023)
by: Bose, Avinandan, et al.
Published: (2023)
Graph Generative Pre-trained Transformer
by: Chen, Xiaohui, et al.
Published: (2025)
by: Chen, Xiaohui, et al.
Published: (2025)
CCS: Controllable and Constrained Sampling with Diffusion Models via Initial Noise Perturbation
by: Song, Bowen, et al.
Published: (2025)
by: Song, Bowen, et al.
Published: (2025)
Similar Items
-
DRiVE: Dynamic Recognition in VEhicles using snnTorch
by: Vora, Heerak, et al.
Published: (2025) -
Towards stable training of parallel continual learning
by: Yuepan, Li, et al.
Published: (2024) -
Make Interval Bound Propagation great again
by: Krukowski, Patryk, et al.
Published: (2024) -
Why Keep Your Doubts to Yourself? Trading Visual Uncertainties in Multi-Agent Bandit Systems
by: Zhang, Jusheng, et al.
Published: (2026) -
Control Tax: The Price of Keeping AI in Check
by: Terekhov, Mikhail, et al.
Published: (2025)