Stochastic Layer-wise Learning: Scalable and Efficient Alternative to Backpropagation
Fuente:
arXiv
Saved in:
| Main Authors: | Yin, Bojian, Corradi, Federico |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Never Reset Again: A Mathematical Framework for Continual Inference in Recurrent Neural Networks
by: Yin, Bojian, et al.
Published: (2024)
by: Yin, Bojian, et al.
Published: (2024)
Traces Propagation: Memory-Efficient and Scalable Forward-Only Learning in Spiking Neural Networks
by: Pes, Lorenzo, et al.
Published: (2025)
by: Pes, Lorenzo, et al.
Published: (2025)
Resource-efficient Layer-wise Federated Self-supervised Learning
by: Tun, Ye Lin, et al.
Published: (2024)
by: Tun, Ye Lin, et al.
Published: (2024)
Dynamic Spectral Backpropagation for Efficient Neural Network Training
by: Muthuraman, Mannmohan
Published: (2025)
by: Muthuraman, Mannmohan
Published: (2025)
Scalable Learning in Structured Recurrent Spiking Neural Networks without Backpropagation
by: Tang, Bo, et al.
Published: (2026)
by: Tang, Bo, et al.
Published: (2026)
Practical Boolean Backpropagation
by: Golbert, Simon
Published: (2025)
by: Golbert, Simon
Published: (2025)
Efficient Knowledge Deletion from Trained Models through Layer-wise Partial Machine Unlearning
by: Gogineni, Vinay Chakravarthi, et al.
Published: (2024)
by: Gogineni, Vinay Chakravarthi, et al.
Published: (2024)
A Layer-wise Analysis of Supervised Fine-Tuning
by: Zhao, Qinghua, et al.
Published: (2026)
by: Zhao, Qinghua, et al.
Published: (2026)
LEAP: Layer-wise Exit-Aware Pretraining for Efficient Transformer Inference
by: Kapadia, Shashank, et al.
Published: (2026)
by: Kapadia, Shashank, et al.
Published: (2026)
Learning without Global Backpropagation via Synergistic Information Distillation
by: Ye, Chenhao, et al.
Published: (2025)
by: Ye, Chenhao, et al.
Published: (2025)
Local Pairwise Distance Matching for Backpropagation-Free Reinforcement Learning
by: Tanneberg, Daniel
Published: (2025)
by: Tanneberg, Daniel
Published: (2025)
SAL: Selective Adaptive Learning for Backpropagation-Free Training with Sparsification
by: Liu, Fanping, et al.
Published: (2026)
by: Liu, Fanping, et al.
Published: (2026)
StreamBP: Memory-Efficient Exact Backpropagation for Long Sequence Training of LLMs
by: Luo, Qijun, et al.
Published: (2025)
by: Luo, Qijun, et al.
Published: (2025)
Fast, Scalable, Energy-Efficient Non-element-wise Matrix Multiplication on FPGA
by: Zhu, Xuqi, et al.
Published: (2024)
by: Zhu, Xuqi, et al.
Published: (2024)
Towards Green AI in Fine-tuning Large Language Models via Adaptive Backpropagation
by: Huang, Kai, et al.
Published: (2023)
by: Huang, Kai, et al.
Published: (2023)
LAVa: Layer-wise KV Cache Eviction with Dynamic Budget Allocation
by: Shen, Yiqun, et al.
Published: (2025)
by: Shen, Yiqun, et al.
Published: (2025)
Information-Theoretic Greedy Layer-wise Training for Traffic Sign Recognition
by: Lyu, Shuyan, et al.
Published: (2025)
by: Lyu, Shuyan, et al.
Published: (2025)
Advancing On-Device Neural Network Training with TinyPropv2: Dynamic, Sparse, and Efficient Backpropagation
by: Rüb, Marcus, et al.
Published: (2024)
by: Rüb, Marcus, et al.
Published: (2024)
Backpropagation-Free Metropolis-Adjusted Langevin Algorithm
by: Cobb, Adam D., et al.
Published: (2025)
by: Cobb, Adam D., et al.
Published: (2025)
DP-LLM: Runtime Model Adaptation with Dynamic Layer-wise Precision Assignment
by: Kwon, Sangwoo, et al.
Published: (2025)
by: Kwon, Sangwoo, et al.
Published: (2025)
LEVI: Generalizable Fine-tuning via Layer-wise Ensemble of Different Views
by: Roh, Yuji, et al.
Published: (2024)
by: Roh, Yuji, et al.
Published: (2024)
An Efficient Training Algorithm for Models with Block-wise Sparsity
by: Zhu, Ding, et al.
Published: (2025)
by: Zhu, Ding, et al.
Published: (2025)
Efficient Sparse Selective-Update RNNs for Long-Range Sequence Modeling
by: Yin, Bojian, et al.
Published: (2026)
by: Yin, Bojian, et al.
Published: (2026)
HKAN: Hierarchical Kolmogorov-Arnold Network without Backpropagation
by: Dudek, Grzegorz, et al.
Published: (2025)
by: Dudek, Grzegorz, et al.
Published: (2025)
Dense Backpropagation Improves Training for Sparse Mixture-of-Experts
by: Panda, Ashwinee, et al.
Published: (2025)
by: Panda, Ashwinee, et al.
Published: (2025)
GRADE: Replacing Policy Gradients with Backpropagation for LLM Alignment
by: Nel, Lukas Abrie
Published: (2025)
by: Nel, Lukas Abrie
Published: (2025)
Exploring the Performance of Perforated Backpropagation through Further Experiments
by: Brenner, Rorry, et al.
Published: (2025)
by: Brenner, Rorry, et al.
Published: (2025)
Beyond Backpropagation: Optimization with Multi-Tangent Forward Gradients
by: Flügel, Katharina, et al.
Published: (2024)
by: Flügel, Katharina, et al.
Published: (2024)
Exploring Layer-wise Information Effectiveness for Post-Training Quantization in Small Language Models
by: Xiao, He, et al.
Published: (2025)
by: Xiao, He, et al.
Published: (2025)
Disentangling Recall and Reasoning in Transformer Models through Layer-wise Attention and Activation Analysis
by: Fartale, Harshwardhan, et al.
Published: (2025)
by: Fartale, Harshwardhan, et al.
Published: (2025)
Fine-tuning Diffusion Policies with Backpropagation Through Diffusion Timesteps
by: Yang, Ningyuan, et al.
Published: (2025)
by: Yang, Ningyuan, et al.
Published: (2025)
A Novel Multimodal RUL Framework for Remaining Useful Life Estimation with Layer-wise Explanations
by: Razzaq, Waleed, et al.
Published: (2025)
by: Razzaq, Waleed, et al.
Published: (2025)
LaRA: Layer-wise Representation Analysis for Detecting Data Contamination in RL Post-Training
by: Gwak, Minju, et al.
Published: (2026)
by: Gwak, Minju, et al.
Published: (2026)
Region-wise stacking ensembles for estimating brain-age using MRI
by: Antonopoulos, Georgios, et al.
Published: (2025)
by: Antonopoulos, Georgios, et al.
Published: (2025)
Principled Approximation Methods for Efficient and Scalable Deep Learning
by: Savarese, Pedro
Published: (2025)
by: Savarese, Pedro
Published: (2025)
Towards Instance-wise Personalized Federated Learning via Semi-Implicit Bayesian Prompt Tuning
by: Ye, Tiandi, et al.
Published: (2025)
by: Ye, Tiandi, et al.
Published: (2025)
MISA: Memory-Efficient LLMs Optimization with Module-wise Importance Sampling
by: Liu, Yuxi, et al.
Published: (2025)
by: Liu, Yuxi, et al.
Published: (2025)
Training Long-Context LLMs Efficiently via Chunk-wise Optimization
by: Li, Wenhao, et al.
Published: (2025)
by: Li, Wenhao, et al.
Published: (2025)
HeadInfer: Memory-Efficient LLM Inference by Head-wise Offloading
by: Luo, Cheng, et al.
Published: (2025)
by: Luo, Cheng, et al.
Published: (2025)
Asynchronous Stochastic Gradient Descent with Decoupled Backpropagation and Layer-Wise Updates
by: Fokam, Cabrel Teguemne, et al.
Published: (2024)
by: Fokam, Cabrel Teguemne, et al.
Published: (2024)
Similar Items
-
Never Reset Again: A Mathematical Framework for Continual Inference in Recurrent Neural Networks
by: Yin, Bojian, et al.
Published: (2024) -
Traces Propagation: Memory-Efficient and Scalable Forward-Only Learning in Spiking Neural Networks
by: Pes, Lorenzo, et al.
Published: (2025) -
Resource-efficient Layer-wise Federated Self-supervised Learning
by: Tun, Ye Lin, et al.
Published: (2024) -
Dynamic Spectral Backpropagation for Efficient Neural Network Training
by: Muthuraman, Mannmohan
Published: (2025) -
Scalable Learning in Structured Recurrent Spiking Neural Networks without Backpropagation
by: Tang, Bo, et al.
Published: (2026)