Rethinking the Potential of Layer Freezing for Efficient DNN Training
Fuente:
arXiv
Saved in:
| Main Authors: | Yang, Chence, Zhang, Ci, Lu, Lei, Tan, Qitao, Li, Sheng, Li, Ao, Tang, Xulong, Huang, Shaoyi, Wang, Jinzhen, Li, Guoming, Li, Jundong, Zhai, Xiaoming, Lu, Jin, Yuan, Geng |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
End-to-End On-Device Quantization-Aware Training for LLMs at Inference Cost
by: Tan, Qitao, et al.
Published: (2025)
by: Tan, Qitao, et al.
Published: (2025)
SmartFRZ: An Efficient Training Framework using Attention-Based Layer Freezing
by: Li, Sheng, et al.
Published: (2024)
by: Li, Sheng, et al.
Published: (2024)
Roots Beneath the Cut: Uncovering the Risk of Concept Revival in Pruning-Based Unlearning for Diffusion Models
by: Zhang, Ci, et al.
Published: (2026)
by: Zhang, Ci, et al.
Published: (2026)
Towards Fast LLM Fine-tuning through Zeroth-Order Optimization with Projected Gradient-Aligned Perturbations
by: Mi, Zhendong, et al.
Published: (2025)
by: Mi, Zhendong, et al.
Published: (2025)
Perturbation-efficient Zeroth-order Optimization for Hardware-friendly On-device Training
by: Tan, Qitao, et al.
Published: (2025)
by: Tan, Qitao, et al.
Published: (2025)
KerZOO: Kernel Function Informed Zeroth-Order Optimization for Accurate and Accelerated LLM Fine-Tuning
by: Mi, Zhendong, et al.
Published: (2025)
by: Mi, Zhendong, et al.
Published: (2025)
Palette: A Modular, Controllable, and Efficient Framework for On-demand Authorized Safety Alignment Relaxation in LLMs
by: Tan, Qitao, et al.
Published: (2026)
by: Tan, Qitao, et al.
Published: (2026)
Layer-wise Weight Selection for Power-Efficient Neural Network Acceleration
by: Fang, Jiaxun, et al.
Published: (2025)
by: Fang, Jiaxun, et al.
Published: (2025)
Improving GPU Multi-Tenancy Through Dynamic Multi-Instance GPU Reconfiguration
by: Wang, Tianyu, et al.
Published: (2024)
by: Wang, Tianyu, et al.
Published: (2024)
A Flexible Programmable Pipeline Parallelism Framework for Efficient DNN Training
by: Jiang, Lijuan, et al.
Published: (2025)
by: Jiang, Lijuan, et al.
Published: (2025)
NestQuant: Post-Training Integer-Nesting Quantization for On-Device DNN
by: Xie, Jianhang, et al.
Published: (2025)
by: Xie, Jianhang, et al.
Published: (2025)
Q-realign: Piggybacking Realignment on Quantization for Safe and Efficient LLM Deployment
by: Tan, Qitao, et al.
Published: (2026)
by: Tan, Qitao, et al.
Published: (2026)
ST-FiT: Inductive Spatial-Temporal Forecasting with Limited Training Data
by: Lei, Zhenyu, et al.
Published: (2024)
by: Lei, Zhenyu, et al.
Published: (2024)
Hardware-Aware DNN Compression for Homogeneous Edge Devices
by: Zhang, Kunlong, et al.
Published: (2025)
by: Zhang, Kunlong, et al.
Published: (2025)
Hardware-Aware DNN Compression for Homogeneous Edge Devices
by: Zhang, Kunlong, et al.
Published: (2025)
by: Zhang, Kunlong, et al.
Published: (2025)
Temporal Aware Pruning for Efficient Diffusion-based Video Generation
by: Li, Sheng, et al.
Published: (2026)
by: Li, Sheng, et al.
Published: (2026)
Heterogeneity-Aware Memory Efficient Federated Learning via Progressive Layer Freezing
by: Yebo, Wu, et al.
Published: (2024)
by: Yebo, Wu, et al.
Published: (2024)
Advancing time series completion via RFAMoE and MDFF
by: Zhang, Ci, et al.
Published: (2025)
by: Zhang, Ci, et al.
Published: (2025)
Two-pion exchange contributions to the relativistic chiral nuclear force at N$^3$LO
by: Lu, Jun-Xu, et al.
Published: (2025)
by: Lu, Jun-Xu, et al.
Published: (2025)
Hypergraph-based Multi-View Action Recognition using Event Cameras
by: Gao, Yue, et al.
Published: (2024)
by: Gao, Yue, et al.
Published: (2024)
Fusing Neural and Physical: Augment Protein Conformation Sampling with Tractable Simulations
by: Lu, Jiarui, et al.
Published: (2024)
by: Lu, Jiarui, et al.
Published: (2024)
Design of Ligand-Binding Proteins with Atomic Flow Matching
by: Liu, Junqi, et al.
Published: (2024)
by: Liu, Junqi, et al.
Published: (2024)
Backdoor Attack Against Vision Transformers via Attention Gradient-Based Image Erosion
by: Guo, Ji, et al.
Published: (2024)
by: Guo, Ji, et al.
Published: (2024)
Layer-wise dynamic rank for compressing large language models
by: Mi, Zhendong, et al.
Published: (2025)
by: Mi, Zhendong, et al.
Published: (2025)
Randomness of Low-Layer Parameters Determines Confusing Samples in Terms of Interaction Representations of a DNN
by: Zhang, Junpeng, et al.
Published: (2025)
by: Zhang, Junpeng, et al.
Published: (2025)
Integrative Pan-Cancer Analysis of RNMT: a Potential Prognostic and Immunological Biomarker
by: Huang, Shuqiang, et al.
Published: (2022)
by: Huang, Shuqiang, et al.
Published: (2022)
PIMCOMP: An End-to-End DNN Compiler for Processing-In-Memory Accelerators
by: Sun, Xiaotian, et al.
Published: (2024)
by: Sun, Xiaotian, et al.
Published: (2024)
AI Gender Bias, Disparities, and Fairness: Does Training Data Matter?
by: Latif, Ehsan, et al.
Published: (2023)
by: Latif, Ehsan, et al.
Published: (2023)
FracTrain: Fractionally Squeezing Bit Savings Both Temporally and Spatially for Efficient DNN Training
by: Fu, Yonggan, et al.
Published: (2020)
by: Fu, Yonggan, et al.
Published: (2020)
Organ Transplantation: Current Status, Challenges, and Future Prospects
by: Xinqiang Li, et al.
Published: (2026)
by: Xinqiang Li, et al.
Published: (2026)
EdgeOL: Efficient in-situ Online Learning on Edge Devices
by: Li, Sheng, et al.
Published: (2024)
by: Li, Sheng, et al.
Published: (2024)
Efficient Dataset Distillation for Pre-Trained Self-Supervised Models via Statistical Flow Matching
by: Xia, Qianxin, et al.
Published: (2026)
by: Xia, Qianxin, et al.
Published: (2026)
JARVIS: An Evidence-Grounded Retrieval System for Interpretable Deceptive Reviews Adjudication
by: Lu, Nan, et al.
Published: (2026)
by: Lu, Nan, et al.
Published: (2026)
Instance-Aware Graph Prompt Learning
by: Li, Jiazheng, et al.
Published: (2024)
by: Li, Jiazheng, et al.
Published: (2024)
ACE: Exploring Activation Cosine Similarity and Variance for Accurate and Calibration-Efficient LLM Pruning
by: Mi, Zhendong, et al.
Published: (2025)
by: Mi, Zhendong, et al.
Published: (2025)
Bit Transition Reduction by Data Transmission Ordering in NoC-based DNN Accelerator
by: Chen, Yizhi, et al.
Published: (2025)
by: Chen, Yizhi, et al.
Published: (2025)
Rethinking Cross-Layer Information Routing in Diffusion Transformers
by: Xu, Chao, et al.
Published: (2026)
by: Xu, Chao, et al.
Published: (2026)
Non-Clifford Fusion: T-Gate Optimization for Quantum Simulation
by: Li, Yingheng, et al.
Published: (2025)
by: Li, Yingheng, et al.
Published: (2025)
Atomic Trajectory Modeling with State Space Models for Biomolecular Dynamics
by: Shi, Liang, et al.
Published: (2026)
by: Shi, Liang, et al.
Published: (2026)
Charge dependent nucleon-nucleon potentials in covariant chiral effective field theory
by: Xiao, Yang, et al.
Published: (2025)
by: Xiao, Yang, et al.
Published: (2025)
Similar Items
-
End-to-End On-Device Quantization-Aware Training for LLMs at Inference Cost
by: Tan, Qitao, et al.
Published: (2025) -
SmartFRZ: An Efficient Training Framework using Attention-Based Layer Freezing
by: Li, Sheng, et al.
Published: (2024) -
Roots Beneath the Cut: Uncovering the Risk of Concept Revival in Pruning-Based Unlearning for Diffusion Models
by: Zhang, Ci, et al.
Published: (2026) -
Towards Fast LLM Fine-tuning through Zeroth-Order Optimization with Projected Gradient-Aligned Perturbations
by: Mi, Zhendong, et al.
Published: (2025) -
Perturbation-efficient Zeroth-order Optimization for Hardware-friendly On-device Training
by: Tan, Qitao, et al.
Published: (2025)