Reassessing Layer Pruning in LLMs: New Insights and Methods
Fuente:
arXiv
Guardado en:
| Autores principales: | Lu, Yao, Cheng, Hao, Fang, Yujie, Wang, Zeyu, Wei, Jiaheng, Xu, Dongwei, Xuan, Qi, Yang, Xiaoniu, Zhu, Zhaowei |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
A Generic Layer Pruning Method for Signal Modulation Recognition Deep Learning Models
por: Lu, Yao, et al.
Publicado: (2024)
por: Lu, Yao, et al.
Publicado: (2024)
SelectMix: Enhancing Label Noise Robustness through Targeted Sample Mixing
por: Liu, Qiuhao, et al.
Publicado: (2025)
por: Liu, Qiuhao, et al.
Publicado: (2025)
Better Reasoning with Less Data: Enhancing VLMs Through Unified Modality Scoring
por: Xu, Mingjie, et al.
Publicado: (2025)
por: Xu, Mingjie, et al.
Publicado: (2025)
RedTest: Towards Measuring Redundancy in Deep Neural Networks Effectively
por: Lu, Yao, et al.
Publicado: (2024)
por: Lu, Yao, et al.
Publicado: (2024)
The Structural Scalpel: Automated Contiguous Layer Pruning for Large Language Models
por: Lu, Yao, et al.
Publicado: (2025)
por: Lu, Yao, et al.
Publicado: (2025)
MDM: Advancing Multi-Domain Distribution Matching for Automatic Modulation Recognition Dataset Synthesis
por: Xu, Dongwei, et al.
Publicado: (2024)
por: Xu, Dongwei, et al.
Publicado: (2024)
FCOS: A Two-Stage Recoverable Model Pruning Framework for Automatic Modulation Recognition
por: Lu, Yao, et al.
Publicado: (2025)
por: Lu, Yao, et al.
Publicado: (2025)
Exploring The Neural Burden In Pruned Models: An Insight Inspired By Neuroscience
por: Wang, Zeyu, et al.
Publicado: (2024)
por: Wang, Zeyu, et al.
Publicado: (2024)
Motion-Aware Adaptive Pixel Pruning for Efficient Local Motion Deblurring
por: Shang, Wei, et al.
Publicado: (2025)
por: Shang, Wei, et al.
Publicado: (2025)
SGLP: A Similarity Guided Fast Layer Partition Pruning for Compressing Large Deep Models
por: Li, Yuqi, et al.
Publicado: (2024)
por: Li, Yuqi, et al.
Publicado: (2024)
ALTER: All-in-One Layer Pruning and Temporal Expert Routing for Efficient Diffusion Generation
por: Yang, Xiaomeng, et al.
Publicado: (2025)
por: Yang, Xiaomeng, et al.
Publicado: (2025)
Improving the Convergence Rate of Ray Search Optimization for Query-Efficient Hard-Label Attacks
por: Xu, Xinjie, et al.
Publicado: (2025)
por: Xu, Xinjie, et al.
Publicado: (2025)
Pluggable Pruning with Contiguous Layer Distillation for Diffusion Transformers
por: Ma, Jian, et al.
Publicado: (2025)
por: Ma, Jian, et al.
Publicado: (2025)
FGP: Feature-Gradient-Prune for Efficient Convolutional Layer Pruning
por: Lv, Qingsong, et al.
Publicado: (2024)
por: Lv, Qingsong, et al.
Publicado: (2024)
Vision Function Layer in Multimodal LLMs
por: Shi, Cheng, et al.
Publicado: (2025)
por: Shi, Cheng, et al.
Publicado: (2025)
Sliding-Window Merging for Compacting Patch-Redundant Layers in LLMs
por: Ding, Xuan, et al.
Publicado: (2025)
por: Ding, Xuan, et al.
Publicado: (2025)
Mixing Signals: Data Augmentation Approach for Deep Learning Based Modulation Recognition
por: Xu, Xinjie, et al.
Publicado: (2022)
por: Xu, Xinjie, et al.
Publicado: (2022)
HIVTP: A Training-Free Method to Improve VLMs Efficiency via Hierarchical Visual Token Pruning Using Middle-Layer-Based Importance Score
por: Xu, Jingqi, et al.
Publicado: (2025)
por: Xu, Jingqi, et al.
Publicado: (2025)
ReDiPrune: Relevance-Diversity Pre-Projection Token Pruning for Efficient Multimodal LLMs
por: Yu, An, et al.
Publicado: (2026)
por: Yu, An, et al.
Publicado: (2026)
Thin-Plate Spline-based Interpolation for Animation Line Inbetweening
por: Zhu, Tianyi, et al.
Publicado: (2024)
por: Zhu, Tianyi, et al.
Publicado: (2024)
Enhancing Layer Attention Efficiency through Pruning Redundant Retrievals
por: Li, Hanze, et al.
Publicado: (2025)
por: Li, Hanze, et al.
Publicado: (2025)
LAPTOP-Diff: Layer Pruning and Normalized Distillation for Compressing Diffusion Models
por: Zhang, Dingkun, et al.
Publicado: (2024)
por: Zhang, Dingkun, et al.
Publicado: (2024)
TP-Spikformer: Token Pruned Spiking Transformer
por: Wei, Wenjie, et al.
Publicado: (2026)
por: Wei, Wenjie, et al.
Publicado: (2026)
An Effective Image Copy-Move Forgery Detection Using Entropy Information
por: Jiang, Li, et al.
Publicado: (2023)
por: Jiang, Li, et al.
Publicado: (2023)
IVC-Prune: Revealing the Implicit Visual Coordinates in LVLMs for Vision Token Pruning
por: Sun, Zhichao, et al.
Publicado: (2026)
por: Sun, Zhichao, et al.
Publicado: (2026)
Decay Pruning Method: Smooth Pruning With a Self-Rectifying Procedure
por: Yang, Minghao, et al.
Publicado: (2024)
por: Yang, Minghao, et al.
Publicado: (2024)
Topology-Aware Layer Pruning for Large Vision-Language Models
por: Zheng, Pengcheng, et al.
Publicado: (2026)
por: Zheng, Pengcheng, et al.
Publicado: (2026)
MMA: Multimodal Memory Agent
por: Lu, Yihao, et al.
Publicado: (2026)
por: Lu, Yihao, et al.
Publicado: (2026)
ReStNet: A Reusable & Stitchable Network for Dynamic Adaptation on IoT Devices
por: Wang, Maoyu, et al.
Publicado: (2025)
por: Wang, Maoyu, et al.
Publicado: (2025)
CorrAdaptor: Adaptive Local Context Learning for Correspondence Pruning
por: Zhu, Wei, et al.
Publicado: (2024)
por: Zhu, Wei, et al.
Publicado: (2024)
CLASP: Class-Adaptive Layer Fusion and Dual-Stage Pruning for Multimodal Large Language Models
por: Dang, Yunkai, et al.
Publicado: (2026)
por: Dang, Yunkai, et al.
Publicado: (2026)
Subgraph Networks Based Contrastive Learning
por: Wang, Jinhuan, et al.
Publicado: (2023)
por: Wang, Jinhuan, et al.
Publicado: (2023)
Towards Fine-Grained Text-to-3D Quality Assessment: A Benchmark and A Two-Stage Rank-Learning Metric
por: Cui, Bingyang, et al.
Publicado: (2025)
por: Cui, Bingyang, et al.
Publicado: (2025)
Benchmarking and Learning Multi-Dimensional Quality Evaluator for Text-to-3D Generation
por: Zhang, Yujie, et al.
Publicado: (2024)
por: Zhang, Yujie, et al.
Publicado: (2024)
Mamba-3D as Masked Autoencoders for Accurate and Data-Efficient Analysis of Medical Ultrasound Videos
por: Zhou, Jiaheng, et al.
Publicado: (2025)
por: Zhou, Jiaheng, et al.
Publicado: (2025)
MCLRL: A Multi-Domain Contrastive Learning with Reinforcement Learning Framework for Few-Shot Modulation Recognition
por: Xu, Dongwei, et al.
Publicado: (2025)
por: Xu, Dongwei, et al.
Publicado: (2025)
Short-LVLM: Compressing and Accelerating Large Vision-Language Models by Pruning Redundant Layers
por: Ma, Ji, et al.
Publicado: (2025)
por: Ma, Ji, et al.
Publicado: (2025)
Selecting and Pruning: A Differentiable Causal Sequentialized State-Space Model for Two-View Correspondence Learning
por: Fang, Xiang, et al.
Publicado: (2025)
por: Fang, Xiang, et al.
Publicado: (2025)
SCOPE: Saliency-Coverage Oriented Token Pruning for Efficient Multimodel LLMs
por: Deng, Jinhong, et al.
Publicado: (2025)
por: Deng, Jinhong, et al.
Publicado: (2025)
First-order State Space Model for Lightweight Image Super-resolution
por: Zhu, Yujie, et al.
Publicado: (2025)
por: Zhu, Yujie, et al.
Publicado: (2025)
Ejemplares similares
-
A Generic Layer Pruning Method for Signal Modulation Recognition Deep Learning Models
por: Lu, Yao, et al.
Publicado: (2024) -
SelectMix: Enhancing Label Noise Robustness through Targeted Sample Mixing
por: Liu, Qiuhao, et al.
Publicado: (2025) -
Better Reasoning with Less Data: Enhancing VLMs Through Unified Modality Scoring
por: Xu, Mingjie, et al.
Publicado: (2025) -
RedTest: Towards Measuring Redundancy in Deep Neural Networks Effectively
por: Lu, Yao, et al.
Publicado: (2024) -
The Structural Scalpel: Automated Contiguous Layer Pruning for Large Language Models
por: Lu, Yao, et al.
Publicado: (2025)