SGLP: A Similarity Guided Fast Layer Partition Pruning for Compressing Large Deep Models
Fuente:
arXiv
Guardado en:
| Autores principales: | Li, Yuqi, Lu, Yao, Dong, Junhao, Dong, Zeyu, Yang, Chuanguang, Yin, Xin, Chen, Yihao, Gou, Jianping, Tian, Yingli, Huang, Tingwen |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
SepPrune: Structured Pruning for Efficient Deep Speech Separation
por: Li, Yuqi, et al.
Publicado: (2025)
por: Li, Yuqi, et al.
Publicado: (2025)
AMMKD: Adaptive Multimodal Multi-teacher Distillation for Lightweight Vision-Language Models
por: Li, Yuqi, et al.
Publicado: (2025)
por: Li, Yuqi, et al.
Publicado: (2025)
Federated Knowledge Distillation for Multi-Model Architectures Lithography Hotspot Detection
por: Li, Yuqi, et al.
Publicado: (2025)
por: Li, Yuqi, et al.
Publicado: (2025)
MMT-ARD: Multimodal Multi-Teacher Adversarial Distillation for Robust Vision-Language Models
por: Li, Yuqi, et al.
Publicado: (2025)
por: Li, Yuqi, et al.
Publicado: (2025)
SRKD: Towards Efficient 3D Point Cloud Segmentation via Structure- and Relation-aware Knowledge Distillation
por: Li, Yuqi, et al.
Publicado: (2025)
por: Li, Yuqi, et al.
Publicado: (2025)
Frequency-Aligned Knowledge Distillation for Lightweight Spatiotemporal Forecasting
por: Li, Yuqi, et al.
Publicado: (2025)
por: Li, Yuqi, et al.
Publicado: (2025)
Distilling Time Series Foundation Models for Efficient Forecasting
por: Li, Yuqi, et al.
Publicado: (2026)
por: Li, Yuqi, et al.
Publicado: (2026)
Prototype-Driven Multi-Feature Generation for Visible-Infrared Person Re-identification
por: Li, Jiarui, et al.
Publicado: (2024)
por: Li, Jiarui, et al.
Publicado: (2024)
GaitProtector: Impersonation-Driven Gait De-Identification via Training-Free Diffusion Latent Optimization
por: Duan, Huiran, et al.
Publicado: (2026)
por: Duan, Huiran, et al.
Publicado: (2026)
The Structural Scalpel: Automated Contiguous Layer Pruning for Large Language Models
por: Lu, Yao, et al.
Publicado: (2025)
por: Lu, Yao, et al.
Publicado: (2025)
A Generic Layer Pruning Method for Signal Modulation Recognition Deep Learning Models
por: Lu, Yao, et al.
Publicado: (2024)
por: Lu, Yao, et al.
Publicado: (2024)
MPQ-DMv2: Flexible Residual Mixed Precision Quantization for Low-Bit Diffusion Models with Temporal Distillation
por: Feng, Weilun, et al.
Publicado: (2025)
por: Feng, Weilun, et al.
Publicado: (2025)
DDTime: Dataset Distillation with Spectral Alignment and Information Bottleneck for Time-Series Forecasting
por: Li, Yuqi, et al.
Publicado: (2025)
por: Li, Yuqi, et al.
Publicado: (2025)
GaitKD: A Universal Decoupled Distillation Framework for Efficient Gait Recognition
por: Li, Yuqi, et al.
Publicado: (2026)
por: Li, Yuqi, et al.
Publicado: (2026)
Spiking Layer-Adaptive Magnitude-based Pruning
por: Wang, Junqiao, et al.
Publicado: (2026)
por: Wang, Junqiao, et al.
Publicado: (2026)
DeepAAT: Deep Automated Aerial Triangulation for Fast UAV-based Mapping
por: Chen, Zequan, et al.
Publicado: (2024)
por: Chen, Zequan, et al.
Publicado: (2024)
Reassessing Layer Pruning in LLMs: New Insights and Methods
por: Lu, Yao, et al.
Publicado: (2024)
por: Lu, Yao, et al.
Publicado: (2024)
SIMPLER: Efficient Foundation Model Adaptation via Similarity-Guided Layer Pruning for Earth Observation
por: Barreiro, Víctor, et al.
Publicado: (2026)
por: Barreiro, Víctor, et al.
Publicado: (2026)
Effective Layer Pruning Through Similarity Metric Perspective
por: Pons, Ian, et al.
Publicado: (2024)
por: Pons, Ian, et al.
Publicado: (2024)
LRCP: Low-Rank Compressibility Guided Visual Token Pruning for Efficient LVLMs
por: Lu, Hongyu, et al.
Publicado: (2026)
por: Lu, Hongyu, et al.
Publicado: (2026)
GPTailor: Large Language Model Pruning Through Layer Cutting and Stitching
por: Su, Guinan, et al.
Publicado: (2025)
por: Su, Guinan, et al.
Publicado: (2025)
Model Compression using Progressive Channel Pruning
por: Guo, Jinyang, et al.
Publicado: (2025)
por: Guo, Jinyang, et al.
Publicado: (2025)
Partition Map-Based Fast Block Partitioning for VVC Inter Coding
por: Feng, Xinmin, et al.
Publicado: (2025)
por: Feng, Xinmin, et al.
Publicado: (2025)
LAPTOP-Diff: Layer Pruning and Normalized Distillation for Compressing Diffusion Models
por: Zhang, Dingkun, et al.
Publicado: (2024)
por: Zhang, Dingkun, et al.
Publicado: (2024)
Adaptive Layer Selection for Layer-Wise Token Pruning in LLM Inference
por: Taniguchi, Rei, et al.
Publicado: (2026)
por: Taniguchi, Rei, et al.
Publicado: (2026)
When Fewer Layers Break More Chains: Layer Pruning Harms Test-Time Scaling in LLMs
por: Wang, Keyu, et al.
Publicado: (2025)
por: Wang, Keyu, et al.
Publicado: (2025)
Short-LVLM: Compressing and Accelerating Large Vision-Language Models by Pruning Redundant Layers
por: Ma, Ji, et al.
Publicado: (2025)
por: Ma, Ji, et al.
Publicado: (2025)
Improving Reasoning Capabilities in Small Models through Mixture-of-Layers Distillation with Stepwise Attention on Key Information
por: Chen, Yao, et al.
Publicado: (2026)
por: Chen, Yao, et al.
Publicado: (2026)
Enhancing Multimodal Large Language Models Complex Reason via Similarity Computation
por: Zhang, Xiaofeng, et al.
Publicado: (2024)
por: Zhang, Xiaofeng, et al.
Publicado: (2024)
GRIDS: Grouped Multiple-Degradation Restoration with Image Degradation Similarity
por: Cao, Shuo, et al.
Publicado: (2024)
por: Cao, Shuo, et al.
Publicado: (2024)
LayerMerge: Neural Network Depth Compression through Layer Pruning and Merging
por: Kim, Jinuk, et al.
Publicado: (2024)
por: Kim, Jinuk, et al.
Publicado: (2024)
Safe Pruning LoRA: Robust Distance-Guided Pruning for Safety Alignment in Adaptation of LLMs
por: Ao, Shuang, et al.
Publicado: (2025)
por: Ao, Shuang, et al.
Publicado: (2025)
Learning Switchable Priors for Neural Image Compression
por: Zhang, Haotian, et al.
Publicado: (2025)
por: Zhang, Haotian, et al.
Publicado: (2025)
IWP: Token Pruning as Implicit Weight Pruning in Large Vision Language Models
por: Lee, Dong-Jae, et al.
Publicado: (2026)
por: Lee, Dong-Jae, et al.
Publicado: (2026)
FDC: Fast KV Dimensionality Compression for Efficient LLM Inference
por: Zhang, Zeyu, et al.
Publicado: (2024)
por: Zhang, Zeyu, et al.
Publicado: (2024)
RAP: KV-Cache Compression via RoPE-Aligned Pruning
por: Xin, Jihao, et al.
Publicado: (2026)
por: Xin, Jihao, et al.
Publicado: (2026)
AquaGS: Fast Underwater Scene Reconstruction with SfM-Free Gaussian Splatting
por: Shi, Junhao, et al.
Publicado: (2025)
por: Shi, Junhao, et al.
Publicado: (2025)
Domain-Specific Pruning of Large Mixture-of-Experts Models with Few-shot Demonstrations
por: Dong, Zican, et al.
Publicado: (2025)
por: Dong, Zican, et al.
Publicado: (2025)
A Simple Linear Patch Revives Layer-Pruned Large Language Models
por: Chen, Xinrui, et al.
Publicado: (2025)
por: Chen, Xinrui, et al.
Publicado: (2025)
Learned Image Compression with Hierarchical Progressive Context Modeling
por: Li, Yuqi, et al.
Publicado: (2025)
por: Li, Yuqi, et al.
Publicado: (2025)
Ejemplares similares
-
SepPrune: Structured Pruning for Efficient Deep Speech Separation
por: Li, Yuqi, et al.
Publicado: (2025) -
AMMKD: Adaptive Multimodal Multi-teacher Distillation for Lightweight Vision-Language Models
por: Li, Yuqi, et al.
Publicado: (2025) -
Federated Knowledge Distillation for Multi-Model Architectures Lithography Hotspot Detection
por: Li, Yuqi, et al.
Publicado: (2025) -
MMT-ARD: Multimodal Multi-Teacher Adversarial Distillation for Robust Vision-Language Models
por: Li, Yuqi, et al.
Publicado: (2025) -
SRKD: Towards Efficient 3D Point Cloud Segmentation via Structure- and Relation-aware Knowledge Distillation
por: Li, Yuqi, et al.
Publicado: (2025)