CFSP: An Efficient Structured Pruning Framework for LLMs with Coarse-to-Fine Activation Information
Fuente:
arXiv
Saved in:
| Main Authors: | Wang, Yuxin, Ma, Minghua, Wang, Zekun, Chen, Jingchang, Fan, Huiming, Shan, Liping, Yang, Qing, Xu, Dongliang, Liu, Ming, Qin, Bing |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
SmartTrim: Adaptive Tokens and Attention Pruning for Efficient Vision-Language Models
by: Wang, Zekun, et al.
Published: (2023)
by: Wang, Zekun, et al.
Published: (2023)
EffiVLM-BENCH: A Comprehensive Benchmark for Evaluating Training-Free Acceleration in Large Vision-Language Models
by: Wang, Zekun, et al.
Published: (2025)
by: Wang, Zekun, et al.
Published: (2025)
Divide-and-Conquer Meets Consensus: Unleashing the Power of Functions in Code Generation
by: Chen, Jingchang, et al.
Published: (2024)
by: Chen, Jingchang, et al.
Published: (2024)
Self-Critique Guided Iterative Reasoning for Multi-hop Question Answering
by: Chu, Zheng, et al.
Published: (2025)
by: Chu, Zheng, et al.
Published: (2025)
MoGU: A Framework for Enhancing Safety of Open-Sourced LLMs While Preserving Their Usability
by: Du, Yanrui, et al.
Published: (2024)
by: Du, Yanrui, et al.
Published: (2024)
Beyond Manually Designed Pruning Policies with Second-Level Performance Prediction: A Pruning Framework for LLMs
by: Ma, Zuxin, et al.
Published: (2025)
by: Ma, Zuxin, et al.
Published: (2025)
Efficient Coarse-to-Fine Diffusion Models with Time Step Sequence Redistribution
by: Tai, Yu-Shan, et al.
Published: (2026)
by: Tai, Yu-Shan, et al.
Published: (2026)
TimeBench: A Comprehensive Evaluation of Temporal Reasoning Abilities in Large Language Models
by: Chu, Zheng, et al.
Published: (2023)
by: Chu, Zheng, et al.
Published: (2023)
PrunePEFT: Iterative Hybrid Pruning for Parameter-Efficient Fine-tuning of LLMs
by: Yu, Tongzhou, et al.
Published: (2025)
by: Yu, Tongzhou, et al.
Published: (2025)
ECoFLaP: Efficient Coarse-to-Fine Layer-Wise Pruning for Vision-Language Models
by: Sung, Yi-Lin, et al.
Published: (2023)
by: Sung, Yi-Lin, et al.
Published: (2023)
SAPT: A Shared Attention Framework for Parameter-Efficient Continual Learning of Large Language Models
by: Zhao, Weixiang, et al.
Published: (2024)
by: Zhao, Weixiang, et al.
Published: (2024)
Coarse-to-Fine Learning of Dynamic Causal Structures
by: Yang, Dezhi, et al.
Published: (2026)
by: Yang, Dezhi, et al.
Published: (2026)
GradPruner: Gradient-Guided Layer Pruning Enabling Efficient Fine-Tuning and Inference for LLMs
by: Huang, Wei, et al.
Published: (2026)
by: Huang, Wei, et al.
Published: (2026)
Sink-Token-Aware Pruning for Fine-Grained Video Understanding in Efficient Video LLMs
by: Kim, Kibum, et al.
Published: (2026)
by: Kim, Kibum, et al.
Published: (2026)
Progressive Binarization with Semi-Structured Pruning for LLMs
by: Yan, Xianglong, et al.
Published: (2025)
by: Yan, Xianglong, et al.
Published: (2025)
MambaScope: Coarse-to-Fine Scoping for Efficient Vision Mamba
by: Liu, Shanhui, et al.
Published: (2025)
by: Liu, Shanhui, et al.
Published: (2025)
Learning Coarse-to-Fine Pruning of Graph Convolutional Networks for Skeleton-based Recognition
by: Sahbi, Hichem
Published: (2024)
by: Sahbi, Hichem
Published: (2024)
OMG-Agent: Toward Robust Missing Modality Generation with Decoupled Coarse-to-Fine Agentic Workflows
by: Dai, Ruiting, et al.
Published: (2026)
by: Dai, Ruiting, et al.
Published: (2026)
A time-to-event three-outcome design for randomized phase II cancer trials
by: Shan, Minghua
Published: (2025)
by: Shan, Minghua
Published: (2025)
Maximum Redundancy Pruning: A Principle-Driven Layerwise Sparsity Allocation for LLMs
by: Gao, Chang, et al.
Published: (2025)
by: Gao, Chang, et al.
Published: (2025)
Shallow Convolution and Parallel Coarse‐To‐Fine Attention for Brain Signal Classification
by: Xiwen Qin, et al.
Published: (2025)
by: Xiwen Qin, et al.
Published: (2025)
Optimal Brain Connection: Towards Efficient Structural Pruning
by: Chen, Shaowu, et al.
Published: (2025)
by: Chen, Shaowu, et al.
Published: (2025)
CF-VLA: Efficient Coarse-to-Fine Action Generation for Vision-Language-Action Policies
by: Du, Fan, et al.
Published: (2026)
by: Du, Fan, et al.
Published: (2026)
Improved Diffusion-based Generative Model with Better Adversarial Robustness
by: Wang, Zekun, et al.
Published: (2025)
by: Wang, Zekun, et al.
Published: (2025)
REDSearcher: A Scalable and Cost-Efficient Framework for Long-Horizon Search Agents
by: Chu, Zheng, et al.
Published: (2026)
by: Chu, Zheng, et al.
Published: (2026)
FoleyDirector: Fine-Grained Temporal Steering for Video-to-Audio Generation via Structured Scripts
by: Li, You, et al.
Published: (2026)
by: Li, You, et al.
Published: (2026)
BrainStratify: Coarse-to-Fine Disentanglement of Intracranial Neural Dynamics
by: Zheng, Hui, et al.
Published: (2025)
by: Zheng, Hui, et al.
Published: (2025)
RankAdaptor: Hierarchical Rank Allocation for Efficient Fine-Tuning Pruned LLMs via Performance Model
by: Zhou, Changhai, et al.
Published: (2024)
by: Zhou, Changhai, et al.
Published: (2024)
Fovea Transformer: Efficient Long-Context Modeling with Structured Fine-to-Coarse Attention
by: He, Ziwei, et al.
Published: (2023)
by: He, Ziwei, et al.
Published: (2023)
Equivariant-Aware Structured Pruning for Efficient Edge Deployment: A Comprehensive Framework with Adaptive Fine-Tuning
by: Alnemari, Mohammed
Published: (2025)
by: Alnemari, Mohammed
Published: (2025)
Continual Gradient Low-Rank Projection Fine-Tuning for LLMs
by: Wang, Chenxu, et al.
Published: (2025)
by: Wang, Chenxu, et al.
Published: (2025)
Pruning for Protection: Increasing Jailbreak Resistance in Aligned LLMs Without Fine-Tuning
by: Hasan, Adib, et al.
Published: (2024)
by: Hasan, Adib, et al.
Published: (2024)
Cordycepin Ameliorates Kainic Acid‐Induced HT22 Cell Neurotoxicity by Activating GPR120‐Mediated Mitophagy
by: Yongzhi San, et al.
Published: (2025)
by: Yongzhi San, et al.
Published: (2025)
Adaptive Fulfillment Systems Under Incomplete Control: A Risk‐Oriented Modeling Framework
by: Lei Yang, et al.
Published: (2025)
by: Lei Yang, et al.
Published: (2025)
BeamAggR: Beam Aggregation Reasoning over Multi-source Knowledge for Multi-hop Question Answering
by: Chu, Zheng, et al.
Published: (2024)
by: Chu, Zheng, et al.
Published: (2024)
HCPM: Hierarchical Candidates Pruning for Efficient Detector-Free Matching
by: Chen, Ying, et al.
Published: (2024)
by: Chen, Ying, et al.
Published: (2024)
CFMS: A Coarse-to-Fine Multimodal Synthesis Framework for Enhanced Tabular Reasoning
by: Huang, Qixian, et al.
Published: (2026)
by: Huang, Qixian, et al.
Published: (2026)
All-in-One Tuning and Structural Pruning for Domain-Specific LLMs
by: Lu, Lei, et al.
Published: (2024)
by: Lu, Lei, et al.
Published: (2024)
ReDiPrune: Relevance-Diversity Pre-Projection Token Pruning for Efficient Multimodal LLMs
by: Yu, An, et al.
Published: (2026)
by: Yu, An, et al.
Published: (2026)
AutoPrune: Each Complexity Deserves a Pruning Policy
by: Wang, Hanshi, et al.
Published: (2025)
by: Wang, Hanshi, et al.
Published: (2025)
Similar Items
-
SmartTrim: Adaptive Tokens and Attention Pruning for Efficient Vision-Language Models
by: Wang, Zekun, et al.
Published: (2023) -
EffiVLM-BENCH: A Comprehensive Benchmark for Evaluating Training-Free Acceleration in Large Vision-Language Models
by: Wang, Zekun, et al.
Published: (2025) -
Divide-and-Conquer Meets Consensus: Unleashing the Power of Functions in Code Generation
by: Chen, Jingchang, et al.
Published: (2024) -
Self-Critique Guided Iterative Reasoning for Multi-hop Question Answering
by: Chu, Zheng, et al.
Published: (2025) -
MoGU: A Framework for Enhancing Safety of Open-Sourced LLMs While Preserving Their Usability
by: Du, Yanrui, et al.
Published: (2024)