PGB: One-Shot Pruning for BERT via Weight Grouping and Permutation
Fuente:
arXiv
Saved in:
| Main Authors: | Lim, Hyemin, Lee, Jaeyeon, Choi, Dong-Wan |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
One-Shot Sensitivity-Aware Mixed Sparsity Pruning for Large Language Models
by: Shao, Hang, et al.
Published: (2023)
by: Shao, Hang, et al.
Published: (2023)
M-Wanda: Improving One-Shot Pruning for Multilingual LLMs
by: Choenni, Rochelle, et al.
Published: (2025)
by: Choenni, Rochelle, et al.
Published: (2025)
Mask-guided BERT for Few Shot Text Classification
by: Liao, Wenxiong, et al.
Published: (2023)
by: Liao, Wenxiong, et al.
Published: (2023)
Prompt-Based One-Shot Exact Length-Controlled Generation with LLMs
by: Xie, Juncheng, et al.
Published: (2025)
by: Xie, Juncheng, et al.
Published: (2025)
DYCP: Dynamic Context Pruning for Long-Form Dialogue with LLMs
by: Choi, Nayoung, et al.
Published: (2026)
by: Choi, Nayoung, et al.
Published: (2026)
NeoBERT: A Next-Generation BERT
by: Breton, Lola Le, et al.
Published: (2025)
by: Breton, Lola Le, et al.
Published: (2025)
All-in-One Tuning and Structural Pruning for Domain-Specific LLMs
by: Lu, Lei, et al.
Published: (2024)
by: Lu, Lei, et al.
Published: (2024)
PruneCD: Contrasting Pruned Self Model to Improve Decoding Factuality
by: Yu, Byeongho, et al.
Published: (2025)
by: Yu, Byeongho, et al.
Published: (2025)
Toward Adaptive Large Language Models Structured Pruning via Hybrid-grained Weight Importance Assessment
by: Liu, Jun, et al.
Published: (2024)
by: Liu, Jun, et al.
Published: (2024)
Systematic Weight Evaluation for Pruning Large Language Models: Enhancing Performance and Sustainability
by: Islam, Ashhadul, et al.
Published: (2025)
by: Islam, Ashhadul, et al.
Published: (2025)
Weighted Grouped Query Attention in Transformers
by: Chinnakonduru, Sai Sena, et al.
Published: (2024)
by: Chinnakonduru, Sai Sena, et al.
Published: (2024)
IM-BERT: Enhancing Robustness of BERT through the Implicit Euler Method
by: Kim, Mihyeon, et al.
Published: (2025)
by: Kim, Mihyeon, et al.
Published: (2025)
KliniskVestBERT: BERT Model Specialised to Norwegian Clinical Texts
by: Autenried, Christian, et al.
Published: (2026)
by: Autenried, Christian, et al.
Published: (2026)
NOVI : Chatbot System for University Novice with BERT and LLMs
by: Nam, Yoonji, et al.
Published: (2024)
by: Nam, Yoonji, et al.
Published: (2024)
KuBERT: Central Kurdish BERT Model and Its Application for Sentiment Analysis
by: Awlla, Kozhin muhealddin, et al.
Published: (2025)
by: Awlla, Kozhin muhealddin, et al.
Published: (2025)
Think Clearly: Improving Reasoning via Redundant Token Pruning
by: Choi, Daewon, et al.
Published: (2025)
by: Choi, Daewon, et al.
Published: (2025)
COPAL: Continual Pruning in Large Language Generative Models
by: Malla, Srikanth, et al.
Published: (2024)
by: Malla, Srikanth, et al.
Published: (2024)
Context Pruning for Coding Agents via Multi-Rubric Latent Reasoning
by: Wang, Jingjing, et al.
Published: (2026)
by: Wang, Jingjing, et al.
Published: (2026)
Grouped Sequency-arranged Rotation: Optimizing Rotation Transformation for Quantization for Free
by: Choi, Euntae, et al.
Published: (2025)
by: Choi, Euntae, et al.
Published: (2025)
Sparser Block-Sparse Attention via Token Permutation
by: Wang, Xinghao, et al.
Published: (2025)
by: Wang, Xinghao, et al.
Published: (2025)
What the Weight?! A Unified Framework for Zero-Shot Knowledge Composition
by: Holtermann, Carolin, et al.
Published: (2024)
by: Holtermann, Carolin, et al.
Published: (2024)
Prompt-based Depth Pruning of Large Language Models
by: Wee, Juyun, et al.
Published: (2025)
by: Wee, Juyun, et al.
Published: (2025)
G-SciEdBERT: A Contextualized LLM for Science Assessment Tasks in German
by: Latif, Ehsan, et al.
Published: (2024)
by: Latif, Ehsan, et al.
Published: (2024)
One-Shot Learning as Instruction Data Prospector for Large Language Models
by: Li, Yunshui, et al.
Published: (2023)
by: Li, Yunshui, et al.
Published: (2023)
Leveraging KV Similarity for Online Structured Pruning in LLMs
by: Lee, Jungmin, et al.
Published: (2025)
by: Lee, Jungmin, et al.
Published: (2025)
Ensemble BERT: A student social network text sentiment classification model based on ensemble learning and BERT architecture
by: Jiang, Kai, et al.
Published: (2024)
by: Jiang, Kai, et al.
Published: (2024)
ProgRAG: Hallucination-Resistant Progressive Retrieval and Reasoning over Knowledge Graphs
by: Park, Minbae, et al.
Published: (2025)
by: Park, Minbae, et al.
Published: (2025)
Rethinking Layer Redundancy: Calibration Matters More Than Search in LLM Depth Pruning
by: Kim, Minkyu, et al.
Published: (2026)
by: Kim, Minkyu, et al.
Published: (2026)
Chinese ModernBERT with Whole-Word Masking
by: Zhao, Zeyu, et al.
Published: (2025)
by: Zhao, Zeyu, et al.
Published: (2025)
Enriched BERT Embeddings for Scholarly Publication Classification
by: Wolff, Benjamin, et al.
Published: (2024)
by: Wolff, Benjamin, et al.
Published: (2024)
Arithmetic Reasoning with LLM: Prolog Generation & Permutation
by: Yang, Xiaocheng, et al.
Published: (2024)
by: Yang, Xiaocheng, et al.
Published: (2024)
TOFA: Training-Free One-Shot Federated Adaptation for Vision-Language Models
by: Zhang, Li, et al.
Published: (2025)
by: Zhang, Li, et al.
Published: (2025)
OneNet: A Fine-Tuning Free Framework for Few-Shot Entity Linking via Large Language Model Prompting
by: Liu, Xukai, et al.
Published: (2024)
by: Liu, Xukai, et al.
Published: (2024)
EvoP: Robust LLM Inference via Evolutionary Pruning
by: Wu, Shangyu, et al.
Published: (2025)
by: Wu, Shangyu, et al.
Published: (2025)
The Super Weight in Large Language Models
by: Yu, Mengxia, et al.
Published: (2024)
by: Yu, Mengxia, et al.
Published: (2024)
BERT-based model for Vietnamese Fact Verification Dataset
by: Tran, Bao, et al.
Published: (2025)
by: Tran, Bao, et al.
Published: (2025)
EuroBERT: Scaling Multilingual Encoders for European Languages
by: Boizard, Nicolas, et al.
Published: (2025)
by: Boizard, Nicolas, et al.
Published: (2025)
BBPOS: BERT-based Part-of-Speech Tagging for Uzbek
by: Bobojonova, Latofat, et al.
Published: (2025)
by: Bobojonova, Latofat, et al.
Published: (2025)
Analyzing Multi-Head Attention on Trojan BERT Models
by: Wang, Jingwei
Published: (2024)
by: Wang, Jingwei
Published: (2024)
BERT-VBD: Vietnamese Multi-Document Summarization Framework
by: Vuong, Tuan-Cuong, et al.
Published: (2024)
by: Vuong, Tuan-Cuong, et al.
Published: (2024)
Similar Items
-
One-Shot Sensitivity-Aware Mixed Sparsity Pruning for Large Language Models
by: Shao, Hang, et al.
Published: (2023) -
M-Wanda: Improving One-Shot Pruning for Multilingual LLMs
by: Choenni, Rochelle, et al.
Published: (2025) -
Mask-guided BERT for Few Shot Text Classification
by: Liao, Wenxiong, et al.
Published: (2023) -
Prompt-Based One-Shot Exact Length-Controlled Generation with LLMs
by: Xie, Juncheng, et al.
Published: (2025) -
DYCP: Dynamic Context Pruning for Long-Form Dialogue with LLMs
by: Choi, Nayoung, et al.
Published: (2026)