Accurate Retraining-free Pruning for Pretrained Encoder-based Language Models
Fuente:
arXiv
Salvato in:
| Autori principali: | Park, Seungcheol, Choi, Hojun, Kang, U |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2023
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Accurate Sublayer Pruning for Large Language Models by Exploiting Latency and Tunability Information
di: Park, Seungcheol, et al.
Pubblicazione: (2025)
di: Park, Seungcheol, et al.
Pubblicazione: (2025)
A Comprehensive Survey of Compression Algorithms for Language Models
di: Park, Seungcheol, et al.
Pubblicazione: (2024)
di: Park, Seungcheol, et al.
Pubblicazione: (2024)
Unifying Uniform and Binary-coding Quantization for Accurate Compression of Large Language Models
di: Park, Seungcheol, et al.
Pubblicazione: (2025)
di: Park, Seungcheol, et al.
Pubblicazione: (2025)
"As Eastern Powers, I will veto." : An Investigation of Nation-level Bias of Large Language Models in International Relations
di: Choi, Jonghyeon, et al.
Pubblicazione: (2025)
di: Choi, Jonghyeon, et al.
Pubblicazione: (2025)
Exploiting Pre-trained Encoder-Decoder Transformers for Sequence-to-Sequence Constituent Parsing
di: Fernández-González, Daniel, et al.
Pubblicazione: (2026)
di: Fernández-González, Daniel, et al.
Pubblicazione: (2026)
Evaluating Pixel Language Models on Non-Standardized Languages
di: Muñoz-Ortiz, Alberto, et al.
Pubblicazione: (2024)
di: Muñoz-Ortiz, Alberto, et al.
Pubblicazione: (2024)
The Superalignment of Superhuman Intelligence with Large Language Models
di: Huang, Minlie, et al.
Pubblicazione: (2024)
di: Huang, Minlie, et al.
Pubblicazione: (2024)
An Unforgeable Publicly Verifiable Watermark for Large Language Models
di: Liu, Aiwei, et al.
Pubblicazione: (2023)
di: Liu, Aiwei, et al.
Pubblicazione: (2023)
Exploring State Tracking Capabilities of Large Language Models
di: Rezaee, Kiamehr, et al.
Pubblicazione: (2025)
di: Rezaee, Kiamehr, et al.
Pubblicazione: (2025)
A Survey of Text Watermarking in the Era of Large Language Models
di: Liu, Aiwei, et al.
Pubblicazione: (2023)
di: Liu, Aiwei, et al.
Pubblicazione: (2023)
Distilling Large Language Models for Efficient Clinical Information Extraction
di: Vedula, Karthik S., et al.
Pubblicazione: (2024)
di: Vedula, Karthik S., et al.
Pubblicazione: (2024)
Adaptive Steering and Remasking for Safe Generation in Diffusion Language Models
di: Lee, Yejin, et al.
Pubblicazione: (2026)
di: Lee, Yejin, et al.
Pubblicazione: (2026)
GATE: Graph-based Adaptive Tool Evolution Across Diverse Tasks
di: Luo, Jianwen, et al.
Pubblicazione: (2025)
di: Luo, Jianwen, et al.
Pubblicazione: (2025)
Towards Effective and Efficient Continual Pre-training of Large Language Models
di: Chen, Jie, et al.
Pubblicazione: (2024)
di: Chen, Jie, et al.
Pubblicazione: (2024)
Setting Standards in Turkish NLP: TR-MMLU for Large Language Model Evaluation
di: Bayram, M. Ali, et al.
Pubblicazione: (2024)
di: Bayram, M. Ali, et al.
Pubblicazione: (2024)
d-TreeRPO: Towards More Reliable Policy Optimization for Diffusion Language Models
di: Pan, Leyi, et al.
Pubblicazione: (2025)
di: Pan, Leyi, et al.
Pubblicazione: (2025)
Direct Large Language Model Alignment Through Self-Rewarding Contrastive Prompt Distillation
di: Liu, Aiwei, et al.
Pubblicazione: (2024)
di: Liu, Aiwei, et al.
Pubblicazione: (2024)
Large Language Models(LLMs) on Tabular Data: Prediction, Generation, and Understanding -- A Survey
di: Fang, Xi, et al.
Pubblicazione: (2024)
di: Fang, Xi, et al.
Pubblicazione: (2024)
WeDLM: Reconciling Diffusion Language Models with Standard Causal Attention for Fast Inference
di: Liu, Aiwei, et al.
Pubblicazione: (2025)
di: Liu, Aiwei, et al.
Pubblicazione: (2025)
Omni-SafetyBench: A Benchmark for Safety Evaluation of Audio-Visual Large Language Models
di: Pan, Leyi, et al.
Pubblicazione: (2025)
di: Pan, Leyi, et al.
Pubblicazione: (2025)
The Paradox of Poetic Intent in Back-Translation: Evaluating the Quality of Large Language Models in Chinese Translation
di: Weigang, Li, et al.
Pubblicazione: (2025)
di: Weigang, Li, et al.
Pubblicazione: (2025)
Unveiling Attractor Cycles in Large Language Models: A Dynamical Systems View of Successive Paraphrasing
di: Wang, Zhilin, et al.
Pubblicazione: (2025)
di: Wang, Zhilin, et al.
Pubblicazione: (2025)
NurValues: Real-World Nursing Values Evaluation for Large Language Models in Clinical Context
di: Yao, Ben, et al.
Pubblicazione: (2025)
di: Yao, Ben, et al.
Pubblicazione: (2025)
Trusted Uncertainty in Large Language Models: A Unified Framework for Confidence Calibration and Risk-Controlled Refusal
di: Oehri, Markus, et al.
Pubblicazione: (2025)
di: Oehri, Markus, et al.
Pubblicazione: (2025)
Math Natural Language Inference: this should be easy!
di: de Paiva, Valeria, et al.
Pubblicazione: (2025)
di: de Paiva, Valeria, et al.
Pubblicazione: (2025)
A Survey on Natural Language Counterfactual Generation
di: Wang, Yongjie, et al.
Pubblicazione: (2024)
di: Wang, Yongjie, et al.
Pubblicazione: (2024)
Culturally-Nuanced Story Generation for Reasoning in Low-Resource Languages: The Case of Javanese and Sundanese
di: Pranida, Salsabila Zahirah, et al.
Pubblicazione: (2025)
di: Pranida, Salsabila Zahirah, et al.
Pubblicazione: (2025)
Tracking linguistic information in transformer-based sentence embeddings through targeted sparsification
di: Nastase, Vivi, et al.
Pubblicazione: (2024)
di: Nastase, Vivi, et al.
Pubblicazione: (2024)
Prior-based Noisy Text Data Filtering: Fast and Strong Alternative For Perplexity
di: Seo, Yeongbin, et al.
Pubblicazione: (2025)
di: Seo, Yeongbin, et al.
Pubblicazione: (2025)
Co-NAML-LSTUR: A Combined Model with Attentive Multi-View Learning and Long- and Short-term User Representations for News Recommendation
di: Nguyen, Minh Hoang, et al.
Pubblicazione: (2025)
di: Nguyen, Minh Hoang, et al.
Pubblicazione: (2025)
Profiling German Text Simplification with Interpretable Model-Fingerprints
di: Klöser, Lars, et al.
Pubblicazione: (2026)
di: Klöser, Lars, et al.
Pubblicazione: (2026)
CLMN: Concept based Language Models via Neural Symbolic Reasoning
di: Yang, Yibo
Pubblicazione: (2025)
di: Yang, Yibo
Pubblicazione: (2025)
On the Robustness of Document-Level Relation Extraction Models to Entity Name Variations
di: Meng, Shiao, et al.
Pubblicazione: (2024)
di: Meng, Shiao, et al.
Pubblicazione: (2024)
PaperAudit-Bench: Benchmarking Error Detection in Research Papers for Critical Automated Peer Review
di: Tu, Songjun, et al.
Pubblicazione: (2026)
di: Tu, Songjun, et al.
Pubblicazione: (2026)
Give it Space! Explicit Disentangling of Positional and Semantic Representations in Encoders
di: Lequeu, Pierre-Antoine, et al.
Pubblicazione: (2026)
di: Lequeu, Pierre-Antoine, et al.
Pubblicazione: (2026)
Prompting Encoder Models for Zero-Shot Classification: A Cross-Domain Study in Italian
di: Auriemma, Serena, et al.
Pubblicazione: (2024)
di: Auriemma, Serena, et al.
Pubblicazione: (2024)
Fine-tuning of Large Language Models for Constituency Parsing Using a Sequence to Sequence Approach
di: Delgado, Francisco Jose Cortes, et al.
Pubblicazione: (2025)
di: Delgado, Francisco Jose Cortes, et al.
Pubblicazione: (2025)
How Human-Like Are Large Language Models? A Register-Aware Linguistic Evaluation Framework
di: Nieth, Björn, et al.
Pubblicazione: (2026)
di: Nieth, Björn, et al.
Pubblicazione: (2026)
Contrasting Linguistic Patterns in Human and LLM-Generated News Text
di: Muñoz-Ortiz, Alberto, et al.
Pubblicazione: (2023)
di: Muñoz-Ortiz, Alberto, et al.
Pubblicazione: (2023)
A Stochastic Analysis of the Linguistic Provenance of English Place Names
di: Dalvean, Michael
Pubblicazione: (2023)
di: Dalvean, Michael
Pubblicazione: (2023)
Documenti analoghi
-
Accurate Sublayer Pruning for Large Language Models by Exploiting Latency and Tunability Information
di: Park, Seungcheol, et al.
Pubblicazione: (2025) -
A Comprehensive Survey of Compression Algorithms for Language Models
di: Park, Seungcheol, et al.
Pubblicazione: (2024) -
Unifying Uniform and Binary-coding Quantization for Accurate Compression of Large Language Models
di: Park, Seungcheol, et al.
Pubblicazione: (2025) -
"As Eastern Powers, I will veto." : An Investigation of Nation-level Bias of Large Language Models in International Relations
di: Choi, Jonghyeon, et al.
Pubblicazione: (2025) -
Exploiting Pre-trained Encoder-Decoder Transformers for Sequence-to-Sequence Constituent Parsing
di: Fernández-González, Daniel, et al.
Pubblicazione: (2026)