Are Compressed Language Models Less Subgroup Robust?
Fuente:
arXiv
Salvato in:
| Autori principali: | Gee, Leonidas, Zugarini, Andrea, Quadrianto, Novi |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Fast Vocabulary Transfer for Language Model Compression
di: Gee, Leonidas, et al.
Pubblicazione: (2024)
di: Gee, Leonidas, et al.
Pubblicazione: (2024)
Multi-word Tokenization for Sequence Compression
di: Gee, Leonidas, et al.
Pubblicazione: (2024)
di: Gee, Leonidas, et al.
Pubblicazione: (2024)
BUSTER: a "BUSiness Transaction Entity Recognition" dataset
di: Zugarini, Andrea, et al.
Pubblicazione: (2024)
di: Zugarini, Andrea, et al.
Pubblicazione: (2024)
Visual-Word Tokenizer: Beyond Fixed Sets of Tokens in Vision Transformers
di: Gee, Leonidas, et al.
Pubblicazione: (2024)
di: Gee, Leonidas, et al.
Pubblicazione: (2024)
Revisiting (Un)Fairness in Recourse by Minimizing Worst-Case Social Burden
di: Barrainkua, Ainhize, et al.
Pubblicazione: (2025)
di: Barrainkua, Ainhize, et al.
Pubblicazione: (2025)
Pay Less Attention to Function Words for Free Robustness of Vision-Language Models
di: Tian, Qiwei, et al.
Pubblicazione: (2025)
di: Tian, Qiwei, et al.
Pubblicazione: (2025)
Strategically Deceptive Model Deployment in Performative Prediction
di: Bautiste, Javier Sanguino, et al.
Pubblicazione: (2025)
di: Bautiste, Javier Sanguino, et al.
Pubblicazione: (2025)
Building a Strong Instruction Language Model for a Less-Resourced Language
di: Vreš, Domen, et al.
Pubblicazione: (2026)
di: Vreš, Domen, et al.
Pubblicazione: (2026)
An energy-based comparative analysis of common approaches to text classification in the Legal domain
di: Gultekin, Sinan, et al.
Pubblicazione: (2023)
di: Gultekin, Sinan, et al.
Pubblicazione: (2023)
Dissecting Performative Prediction: A Comprehensive Survey
di: Kehrenberg, Thomas, et al.
Pubblicazione: (2026)
di: Kehrenberg, Thomas, et al.
Pubblicazione: (2026)
Dancing in the Shadows: Harnessing Ambiguity for Fairer Classifiers
di: Barrainkua, Ainhize, et al.
Pubblicazione: (2024)
di: Barrainkua, Ainhize, et al.
Pubblicazione: (2024)
Proxy Compression for Language Modeling
di: Zheng, Lin, et al.
Pubblicazione: (2026)
di: Zheng, Lin, et al.
Pubblicazione: (2026)
Does Training on Synthetic Data Make Models Less Robust?
di: Zhang, Lingze, et al.
Pubblicazione: (2025)
di: Zhang, Lingze, et al.
Pubblicazione: (2025)
Safe Fairness Guarantees Without Demographics in Classification: Spectral Uncertainty Set Perspective
di: Barrainkua, Ainhize, et al.
Pubblicazione: (2026)
di: Barrainkua, Ainhize, et al.
Pubblicazione: (2026)
Less but Better: Parameter-Efficient Fine-Tuning of Large Language Models for Personality Detection
di: Shen, Lingzhi, et al.
Pubblicazione: (2025)
di: Shen, Lingzhi, et al.
Pubblicazione: (2025)
Less Diverse, Less Safe: The Indirect But Pervasive Risk of Test-Time Scaling in Large Language Models
di: Nahin, Shahriar Kabir, et al.
Pubblicazione: (2025)
di: Nahin, Shahriar Kabir, et al.
Pubblicazione: (2025)
Less is KEN: a Universal and Simple Non-Parametric Pruning Algorithm for Large Language Models
di: Mastromattei, Michele, et al.
Pubblicazione: (2024)
di: Mastromattei, Michele, et al.
Pubblicazione: (2024)
Privacy and Accuracy Implications of Model Complexity and Integration in Heterogeneous Federated Learning
di: Németh, Gergely Dániel, et al.
Pubblicazione: (2023)
di: Németh, Gergely Dániel, et al.
Pubblicazione: (2023)
Less is More: Local Intrinsic Dimensions of Contextual Language Models
di: Ruppik, Benjamin Matthias, et al.
Pubblicazione: (2025)
di: Ruppik, Benjamin Matthias, et al.
Pubblicazione: (2025)
Compressed Context Memory For Online Language Model Interaction
di: Kim, Jang-Hyun, et al.
Pubblicazione: (2023)
di: Kim, Jang-Hyun, et al.
Pubblicazione: (2023)
Extreme Compression of Large Language Models via Additive Quantization
di: Egiazarian, Vage, et al.
Pubblicazione: (2024)
di: Egiazarian, Vage, et al.
Pubblicazione: (2024)
Radio: Rate-Distortion Optimization for Large Language Model Compression
di: Young, Sean I.
Pubblicazione: (2025)
di: Young, Sean I.
Pubblicazione: (2025)
MiniDisc: Minimal Distillation Schedule for Language Model Compression
di: Zhang, Chen, et al.
Pubblicazione: (2022)
di: Zhang, Chen, et al.
Pubblicazione: (2022)
Layer-wise Importance Matters: Less Memory for Better Performance in Parameter-efficient Fine-tuning of Large Language Models
di: Yao, Kai, et al.
Pubblicazione: (2024)
di: Yao, Kai, et al.
Pubblicazione: (2024)
SliceGPT: Compress Large Language Models by Deleting Rows and Columns
di: Ashkboos, Saleh, et al.
Pubblicazione: (2024)
di: Ashkboos, Saleh, et al.
Pubblicazione: (2024)
FoldGPT: Simple and Effective Large Language Model Compression Scheme
di: Liu, Songwei, et al.
Pubblicazione: (2024)
di: Liu, Songwei, et al.
Pubblicazione: (2024)
Foundations of Large Language Model Compression -- Part 1: Weight Quantization
di: Young, Sean I.
Pubblicazione: (2024)
di: Young, Sean I.
Pubblicazione: (2024)
Iterative Layer-wise Distillation for Efficient Compression of Large Language Models
di: Kovalev, Grigory, et al.
Pubblicazione: (2025)
di: Kovalev, Grigory, et al.
Pubblicazione: (2025)
ILRe: Intermediate Layer Retrieval for Context Compression in Causal Language Models
di: Liang, Manlai, et al.
Pubblicazione: (2025)
di: Liang, Manlai, et al.
Pubblicazione: (2025)
Robust and Scalable Model Editing for Large Language Models
di: Chen, Yingfa, et al.
Pubblicazione: (2024)
di: Chen, Yingfa, et al.
Pubblicazione: (2024)
Model Hemorrhage and the Robustness Limits of Large Language Models
di: Ma, Ziyang, et al.
Pubblicazione: (2025)
di: Ma, Ziyang, et al.
Pubblicazione: (2025)
On the Compressibility of Quantized Large Language Models
di: Mao, Yu, et al.
Pubblicazione: (2024)
di: Mao, Yu, et al.
Pubblicazione: (2024)
Basis Sharing: Cross-Layer Parameter Sharing for Large Language Model Compression
di: Wang, Jingcun, et al.
Pubblicazione: (2024)
di: Wang, Jingcun, et al.
Pubblicazione: (2024)
PocketLLM: Ultimate Compression of Large Language Models via Meta Networks
di: Tian, Ye, et al.
Pubblicazione: (2025)
di: Tian, Ye, et al.
Pubblicazione: (2025)
LoRA Learns Less and Forgets Less
di: Biderman, Dan, et al.
Pubblicazione: (2024)
di: Biderman, Dan, et al.
Pubblicazione: (2024)
Energy-Based Reward Models for Robust Language Model Alignment
di: Lochab, Anamika, et al.
Pubblicazione: (2025)
di: Lochab, Anamika, et al.
Pubblicazione: (2025)
SVD-LLM: Truncation-aware Singular Value Decomposition for Large Language Model Compression
di: Wang, Xin, et al.
Pubblicazione: (2024)
di: Wang, Xin, et al.
Pubblicazione: (2024)
ByteFlow: Language Modeling through Adaptive Byte Compression without a Tokenizer
di: Deng, Chunyuan, et al.
Pubblicazione: (2026)
di: Deng, Chunyuan, et al.
Pubblicazione: (2026)
Lossless Compression of Large Language Model-Generated Text via Next-Token Prediction
di: Mao, Yu, et al.
Pubblicazione: (2025)
di: Mao, Yu, et al.
Pubblicazione: (2025)
Saten: Sparse Augmented Tensor Networks for Post-Training Compression of Large Language Models
di: Solgi, Ryan, et al.
Pubblicazione: (2025)
di: Solgi, Ryan, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Fast Vocabulary Transfer for Language Model Compression
di: Gee, Leonidas, et al.
Pubblicazione: (2024) -
Multi-word Tokenization for Sequence Compression
di: Gee, Leonidas, et al.
Pubblicazione: (2024) -
BUSTER: a "BUSiness Transaction Entity Recognition" dataset
di: Zugarini, Andrea, et al.
Pubblicazione: (2024) -
Visual-Word Tokenizer: Beyond Fixed Sets of Tokens in Vision Transformers
di: Gee, Leonidas, et al.
Pubblicazione: (2024) -
Revisiting (Un)Fairness in Recourse by Minimizing Worst-Case Social Burden
di: Barrainkua, Ainhize, et al.
Pubblicazione: (2025)