The Impact of Quantization on the Robustness of Transformer-based Text Classifiers
Fuente:
arXiv
Saved in:
| Main Authors: | Neshaei, Seyed Parsa, Boreshban, Yasaman, Ghassem-Sani, Gholamreza, Mirroshandel, Seyed Abolghasem |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Towards Modeling Learner Performance with Large Language Models
by: Neshaei, Seyed Parsa, et al.
Published: (2024)
by: Neshaei, Seyed Parsa, et al.
Published: (2024)
Vision Transformers in Precision Agriculture: A Comprehensive Survey
by: Mehdipour, Saber, et al.
Published: (2025)
by: Mehdipour, Saber, et al.
Published: (2025)
Seq2Seq2Seq: Lossless Data Compression via Discrete Latent Transformers and Reinforcement Learning
by: Khodabandeh, Mahdi, et al.
Published: (2026)
by: Khodabandeh, Mahdi, et al.
Published: (2026)
Moving Beyond Review: Applying Language Models to Planning and Translation in Reflection
by: Neshaei, Seyed Parsa, et al.
Published: (2026)
by: Neshaei, Seyed Parsa, et al.
Published: (2026)
Using Large Multimodal Models to Extract Knowledge Components for Knowledge Tracing from Multimedia Question Information
by: Moon, Hyeongdon, et al.
Published: (2024)
by: Moon, Hyeongdon, et al.
Published: (2024)
Active Few-Shot Learning for Text Classification
by: Ahmadnia, Saeed, et al.
Published: (2025)
by: Ahmadnia, Saeed, et al.
Published: (2025)
From Semantic Roles to Opinion Roles: SRL Data Extraction for Multi-Task and Transfer Learning in Low-Resource ORL
by: Galdiani, Amirmohammad Omidi, et al.
Published: (2025)
by: Galdiani, Amirmohammad Omidi, et al.
Published: (2025)
Views Are My Own, but Also Yours: Benchmarking Theory of Mind Using Common Ground
by: Soubki, Adil, et al.
Published: (2024)
by: Soubki, Adil, et al.
Published: (2024)
Memory-Efficient Training for Text-Dependent SV with Independent Pre-trained Models
by: Farokh, Seyed Ali, et al.
Published: (2024)
by: Farokh, Seyed Ali, et al.
Published: (2024)
Thinking into the Future: Latent Lookahead Training for Transformers
by: Noci, Lorenzo, et al.
Published: (2026)
by: Noci, Lorenzo, et al.
Published: (2026)
Review, Remask, Refine (R3): Process-Guided Block Diffusion for Text Generation
by: Mounier, Nikita, et al.
Published: (2025)
by: Mounier, Nikita, et al.
Published: (2025)
One Goal, Many Challenges: Robust Preference Optimization Amid Content-Aware and Multi-Source Noise
by: Afzali, Amirabbas, et al.
Published: (2025)
by: Afzali, Amirabbas, et al.
Published: (2025)
Contrast-CAT: Contrasting Activations for Enhanced Interpretability in Transformer-based Text Classifiers
by: Han, Sungmin, et al.
Published: (2025)
by: Han, Sungmin, et al.
Published: (2025)
Word Sense Disambiguation in Persian: Can AI Finally Get It Right?
by: Ayyoubzadeh, Seyed Moein, et al.
Published: (2024)
by: Ayyoubzadeh, Seyed Moein, et al.
Published: (2024)
Logit Reweighting for Topic-Focused Summarization
by: Braun, Joschka, et al.
Published: (2025)
by: Braun, Joschka, et al.
Published: (2025)
Beyond Multiple Choice: Evaluating Steering Vectors for Summarization
by: Braun, Joschka, et al.
Published: (2025)
by: Braun, Joschka, et al.
Published: (2025)
Transparent Neighborhood Approximation for Text Classifier Explanation
by: Cai, Yi, et al.
Published: (2024)
by: Cai, Yi, et al.
Published: (2024)
Flip-Flop Consistency: Unsupervised Training for Robustness to Prompt Perturbations in LLMs
by: Hejabi, Parsa, et al.
Published: (2025)
by: Hejabi, Parsa, et al.
Published: (2025)
DISCERN: Decoding Systematic Errors in Natural Language for Text Classifiers
by: Menon, Rakesh R., et al.
Published: (2024)
by: Menon, Rakesh R., et al.
Published: (2024)
Robust Training of Vector Quantized Bottleneck Models
by: Łańcucki, Adrian, et al.
Published: (2020)
by: Łańcucki, Adrian, et al.
Published: (2020)
Explaining Text Classifiers with Counterfactual Representations
by: Lemberger, Pirmin, et al.
Published: (2024)
by: Lemberger, Pirmin, et al.
Published: (2024)
On the Quantization Robustness of Diffusion Language Models in Coding Benchmarks
by: Gupta, Aarav, et al.
Published: (2026)
by: Gupta, Aarav, et al.
Published: (2026)
FrameQuant: Flexible Low-Bit Quantization for Transformers
by: Adepu, Harshavardhan, et al.
Published: (2024)
by: Adepu, Harshavardhan, et al.
Published: (2024)
Towards Robust Few-Shot Text Classification Using Transformer Architectures and Dual Loss Strategies
by: Han, Xu, et al.
Published: (2025)
by: Han, Xu, et al.
Published: (2025)
Chain-of-Defensive-Thought: Structured Reasoning Elicits Robustness in Large Language Models against Reference Corruption
by: Wang, Wenxiao, et al.
Published: (2025)
by: Wang, Wenxiao, et al.
Published: (2025)
PSO Fuzzy XGBoost Classifier Boosted with Neural Gas Features on EEG Signals in Emotion Recognition
by: Mousavi, Seyed Muhammad Hossein
Published: (2024)
by: Mousavi, Seyed Muhammad Hossein
Published: (2024)
ylmmcl at Multilingual Text Detoxification 2025: Lexicon-Guided Detoxification and Classifier-Gated Rewriting
by: Lai-Lopez, Nicole, et al.
Published: (2025)
by: Lai-Lopez, Nicole, et al.
Published: (2025)
Geometry-Aware Decoding with Wasserstein-Regularized Truncation and Mass Penalties for Large Language Models
by: Davoodi, Arash Gholami, et al.
Published: (2026)
by: Davoodi, Arash Gholami, et al.
Published: (2026)
LATMiX: Learnable Affine Transformations for Microscaling Quantization of LLMs
by: Gordon, Ofir, et al.
Published: (2026)
by: Gordon, Ofir, et al.
Published: (2026)
Memory-Augmented Transformers: A Systematic Review from Neuroscience Principles to Enhanced Model Architectures
by: Omidi, Parsa, et al.
Published: (2025)
by: Omidi, Parsa, et al.
Published: (2025)
Self-Reported Confidence of Large Language Models in Gastroenterology: Analysis of Commercial, Open-Source, and Quantized Models
by: Naderi, Nariman, et al.
Published: (2025)
by: Naderi, Nariman, et al.
Published: (2025)
I Can't Believe It's Not Robust: Catastrophic Collapse of Safety Classifiers under Embedding Drift
by: Sahoo, Subramanyam, et al.
Published: (2026)
by: Sahoo, Subramanyam, et al.
Published: (2026)
Transformers Struggle to Learn to Search
by: Saparov, Abulhair, et al.
Published: (2024)
by: Saparov, Abulhair, et al.
Published: (2024)
Mitigating the Impact of Outlier Channels for Language Model Quantization with Activation Regularization
by: Nrusimha, Aniruddha, et al.
Published: (2024)
by: Nrusimha, Aniruddha, et al.
Published: (2024)
DALLMi: Domain Adaption for LLM-based Multi-label Classifier
by: Beţianu, Miruna, et al.
Published: (2024)
by: Beţianu, Miruna, et al.
Published: (2024)
Explaining Text Similarity in Transformer Models
by: Vasileiou, Alexandros, et al.
Published: (2024)
by: Vasileiou, Alexandros, et al.
Published: (2024)
Through a Compressed Lens: Investigating The Impact of Quantization on Factual Knowledge Recall
by: Wang, Qianli, et al.
Published: (2025)
by: Wang, Qianli, et al.
Published: (2025)
Enhanced Labeling Technique for Reddit Text and Fine-Tuned Longformer Models for Classifying Depression Severity in English and Luganda
by: Kimera, Richard, et al.
Published: (2024)
by: Kimera, Richard, et al.
Published: (2024)
LLMs Are Not Intelligent Thinkers: Introducing Mathematical Topic Tree Benchmark for Comprehensive Evaluation of LLMs
by: Davoodi, Arash Gholami, et al.
Published: (2024)
by: Davoodi, Arash Gholami, et al.
Published: (2024)
IRIS: An Iterative and Integrated Framework for Verifiable Causal Discovery in the Absence of Tabular Data
by: Feng, Tao, et al.
Published: (2025)
by: Feng, Tao, et al.
Published: (2025)
Similar Items
-
Towards Modeling Learner Performance with Large Language Models
by: Neshaei, Seyed Parsa, et al.
Published: (2024) -
Vision Transformers in Precision Agriculture: A Comprehensive Survey
by: Mehdipour, Saber, et al.
Published: (2025) -
Seq2Seq2Seq: Lossless Data Compression via Discrete Latent Transformers and Reinforcement Learning
by: Khodabandeh, Mahdi, et al.
Published: (2026) -
Moving Beyond Review: Applying Language Models to Planning and Translation in Reflection
by: Neshaei, Seyed Parsa, et al.
Published: (2026) -
Using Large Multimodal Models to Extract Knowledge Components for Knowledge Tracing from Multimedia Question Information
by: Moon, Hyeongdon, et al.
Published: (2024)