Can Post-Training Quantization Benefit from an Additional QLoRA Integration?
Fuente:
arXiv
Saved in:
| Main Authors: | Zhu, Xiliang, Khasanova, Elena, Chen, Cheng |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Bailong: Bilingual Transfer Learning based on QLoRA and Zip-tie Embedding
by: Chen, Lung-Chuan, et al.
Published: (2024)
by: Chen, Lung-Chuan, et al.
Published: (2024)
Internalizing Tool Knowledge in Small Language Models via QLoRA Fine-Tuning
by: Shemla, Yuval, et al.
Published: (2026)
by: Shemla, Yuval, et al.
Published: (2026)
Speaker attribution in German parliamentary debates with QLoRA-adapted large language models
by: Bornheim, Tobias, et al.
Published: (2023)
by: Bornheim, Tobias, et al.
Published: (2023)
Output Composability of QLoRA PEFT Modules for Plug-and-Play Attribute-Controlled Text Generation
by: Lorandi, Michela, et al.
Published: (2026)
by: Lorandi, Michela, et al.
Published: (2026)
Harnessing Earnings Reports for Stock Predictions: A QLoRA-Enhanced LLM Approach
by: Ni, Haowei, et al.
Published: (2024)
by: Ni, Haowei, et al.
Published: (2024)
From Baselines to Preferences: A Comparative Study of LoRA/QLoRA and Preference Optimization for Mental Health Text Classification
by: Arcan, Mihael
Published: (2026)
by: Arcan, Mihael
Published: (2026)
Exploring Fact Memorization and Style Imitation in LLMs Using QLoRA: An Experimental Study and Quality Assessment Methods
by: Vyborov, Eugene, et al.
Published: (2024)
by: Vyborov, Eugene, et al.
Published: (2024)
QU-NLP at QIAS 2026: Multi-Stage QLoRA Fine-Tuning for Arabic Islamic Inheritance Reasoning
by: AL-Smadi, Mohammad
Published: (2026)
by: AL-Smadi, Mohammad
Published: (2026)
CARE: A QLoRA-Fine Tuned Multi-Domain Chatbot With Fast Learning On Minimal Hardware
by: Dutta, Ankit, et al.
Published: (2025)
by: Dutta, Ankit, et al.
Published: (2025)
Adapting Large Language Models to a Low-Resource Agglutinative Language: A Comparative Study of LoRA and QLoRA for Bashkir
by: Arabov, Mullosharaf K., et al.
Published: (2026)
by: Arabov, Mullosharaf K., et al.
Published: (2026)
Lightweight Clinical Decision Support System using QLoRA-Fine-Tuned LLMs and Retrieval-Augmented Generation
by: Ansari, Mohammad Shoaib, et al.
Published: (2025)
by: Ansari, Mohammad Shoaib, et al.
Published: (2025)
Fine-Tuning Large Language Models with QLoRA for Offensive Language Detection in Roman Urdu-English Code-Mixed Text
by: Hussain, Nisar, et al.
Published: (2025)
by: Hussain, Nisar, et al.
Published: (2025)
Overcoming linguistic barriers in code assistants: creating a QLoRA adapter to improve support for Russian-language code writing instructions
by: Pronin, C. B., et al.
Published: (2024)
by: Pronin, C. B., et al.
Published: (2024)
Parameter-Efficient Fine-Tuning for HAR: Integrating LoRA and QLoRA into Transformer Models
by: Seregina, Irina, et al.
Published: (2025)
by: Seregina, Irina, et al.
Published: (2025)
QU-NLP at ArchEHR-QA 2026: Two-Stage QLoRA Fine-Tuning of Qwen3-4B for Patient-Oriented Clinical Question Answering and Evidence Sentence Alignment
by: AL-Smadi, Mohammad
Published: (2026)
by: AL-Smadi, Mohammad
Published: (2026)
Tiny Titans: Can Smaller Large Language Models Punch Above Their Weight in the Real World for Meeting Summarization?
by: Fu, Xue-Yong, et al.
Published: (2024)
by: Fu, Xue-Yong, et al.
Published: (2024)
EdgeNav-QE: QLoRA Quantization and Dynamic Early Exit for LAM-based Navigation on Edge Devices
by: Liu, Mengyun, et al.
Published: (2026)
by: Liu, Mengyun, et al.
Published: (2026)
Efficient Telecom Specific LLM: TSLAM-Mini with QLoRA and Digital Twin Data
by: Ethiraj, Vignesh, et al.
Published: (2025)
by: Ethiraj, Vignesh, et al.
Published: (2025)
Viz: A QLoRA-based Copyright Marketplace for Legally Compliant Generative AI
by: Sarkar, Dipankar
Published: (2023)
by: Sarkar, Dipankar
Published: (2023)
LLaMA-XR: A Novel Framework for Radiology Report Generation using LLaMA and QLoRA Fine Tuning
by: Jahangir, Md. Zihad Bin, et al.
Published: (2025)
by: Jahangir, Md. Zihad Bin, et al.
Published: (2025)
Profiling LoRA/QLoRA Fine-Tuning Efficiency on Consumer GPUs: An RTX 4060 Case Study
by: Avinash, MSR
Published: (2025)
by: Avinash, MSR
Published: (2025)
DACIP-RC: Domain Adaptive Continual Instruction Pre-Training via Reading Comprehension on Business Conversations
by: Khasanova, Elena, et al.
Published: (2025)
by: Khasanova, Elena, et al.
Published: (2025)
ROMA: a Read-Only-Memory-based Accelerator for QLoRA-based On-Device LLM
by: Wang, Wenqiang, et al.
Published: (2025)
by: Wang, Wenqiang, et al.
Published: (2025)
Query-OPT: Optimizing Inference of Large Language Models via Multi-Query Instructions in Meeting Summarization
by: Laskar, Md Tahmid Rahman, et al.
Published: (2024)
by: Laskar, Md Tahmid Rahman, et al.
Published: (2024)
The Uneven Impact of Post-Training Quantization in Machine Translation
by: Marie, Benjamin, et al.
Published: (2025)
by: Marie, Benjamin, et al.
Published: (2025)
QuantMoE-Bench: Examining Post-Training Quantization for Mixture-of-Experts
by: Li, Pingzhi, et al.
Published: (2024)
by: Li, Pingzhi, et al.
Published: (2024)
How Accurate Are LLMs at Multi-Question Answering on Conversational Transcripts?
by: Zhu, Xiliang, et al.
Published: (2025)
by: Zhu, Xiliang, et al.
Published: (2025)
DACP: Domain-Adaptive Continual Pre-Training of Large Language Models for Phone Conversation Summarization
by: Fu, Xue-Yong, et al.
Published: (2025)
by: Fu, Xue-Yong, et al.
Published: (2025)
Can Post-Training Transform LLMs into Causal Reasoners?
by: Chen, Junqi, et al.
Published: (2026)
by: Chen, Junqi, et al.
Published: (2026)
Scaling Laws for Post Training Quantized Large Language Models
by: Xu, Zifei, et al.
Published: (2024)
by: Xu, Zifei, et al.
Published: (2024)
QuAILoRA: Quantization-Aware Initialization for LoRA
by: Lawton, Neal, et al.
Published: (2024)
by: Lawton, Neal, et al.
Published: (2024)
VLMQ: Token Saliency-Driven Post-Training Quantization for Vision-language Models
by: Xue, Yufei, et al.
Published: (2025)
by: Xue, Yufei, et al.
Published: (2025)
AdpQ: A Zero-shot Calibration Free Adaptive Post Training Quantization Method for LLMs
by: Ghaffari, Alireza, et al.
Published: (2024)
by: Ghaffari, Alireza, et al.
Published: (2024)
The Model Arena for Cross-lingual Sentiment Analysis: A Comparative Study in the Era of Large Language Models
by: Zhu, Xiliang, et al.
Published: (2024)
by: Zhu, Xiliang, et al.
Published: (2024)
Enhancing Post-Training Quantization via Future Activation Awareness
by: Lv, Zheqi, et al.
Published: (2026)
by: Lv, Zheqi, et al.
Published: (2026)
Exploring parameter-efficient fine-tuning (PEFT) of billion-parameter vision models with QLoRA and DoRA: insights into generalization for limited-data image classification under a 98:1 test-to-train regime
by: Yang, Haiyu, et al.
Published: (2026)
by: Yang, Haiyu, et al.
Published: (2026)
EAQuant: Enhancing Post-Training Quantization for MoE Models via Expert-Aware Optimization
by: Fu, Zhongqian, et al.
Published: (2025)
by: Fu, Zhongqian, et al.
Published: (2025)
BiLLM: Pushing the Limit of Post-Training Quantization for LLMs
by: Huang, Wei, et al.
Published: (2024)
by: Huang, Wei, et al.
Published: (2024)
SignRoundV2: Toward Closing the Performance Gap in Extremely Low-Bit Post-Training Quantization for LLMs
by: Cheng, Wenhua, et al.
Published: (2025)
by: Cheng, Wenhua, et al.
Published: (2025)
AlphaLoRA: Assigning LoRA Experts Based on Layer Training Quality
by: Qing, Peijun, et al.
Published: (2024)
by: Qing, Peijun, et al.
Published: (2024)
Similar Items
-
Bailong: Bilingual Transfer Learning based on QLoRA and Zip-tie Embedding
by: Chen, Lung-Chuan, et al.
Published: (2024) -
Internalizing Tool Knowledge in Small Language Models via QLoRA Fine-Tuning
by: Shemla, Yuval, et al.
Published: (2026) -
Speaker attribution in German parliamentary debates with QLoRA-adapted large language models
by: Bornheim, Tobias, et al.
Published: (2023) -
Output Composability of QLoRA PEFT Modules for Plug-and-Play Attribute-Controlled Text Generation
by: Lorandi, Michela, et al.
Published: (2026) -
Harnessing Earnings Reports for Stock Predictions: A QLoRA-Enhanced LLM Approach
by: Ni, Haowei, et al.
Published: (2024)