ProjQ: Project-and-Quantize for Adapter-Aware LLM Compression
Fuente:
arXiv
Saved in:
| Main Authors: | Yu, Wneya, Zhang, Chao, Wang, Li, Lasaulce, Samson, Debbah, Merouane |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
BAQ: Efficient Bit Allocation Quantization for Large Language Models
by: Zhang, Chao, et al.
Published: (2025)
by: Zhang, Chao, et al.
Published: (2025)
Strategic Federated Learning: Application to Smart Meter Data Clustering
by: Mohamad, Hassan, et al.
Published: (2024)
by: Mohamad, Hassan, et al.
Published: (2024)
Goal-Oriented State Information Compression for Linear Dynamical System Control
by: Wang, Li, et al.
Published: (2024)
by: Wang, Li, et al.
Published: (2024)
RF-GPT: Teaching AI to See the Wireless World
by: Zou, Hang, et al.
Published: (2026)
by: Zou, Hang, et al.
Published: (2026)
Per-antenna power constraints: constructing Pareto-optimal precoders with cubic complexity under non-negligible noise conditions
by: Petrov, Sergey, et al.
Published: (2025)
by: Petrov, Sergey, et al.
Published: (2025)
Quantization-Aware Collaborative Inference for Large Embodied AI Models
by: Lyu, Zhonghao, et al.
Published: (2026)
by: Lyu, Zhonghao, et al.
Published: (2026)
Neural Precoding in Complex Projective Spaces
by: Abdullah, Zaid, et al.
Published: (2026)
by: Abdullah, Zaid, et al.
Published: (2026)
Large Language Models as Bidding Agents in Repeated HetNet Auction
by: Lotfi, Ismail, et al.
Published: (2026)
by: Lotfi, Ismail, et al.
Published: (2026)
From LLM Reasoning to Autonomous AI Agents: A Comprehensive Review
by: Ferrag, Mohamed Amine, et al.
Published: (2025)
by: Ferrag, Mohamed Amine, et al.
Published: (2025)
DAQ: Delta-Aware Quantization for Post-Training LLM Weight Compression
by: Yu, Xiaoming, et al.
Published: (2026)
by: Yu, Xiaoming, et al.
Published: (2026)
Communication-Efficient Zero-Order and First-Order Federated Learning Methods over Wireless Networks
by: Assaad, Mohamad, et al.
Published: (2025)
by: Assaad, Mohamad, et al.
Published: (2025)
Reading Radio from Camera: Visually-Grounded, Lightweight, and Interpretable RSSI Prediction
by: Yan, Sen, et al.
Published: (2025)
by: Yan, Sen, et al.
Published: (2025)
Non-Identical Diffusion Models in MIMO-OFDM Channel Generation
by: Yang, Yuzhi, et al.
Published: (2025)
by: Yang, Yuzhi, et al.
Published: (2025)
Reasoning Beyond Limits: Advances and Open Problems for LLMs
by: Ferrag, Mohamed Amine, et al.
Published: (2025)
by: Ferrag, Mohamed Amine, et al.
Published: (2025)
Scaling-Aware Adapter for Structure-Grounded LLM Reasoning
by: Jing, Zihao, et al.
Published: (2026)
by: Jing, Zihao, et al.
Published: (2026)
Lightweight Fairness for LLM-Based Recommendations via Kernelized Projection and Gated Adapters
by: Cui, Nan, et al.
Published: (2026)
by: Cui, Nan, et al.
Published: (2026)
Q-Adapter: Customizing Pre-trained LLMs to New Preferences with Forgetting Mitigation
by: Li, Yi-Chen, et al.
Published: (2024)
by: Li, Yi-Chen, et al.
Published: (2024)
AutoQRA: Joint Optimization of Mixed-Precision Quantization and Low-rank Adapters for Efficient LLM Fine-Tuning
by: Zhou, Changhai, et al.
Published: (2026)
by: Zhou, Changhai, et al.
Published: (2026)
Generative AI for RF Sensing in IoT systems
by: Wang, Li, et al.
Published: (2024)
by: Wang, Li, et al.
Published: (2024)
Quantization-Aware Regularizers for Deep Neural Networks Compression
by: Malchiodi, Dario, et al.
Published: (2026)
by: Malchiodi, Dario, et al.
Published: (2026)
GenAINet: Enabling Wireless Collective Intelligence via Knowledge Transfer and Reasoning
by: Zou, Hang, et al.
Published: (2024)
by: Zou, Hang, et al.
Published: (2024)
Leech Lattice Vector Quantization for Efficient LLM Compression
by: van der Ouderaa, Tycho F. A., et al.
Published: (2026)
by: van der Ouderaa, Tycho F. A., et al.
Published: (2026)
Learning Grouped Lattice Vector Quantizers for Low-Bit LLM Compression
by: Zhang, Xi, et al.
Published: (2025)
by: Zhang, Xi, et al.
Published: (2025)
WirelessMathBench: A Mathematical Modeling Benchmark for LLMs in Wireless Communications
by: Li, Xin, et al.
Published: (2025)
by: Li, Xin, et al.
Published: (2025)
Seeing Radio: From Zero RF Priors to Explainable Modulation Recognition with Vision Language Models
by: Zou, Hang, et al.
Published: (2026)
by: Zou, Hang, et al.
Published: (2026)
TesseraQ: Ultra Low-Bit LLM Post-Training Quantization with Block Reconstruction
by: Li, Yuhang, et al.
Published: (2024)
by: Li, Yuhang, et al.
Published: (2024)
AWP: Activation-Aware Weight Pruning and Quantization with Projected Gradient Descent
by: Liu, Jing, et al.
Published: (2025)
by: Liu, Jing, et al.
Published: (2025)
Q-GaLore: Quantized GaLore with INT4 Projection and Layer-Adaptive Low-Rank Gradients
by: Zhang, Zhenyu, et al.
Published: (2024)
by: Zhang, Zhenyu, et al.
Published: (2024)
Data-driven Energy Efficiency Modelling in Large-scale Networks: An Expert Knowledge and ML-based Approach
by: López-Pérez, David, et al.
Published: (2023)
by: López-Pérez, David, et al.
Published: (2023)
Low-Rank Adapters Meet Neural Architecture Search for LLM Compression
by: Muñoz, J. Pablo, et al.
Published: (2025)
by: Muñoz, J. Pablo, et al.
Published: (2025)
Robust Image Semantic Coding with Learnable CSI Fusion Masking over MIMO Fading Channels
by: Xie, Bingyan, et al.
Published: (2024)
by: Xie, Bingyan, et al.
Published: (2024)
Q-realign: Piggybacking Realignment on Quantization for Safe and Efficient LLM Deployment
by: Tan, Qitao, et al.
Published: (2026)
by: Tan, Qitao, et al.
Published: (2026)
Multi-Agent Deep Reinforcement Learning for Safe Autonomous Driving with RICS-Assisted MEC
by: Zhang, Xueyao, et al.
Published: (2025)
by: Zhang, Xueyao, et al.
Published: (2025)
Federated Attention: A Distributed Paradigm for Collaborative LLM Inference over Edge Networks
by: Deng, Xiumei, et al.
Published: (2025)
by: Deng, Xiumei, et al.
Published: (2025)
WinQ: Accelerating Quantization-Aware Training of Language Models Around Saddle Points
by: Li, Dongyue, et al.
Published: (2026)
by: Li, Dongyue, et al.
Published: (2026)
DyQ-VLA: Temporal-Dynamic-Aware Quantization for Embodied Vision-Language-Action Models
by: Zheng, Zihao, et al.
Published: (2026)
by: Zheng, Zihao, et al.
Published: (2026)
Can LLMs Revolutionize the Design of Explainable and Efficient TinyML Models?
by: Zeinaty, Christophe El, et al.
Published: (2025)
by: Zeinaty, Christophe El, et al.
Published: (2025)
LoQT: Low-Rank Adapters for Quantized Pretraining
by: Loeschcke, Sebastian, et al.
Published: (2024)
by: Loeschcke, Sebastian, et al.
Published: (2024)
QET: Enhancing Quantized LLM Parameters and KV cache Compression through Element Substitution and Residual Clustering
by: Wang, Yanshu, et al.
Published: (2024)
by: Wang, Yanshu, et al.
Published: (2024)
A2Q+: Improving Accumulator-Aware Weight Quantization
by: Colbert, Ian, et al.
Published: (2024)
by: Colbert, Ian, et al.
Published: (2024)
Similar Items
-
BAQ: Efficient Bit Allocation Quantization for Large Language Models
by: Zhang, Chao, et al.
Published: (2025) -
Strategic Federated Learning: Application to Smart Meter Data Clustering
by: Mohamad, Hassan, et al.
Published: (2024) -
Goal-Oriented State Information Compression for Linear Dynamical System Control
by: Wang, Li, et al.
Published: (2024) -
RF-GPT: Teaching AI to See the Wireless World
by: Zou, Hang, et al.
Published: (2026) -
Per-antenna power constraints: constructing Pareto-optimal precoders with cubic complexity under non-negligible noise conditions
by: Petrov, Sergey, et al.
Published: (2025)