LQA: A Lightweight Quantized-Adaptive Framework for Vision-Language Models on the Edge
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wang, Xin, Jia, Hong, Zhou, Hualin, Wang, Sheng Guang, Zhang, Yu, Dang, Ting, Gu, Tao |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
EdgeRazor: A Lightweight Framework for Large Language Models via Mixed-Precision Quantization-Aware Distillation
von: Zhang, Shu-Hao, et al.
Veröffentlicht: (2026)
von: Zhang, Shu-Hao, et al.
Veröffentlicht: (2026)
Efficient and Personalized Mobile Health Event Prediction via Small Language Models
von: Wang, Xin, et al.
Veröffentlicht: (2024)
von: Wang, Xin, et al.
Veröffentlicht: (2024)
Scaling Auditory Cognition via Test-Time Compute in Audio Language Models
von: Dang, Ting, et al.
Veröffentlicht: (2025)
von: Dang, Ting, et al.
Veröffentlicht: (2025)
Aligned Vector Quantization for Edge-Cloud Collabrative Vision-Language Models
von: Liu, Xiao, et al.
Veröffentlicht: (2024)
von: Liu, Xiao, et al.
Veröffentlicht: (2024)
HealthSLM-Bench: Benchmarking Small Language Models for Mobile and Wearable Healthcare Monitoring
von: Wang, Xin, et al.
Veröffentlicht: (2025)
von: Wang, Xin, et al.
Veröffentlicht: (2025)
AER-LLM: Ambiguity-aware Emotion Recognition Leveraging Large Language Models
von: Hong, Xin, et al.
Veröffentlicht: (2024)
von: Hong, Xin, et al.
Veröffentlicht: (2024)
MBQ: Modality-Balanced Quantization for Large Vision-Language Models
von: Li, Shiyao, et al.
Veröffentlicht: (2024)
von: Li, Shiyao, et al.
Veröffentlicht: (2024)
ActionFlow: A Pipelined Action Acceleration for Vision Language Models on Edge
von: Dai, Yuntao, et al.
Veröffentlicht: (2025)
von: Dai, Yuntao, et al.
Veröffentlicht: (2025)
VEQ: Modality-Adaptive Quantization for MoE Vision-Language Models
von: Qin, Guangshuo, et al.
Veröffentlicht: (2026)
von: Qin, Guangshuo, et al.
Veröffentlicht: (2026)
FlexQuant: Elastic Quantization Framework for Locally Hosted LLM on Edge Devices
von: Chai, Yuji, et al.
Veröffentlicht: (2025)
von: Chai, Yuji, et al.
Veröffentlicht: (2025)
Quant Experts: Token-aware Adaptive Error Reconstruction with Mixture of Experts for Large Vision-Language Models Quantization
von: Jia, Chenwei, et al.
Veröffentlicht: (2026)
von: Jia, Chenwei, et al.
Veröffentlicht: (2026)
AppVLM: A Lightweight Vision Language Model for Online App Control
von: Papoudakis, Georgios, et al.
Veröffentlicht: (2025)
von: Papoudakis, Georgios, et al.
Veröffentlicht: (2025)
P$^2$-ViT: Power-of-Two Post-Training Quantization and Acceleration for Fully Quantized Vision Transformer
von: Shi, Huihong, et al.
Veröffentlicht: (2024)
von: Shi, Huihong, et al.
Veröffentlicht: (2024)
Disentangling Reasoning in Large Audio-Language Models for Ambiguous Emotion Prediction
von: Yu, Xiaofeng, et al.
Veröffentlicht: (2026)
von: Yu, Xiaofeng, et al.
Veröffentlicht: (2026)
Model Specific Task Similarity for Vision Language Model Selection via Layer Conductance
von: Yang, Wei, et al.
Veröffentlicht: (2026)
von: Yang, Wei, et al.
Veröffentlicht: (2026)
Vision-Language Model Purified Semi-Supervised Semantic Segmentation for Remote Sensing Images
von: Wang, Shanwen, et al.
Veröffentlicht: (2026)
von: Wang, Shanwen, et al.
Veröffentlicht: (2026)
LightPROF: A Lightweight Reasoning Framework for Large Language Model on Knowledge Graph
von: Ao, Tu, et al.
Veröffentlicht: (2025)
von: Ao, Tu, et al.
Veröffentlicht: (2025)
UniQL: Unified Quantization and Low-rank Compression for Adaptive Edge LLMs
von: Chiang, Hung-Yueh, et al.
Veröffentlicht: (2025)
von: Chiang, Hung-Yueh, et al.
Veröffentlicht: (2025)
CoVSpec: Efficient Device-Edge Co-Inference for Vision-Language Models via Speculative Decoding
von: Jia, Yuanyuan, et al.
Veröffentlicht: (2026)
von: Jia, Yuanyuan, et al.
Veröffentlicht: (2026)
Quantized Large Language Models in Biomedical Natural Language Processing: Evaluation and Recommendation
von: Zhan, Zaifu, et al.
Veröffentlicht: (2025)
von: Zhan, Zaifu, et al.
Veröffentlicht: (2025)
Seeing Through Experts Eyes A Foundational Vision Language Model Trained on Radiologists Gaze and Reasoning
von: Lee, Kinhei, et al.
Veröffentlicht: (2026)
von: Lee, Kinhei, et al.
Veröffentlicht: (2026)
Adaptive-Solver Framework for Dynamic Strategy Selection in Large Language Model Reasoning
von: Zhou, Jianpeng, et al.
Veröffentlicht: (2023)
von: Zhou, Jianpeng, et al.
Veröffentlicht: (2023)
Privacy-Preserving SAM Quantization for Efficient Edge Intelligence in Healthcare
von: Li, Zhikai, et al.
Veröffentlicht: (2024)
von: Li, Zhikai, et al.
Veröffentlicht: (2024)
ACPO: Adaptive Curriculum Policy Optimization for Aligning Vision-Language Models in Complex Reasoning
von: Wang, Yunhao, et al.
Veröffentlicht: (2025)
von: Wang, Yunhao, et al.
Veröffentlicht: (2025)
Menta: A Small Language Model for On-Device Mental Health Prediction
von: Zhang, Tianyi, et al.
Veröffentlicht: (2025)
von: Zhang, Tianyi, et al.
Veröffentlicht: (2025)
Evaluating Quantized Large Language Models
von: Li, Shiyao, et al.
Veröffentlicht: (2024)
von: Li, Shiyao, et al.
Veröffentlicht: (2024)
APreQEL: Adaptive Mixed Precision Quantization For Edge LLMs
von: Bouzouad, Meriem, et al.
Veröffentlicht: (2026)
von: Bouzouad, Meriem, et al.
Veröffentlicht: (2026)
VPTQ: Extreme Low-bit Vector Post-Training Quantization for Large Language Models
von: Liu, Yifei, et al.
Veröffentlicht: (2024)
von: Liu, Yifei, et al.
Veröffentlicht: (2024)
QAPruner: Quantization-Aware Vision Token Pruning for Multimodal Large Language Models
von: Wang, Xinhao, et al.
Veröffentlicht: (2026)
von: Wang, Xinhao, et al.
Veröffentlicht: (2026)
Fine-Grained Post-Training Quantization for Large Vision Language Models with Quantization-Aware Integrated Gradients
von: Xiang, Ziwei, et al.
Veröffentlicht: (2026)
von: Xiang, Ziwei, et al.
Veröffentlicht: (2026)
Edge Intelligence Optimization for Large Language Model Inference with Batching and Quantization
von: Zhang, Xinyuan, et al.
Veröffentlicht: (2024)
von: Zhang, Xinyuan, et al.
Veröffentlicht: (2024)
GeoWorld-VLM: Geometry from World Models for Vision-Language Models
von: Gu, Renjie, et al.
Veröffentlicht: (2026)
von: Gu, Renjie, et al.
Veröffentlicht: (2026)
Adaptive AI Agent Placement and Migration in Edge Intelligence Systems
von: Wang, Xingdan, et al.
Veröffentlicht: (2025)
von: Wang, Xingdan, et al.
Veröffentlicht: (2025)
Navigation-GPT: A Robust and Adaptive Framework Utilizing Large Language Models for Navigation Applications
von: Ma, Feng, et al.
Veröffentlicht: (2025)
von: Ma, Feng, et al.
Veröffentlicht: (2025)
Mitigating Entangled Steering in Large Vision-Language Models for Hallucination Reduction
von: Zhang, Yuanhong, et al.
Veröffentlicht: (2026)
von: Zhang, Yuanhong, et al.
Veröffentlicht: (2026)
Adversarial Prompt Tuning for Vision-Language Models
von: Zhang, Jiaming, et al.
Veröffentlicht: (2023)
von: Zhang, Jiaming, et al.
Veröffentlicht: (2023)
Lightweight LLM Agent Memory with Small Language Models
von: Zhang, Jiaquan, et al.
Veröffentlicht: (2026)
von: Zhang, Jiaquan, et al.
Veröffentlicht: (2026)
Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models
von: Lian, Chenyu, et al.
Veröffentlicht: (2025)
von: Lian, Chenyu, et al.
Veröffentlicht: (2025)
PPU-Bench:Real World Benchmark for Personalized Partial Unlearning in Vision Language Models
von: Guang, Jiahui, et al.
Veröffentlicht: (2026)
von: Guang, Jiahui, et al.
Veröffentlicht: (2026)
Q-resafe: Assessing Safety Risks and Quantization-aware Safety Patching for Quantized Large Language Models
von: Chen, Kejia, et al.
Veröffentlicht: (2025)
von: Chen, Kejia, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
EdgeRazor: A Lightweight Framework for Large Language Models via Mixed-Precision Quantization-Aware Distillation
von: Zhang, Shu-Hao, et al.
Veröffentlicht: (2026) -
Efficient and Personalized Mobile Health Event Prediction via Small Language Models
von: Wang, Xin, et al.
Veröffentlicht: (2024) -
Scaling Auditory Cognition via Test-Time Compute in Audio Language Models
von: Dang, Ting, et al.
Veröffentlicht: (2025) -
Aligned Vector Quantization for Edge-Cloud Collabrative Vision-Language Models
von: Liu, Xiao, et al.
Veröffentlicht: (2024) -
HealthSLM-Bench: Benchmarking Small Language Models for Mobile and Wearable Healthcare Monitoring
von: Wang, Xin, et al.
Veröffentlicht: (2025)