Less Precise Can Be More Reliable: A Systematic Evaluation of Quantization's Impact on VLMs Beyond Accuracy
Fuente:
arXiv
Saved in:
| Main Authors: | Bouguerra, Aymen, Montoya, Daniel, Gomez-Villa, Alexandra, Mraidha, Chokri, Arnez, Fabio |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
FindMeIfYouCan: Bringing Open Set metrics to $\textit{near} $, $ \textit{far} $ and $\textit{farther}$ Out-of-Distribution Object Detection
by: Montoya, Daniel, et al.
Published: (2025)
by: Montoya, Daniel, et al.
Published: (2025)
The Map of Misbelief: Tracing Intrinsic and Extrinsic Hallucinations Through Attention Patterns
by: Hajji, Elyes, et al.
Published: (2025)
by: Hajji, Elyes, et al.
Published: (2025)
Oracle-Guided Soft Shielding for Safe Move Prediction in Chess
by: Rajendran, Prajit T, et al.
Published: (2026)
by: Rajendran, Prajit T, et al.
Published: (2026)
Continual Learning for VLMs: A Survey and Taxonomy Beyond Forgetting
by: Liu, Yuyang, et al.
Published: (2025)
by: Liu, Yuyang, et al.
Published: (2025)
LIME: Less Is More for MLLM Evaluation
by: Zhu, King, et al.
Published: (2024)
by: Zhu, King, et al.
Published: (2024)
CIVET: Systematic Evaluation of Understanding in VLMs
by: Rizzoli, Massimo, et al.
Published: (2025)
by: Rizzoli, Massimo, et al.
Published: (2025)
More Thought, Less Accuracy? On the Dual Nature of Reasoning in Vision-Language Models
by: Tian, Xinyu, et al.
Published: (2025)
by: Tian, Xinyu, et al.
Published: (2025)
Send Less, Perceive More: Masked Quantized Point Cloud Communication for Loss-Tolerant Collaborative Perception
by: Xu, Sheng, et al.
Published: (2026)
by: Xu, Sheng, et al.
Published: (2026)
Evaluating the Impact of Post-Training Quantization on Reliable VQA with Multimodal LLMs
by: Kurz, Paul Jonas, et al.
Published: (2026)
by: Kurz, Paul Jonas, et al.
Published: (2026)
Right this way: Can VLMs Guide Us to See More to Answer Questions?
by: Liu, Li, et al.
Published: (2024)
by: Liu, Li, et al.
Published: (2024)
Better Reasoning with Less Data: Enhancing VLMs Through Unified Modality Scoring
by: Xu, Mingjie, et al.
Published: (2025)
by: Xu, Mingjie, et al.
Published: (2025)
WindowQuant: Mixed-Precision KV Cache Quantization based on Window-Level Similarity for VLMs Inference Optimization
by: Tao, Wei, et al.
Published: (2026)
by: Tao, Wei, et al.
Published: (2026)
Less is More: Discovering Concise Network Explanations
by: Kondapaneni, Neehar, et al.
Published: (2024)
by: Kondapaneni, Neehar, et al.
Published: (2024)
Less-to-More Generalization: Unlocking More Controllability by In-Context Generation
by: Wu, Shaojin, et al.
Published: (2025)
by: Wu, Shaojin, et al.
Published: (2025)
Beyond the Linear Separability Ceiling: Aligning Representations in VLMs
by: Vompa, Enrico, et al.
Published: (2025)
by: Vompa, Enrico, et al.
Published: (2025)
Rethinking Semi-supervised Segmentation Beyond Accuracy: Reliability and Robustness
by: Landgraf, Steven, et al.
Published: (2025)
by: Landgraf, Steven, et al.
Published: (2025)
Less Gaussians, Texture More: 4K Feed-Forward Textured Splatting
by: Lao, Yixing, et al.
Published: (2026)
by: Lao, Yixing, et al.
Published: (2026)
Beyond Accuracy: Evaluating Grounded Visual Evidence in Thinking with Images
by: Li, Xuchen, et al.
Published: (2026)
by: Li, Xuchen, et al.
Published: (2026)
Beyond Accuracy: Evaluating Visual Grounding In Multimodal Medical Reasoning
by: Zafar, Anas, et al.
Published: (2026)
by: Zafar, Anas, et al.
Published: (2026)
What Makes VLMs Robust? Towards Reconciling Robustness and Accuracy in Vision-Language Models
by: Nie, Sen, et al.
Published: (2026)
by: Nie, Sen, et al.
Published: (2026)
ProSR: Process-Shaped Spatial Reasoning for Reliable Chain-of-Thought in VLMs
by: Li, Jiangyang, et al.
Published: (2026)
by: Li, Jiangyang, et al.
Published: (2026)
Probing the Reliability of Driving VLMs: From Inconsistent Responses to Grounded Temporal Reasoning
by: Chang, Chun-Peng, et al.
Published: (2026)
by: Chang, Chun-Peng, et al.
Published: (2026)
Less is More: The Influence of Pruning on the Explainability of CNNs
by: Merkle, Florian, et al.
Published: (2023)
by: Merkle, Florian, et al.
Published: (2023)
Leaner Transformers: More Heads, Less Depth
by: Saratchandran, Hemanth, et al.
Published: (2025)
by: Saratchandran, Hemanth, et al.
Published: (2025)
Less is More: Token Context-aware Learning for Object Tracking
by: Xu, Chenlong, et al.
Published: (2025)
by: Xu, Chenlong, et al.
Published: (2025)
Less is More: Improving Motion Diffusion Models with Sparse Keyframes
by: Bae, Jinseok, et al.
Published: (2025)
by: Bae, Jinseok, et al.
Published: (2025)
Less Is More: An Explainable AI Framework for Lightweight Malaria Classification
by: Kafi, Md Abdullah Al, et al.
Published: (2025)
by: Kafi, Md Abdullah Al, et al.
Published: (2025)
Think Twice to See More: Iterative Visual Reasoning in Medical VLMs
by: Chen, Kaitao, et al.
Published: (2025)
by: Chen, Kaitao, et al.
Published: (2025)
NumColor: Precise Numeric Color Control in Text-to-Image Generation
by: Butt, Muhammad Atif, et al.
Published: (2026)
by: Butt, Muhammad Atif, et al.
Published: (2026)
Seeing More with Less: Video Capsule Endoscopy with Multi-Task Learning
by: Werner, Julia, et al.
Published: (2025)
by: Werner, Julia, et al.
Published: (2025)
See More, Change Less: Anatomy-Aware Diffusion for Contrast Enhancement
by: Liu, Junqi, et al.
Published: (2025)
by: Liu, Junqi, et al.
Published: (2025)
Dinomaly: The Less Is More Philosophy in Multi-Class Unsupervised Anomaly Detection
by: Guo, Jia, et al.
Published: (2024)
by: Guo, Jia, et al.
Published: (2024)
Beyond GSD-as-Token: Continuous Scale Conditioning for Remote Sensing VLMs
by: Zhang, Song, et al.
Published: (2026)
by: Zhang, Song, et al.
Published: (2026)
VLMs have Tunnel Vision: Evaluating Nonlocal Visual Reasoning in Leading VLMs
by: Berman, Shmuel, et al.
Published: (2025)
by: Berman, Shmuel, et al.
Published: (2025)
Probabilistic Precision and Recall Towards Reliable Evaluation of Generative Models
by: Park, Dogyun, et al.
Published: (2023)
by: Park, Dogyun, et al.
Published: (2023)
VisualActBench: Can VLMs See and Act like a Human?
by: Zhang, Daoan, et al.
Published: (2025)
by: Zhang, Daoan, et al.
Published: (2025)
Can VLMs be used on videos for action recognition? LLMs are Visual Reasoning Coordinators
by: Lunia, Harsh
Published: (2024)
by: Lunia, Harsh
Published: (2024)
Less is More: Efficient Point Cloud Reconstruction via Multi-Head Decoders
by: Alonso, Pedro, et al.
Published: (2025)
by: Alonso, Pedro, et al.
Published: (2025)
See More, Store Less: Memory-Efficient Resolution for Video Moment Retrieval
by: Jeon, Mingyu, et al.
Published: (2026)
by: Jeon, Mingyu, et al.
Published: (2026)
MeMix: Writing Less, Remembering More for Streaming 3D Reconstruction
by: Dong, Jiacheng, et al.
Published: (2026)
by: Dong, Jiacheng, et al.
Published: (2026)
Similar Items
-
FindMeIfYouCan: Bringing Open Set metrics to $\textit{near} $, $ \textit{far} $ and $\textit{farther}$ Out-of-Distribution Object Detection
by: Montoya, Daniel, et al.
Published: (2025) -
The Map of Misbelief: Tracing Intrinsic and Extrinsic Hallucinations Through Attention Patterns
by: Hajji, Elyes, et al.
Published: (2025) -
Oracle-Guided Soft Shielding for Safe Move Prediction in Chess
by: Rajendran, Prajit T, et al.
Published: (2026) -
Continual Learning for VLMs: A Survey and Taxonomy Beyond Forgetting
by: Liu, Yuyang, et al.
Published: (2025) -
LIME: Less Is More for MLLM Evaluation
by: Zhu, King, et al.
Published: (2024)