Know What You do Not Know: Verbalized Uncertainty Estimation Robustness on Corrupted Images in Vision-Language Models
Fuente:
arXiv
Saved in:
| Main Authors: | Borszukovszki, Mirko, de Jong, Ivo Pascal, Valdenegro-Toro, Matias |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Overconfidence is Key: Verbalized Uncertainty Evaluation in Large Language and Vision-Language Models
by: Groot, Tobias, et al.
Published: (2024)
by: Groot, Tobias, et al.
Published: (2024)
Uncertainty Quantification for cross-subject Motor Imagery classification
by: Manivannan, Prithviraj, et al.
Published: (2024)
by: Manivannan, Prithviraj, et al.
Published: (2024)
Can Bayesian Neural Networks Explicitly Model Input Uncertainty?
by: Valdenegro-Toro, Matias, et al.
Published: (2025)
by: Valdenegro-Toro, Matias, et al.
Published: (2025)
Visually Dehallucinative Instruction Generation: Know What You Don't Know
by: Cha, Sungguk, et al.
Published: (2024)
by: Cha, Sungguk, et al.
Published: (2024)
Uncertainty Estimation for Super-Resolution using ESRGAN
by: Adapa, Maniraj Sai, et al.
Published: (2024)
by: Adapa, Maniraj Sai, et al.
Published: (2024)
All You Need to Know About Training Image Retrieval Models
by: Berton, Gabriele, et al.
Published: (2025)
by: Berton, Gabriele, et al.
Published: (2025)
Specifying What You Know or Not for Multi-Label Class-Incremental Learning
by: Zhang, Aoting, et al.
Published: (2025)
by: Zhang, Aoting, et al.
Published: (2025)
You Never Know: Quantization Induces Inconsistent Biases in Vision-Language Foundation Models
by: Slyman, Eric, et al.
Published: (2024)
by: Slyman, Eric, et al.
Published: (2024)
The VVAD-LRS3 Dataset for Visual Voice Activity Detection
by: Lubitz, Adrian, et al.
Published: (2021)
by: Lubitz, Adrian, et al.
Published: (2021)
Analysing the Robustness of Vision-Language-Models to Common Corruptions
by: Usama, Muhammad, et al.
Published: (2025)
by: Usama, Muhammad, et al.
Published: (2025)
World Models That Know When They Don't Know - Controllable Video Generation with Calibrated Uncertainty
by: Mei, Zhiting, et al.
Published: (2025)
by: Mei, Zhiting, et al.
Published: (2025)
Know3D: Prompting 3D Generation with Knowledge from Vision-Language Models
by: Chen, Wenyue, et al.
Published: (2026)
by: Chen, Wenyue, et al.
Published: (2026)
Generative Models: What Do They Know? Do They Know Things? Let's Find Out!
by: Du, Xiaodan, et al.
Published: (2023)
by: Du, Xiaodan, et al.
Published: (2023)
When Seeing Overrides Knowing: Disentangling Knowledge Conflicts in Vision-Language Models
by: Ortu, Francesco, et al.
Published: (2025)
by: Ortu, Francesco, et al.
Published: (2025)
Robust Deepfake Detection for Electronic Know Your Customer Systems Using Registered Images
by: Amada, Takuma, et al.
Published: (2025)
by: Amada, Takuma, et al.
Published: (2025)
Air-Know: Arbiter-Calibrated Knowledge-Internalizing Robust Network for Composed Image Retrieval
by: Fu, Zhiheng, et al.
Published: (2026)
by: Fu, Zhiheng, et al.
Published: (2026)
Magnet: We Never Know How Text-to-Image Diffusion Models Work, Until We Learn How Vision-Language Models Function
by: Zhuang, Chenyi, et al.
Published: (2024)
by: Zhuang, Chenyi, et al.
Published: (2024)
Unified Uncertainties: Combining Input, Data and Model Uncertainty into a Single Formulation
by: Valdenegro-Toro, Matias, et al.
Published: (2024)
by: Valdenegro-Toro, Matias, et al.
Published: (2024)
Unconsciously Forget: Mitigating Memorization; Without Knowing What is being Memorized
by: Jin, Er, et al.
Published: (2025)
by: Jin, Er, et al.
Published: (2025)
Guard Me If You Know Me: Protecting Specific Face-Identity from Deepfakes
by: Lin, Kaiqing, et al.
Published: (2025)
by: Lin, Kaiqing, et al.
Published: (2025)
Know Your Neighbors: Improving Single-View Reconstruction via Spatial Vision-Language Reasoning
by: Li, Rui, et al.
Published: (2024)
by: Li, Rui, et al.
Published: (2024)
Improving Robustness of Vision-Language-Action Models by Restoring Corrupted Visual Inputs
by: Orjuela, Daniel Yezid Guarnizo, et al.
Published: (2026)
by: Orjuela, Daniel Yezid Guarnizo, et al.
Published: (2026)
The First to Know: How Token Distributions Reveal Hidden Knowledge in Large Vision-Language Models?
by: Zhao, Qinyu, et al.
Published: (2024)
by: Zhao, Qinyu, et al.
Published: (2024)
Know-Show: Benchmarking Video-Language Models on Spatio-Temporal Grounded Reasoning
by: Sugandhika, Chinthani, et al.
Published: (2025)
by: Sugandhika, Chinthani, et al.
Published: (2025)
SeTformer is What You Need for Vision and Language
by: Shamsolmoali, Pourya, et al.
Published: (2024)
by: Shamsolmoali, Pourya, et al.
Published: (2024)
A Survey on the Robustness of Computer Vision Models against Common Corruptions
by: Wang, Shunxin, et al.
Published: (2023)
by: Wang, Shunxin, et al.
Published: (2023)
Explorations of the Softmax Space: Knowing When the Neural Network Doesn't Know
by: Sikar, Daniel, et al.
Published: (2025)
by: Sikar, Daniel, et al.
Published: (2025)
Get What You Want, Not What You Don't: Image Content Suppression for Text-to-Image Diffusion Models
by: Li, Senmao, et al.
Published: (2024)
by: Li, Senmao, et al.
Published: (2024)
Adaptive Prompt Tuning: Vision Guided Prompt Tuning with Cross-Attention for Fine-Grained Few-Shot Learning
by: Brouwer, Eric, et al.
Published: (2024)
by: Brouwer, Eric, et al.
Published: (2024)
Do You Know Where Your Camera Is? View-Invariant Policy Learning with Camera Conditioning
by: Jiang, Tianchong, et al.
Published: (2025)
by: Jiang, Tianchong, et al.
Published: (2025)
VL-Uncertainty: Detecting Hallucination in Large Vision-Language Model via Uncertainty Estimation
by: Zhang, Ruiyang, et al.
Published: (2024)
by: Zhang, Ruiyang, et al.
Published: (2024)
Diffusion Knows Transparency: Repurposing Video Diffusion for Transparent Object Depth and Normal Estimation
by: Xu, Shaocong, et al.
Published: (2025)
by: Xu, Shaocong, et al.
Published: (2025)
Let 2D Diffusion Model Know 3D-Consistency for Robust Text-to-3D Generation
by: Seo, Junyoung, et al.
Published: (2023)
by: Seo, Junyoung, et al.
Published: (2023)
What You Have is What You Track: Adaptive and Robust Multimodal Tracking
by: Tan, Yuedong, et al.
Published: (2025)
by: Tan, Yuedong, et al.
Published: (2025)
The Marine Debris Forward-Looking Sonar Datasets
by: Valdenegro-Toro, Matias, et al.
Published: (2025)
by: Valdenegro-Toro, Matias, et al.
Published: (2025)
Is What You Ask For What You Get? Investigating Concept Associations in Text-to-Image Models
by: Magid, Salma Abdel, et al.
Published: (2024)
by: Magid, Salma Abdel, et al.
Published: (2024)
Intra-Class Probabilistic Embeddings for Uncertainty Estimation in Vision-Language Models
by: Lin, Zhenxiang, et al.
Published: (2025)
by: Lin, Zhenxiang, et al.
Published: (2025)
Image Corruption-Inspired Membership Inference Attacks against Large Vision-Language Models
by: Wu, Zongyu, et al.
Published: (2025)
by: Wu, Zongyu, et al.
Published: (2025)
Text Embedding Knows How to Quantize Text-Guided Diffusion Models
by: Lee, Hongjae, et al.
Published: (2025)
by: Lee, Hongjae, et al.
Published: (2025)
RadProPoser: Probabilistic Radar Tensor Human Pose Estimation That Knows Its Limits
by: Mueller, Jonas Leo, et al.
Published: (2025)
by: Mueller, Jonas Leo, et al.
Published: (2025)
Similar Items
-
Overconfidence is Key: Verbalized Uncertainty Evaluation in Large Language and Vision-Language Models
by: Groot, Tobias, et al.
Published: (2024) -
Uncertainty Quantification for cross-subject Motor Imagery classification
by: Manivannan, Prithviraj, et al.
Published: (2024) -
Can Bayesian Neural Networks Explicitly Model Input Uncertainty?
by: Valdenegro-Toro, Matias, et al.
Published: (2025) -
Visually Dehallucinative Instruction Generation: Know What You Don't Know
by: Cha, Sungguk, et al.
Published: (2024) -
Uncertainty Estimation for Super-Resolution using ESRGAN
by: Adapa, Maniraj Sai, et al.
Published: (2024)