Object-Level Verbalized Confidence Calibration in Vision-Language Models via Semantic Perturbation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhao, Yunpu, Zhang, Rui, Xiao, Junbin, Hou, Ruibo, Guo, Jiaming, Zhang, Zihao, Hao, Yifan, Chen, Yunji |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Sycophancy in Vision-Language Models: A Systematic Analysis and an Inference-Time Mitigation Framework
von: Zhao, Yunpu, et al.
Veröffentlicht: (2024)
von: Zhao, Yunpu, et al.
Veröffentlicht: (2024)
Decision-Driven Semantic Object Exploration for Legged Robots via Confidence-Calibrated Perception and Topological Subgoal Selection
von: Zhao, Guoyang, et al.
Veröffentlicht: (2025)
von: Zhao, Guoyang, et al.
Veröffentlicht: (2025)
EmoCaliber: Advancing Reliable Visual Emotion Comprehension via Confidence Verbalization and Calibration
von: Wu, Daiqing, et al.
Veröffentlicht: (2025)
von: Wu, Daiqing, et al.
Veröffentlicht: (2025)
VL-Calibration: Decoupled Confidence Calibration for Large Vision-Language Models Reasoning
von: Xiao, Wenyi, et al.
Veröffentlicht: (2026)
von: Xiao, Wenyi, et al.
Veröffentlicht: (2026)
DA-Ada: Learning Domain-Aware Adapter for Domain Adaptive Object Detection
von: Li, Haochen, et al.
Veröffentlicht: (2024)
von: Li, Haochen, et al.
Veröffentlicht: (2024)
Boosting 3D Object Detection with Semantic-Aware Multi-Branch Framework
von: Jing, Hao, et al.
Veröffentlicht: (2024)
von: Jing, Hao, et al.
Veröffentlicht: (2024)
World-Consistent Data Generation for Vision-and-Language Navigation
von: Zhong, Yu, et al.
Veröffentlicht: (2024)
von: Zhong, Yu, et al.
Veröffentlicht: (2024)
Seeing is Believing, but How Much? A Comprehensive Analysis of Verbalized Calibration in Vision-Language Models
von: Xuan, Weihao, et al.
Veröffentlicht: (2025)
von: Xuan, Weihao, et al.
Veröffentlicht: (2025)
Instinct vs. Reflection: Unifying Token and Verbalized Confidence in Multimodal Large Models
von: Dang, Yunkai, et al.
Veröffentlicht: (2026)
von: Dang, Yunkai, et al.
Veröffentlicht: (2026)
Think-as-You-See: Streaming Chain-of-Thought Reasoning for Large Vision-Language Models
von: Zhang, Jialiang, et al.
Veröffentlicht: (2026)
von: Zhang, Jialiang, et al.
Veröffentlicht: (2026)
Assessing and Understanding Creativity in Large Language Models
von: Zhao, Yunpu, et al.
Veröffentlicht: (2024)
von: Zhao, Yunpu, et al.
Veröffentlicht: (2024)
Dual-View Pyramid Pooling in Deep Neural Networks for Improved Medical Image Classification and Confidence Calibration
von: Zhang, Xiaoqing, et al.
Veröffentlicht: (2024)
von: Zhang, Xiaoqing, et al.
Veröffentlicht: (2024)
On Robustness of Vision-Language-Action Model against Multi-Modal Perturbations
von: Guo, Jianing, et al.
Veröffentlicht: (2025)
von: Guo, Jianing, et al.
Veröffentlicht: (2025)
Mitigating Object Hallucinations in Large Vision-Language Models via Attention Calibration
von: Zhu, Younan, et al.
Veröffentlicht: (2025)
von: Zhu, Younan, et al.
Veröffentlicht: (2025)
Consistency Calibration: Improving Uncertainty Calibration via Consistency among Perturbed Neighbors
von: Tao, Linwei, et al.
Veröffentlicht: (2024)
von: Tao, Linwei, et al.
Veröffentlicht: (2024)
On the Consistency of Video Large Language Models in Temporal Comprehension
von: Jung, Minjoon, et al.
Veröffentlicht: (2024)
von: Jung, Minjoon, et al.
Veröffentlicht: (2024)
Cross-Level Sensor Fusion with Object Lists via Transformer for 3D Object Detection
von: Liu, Xiangzhong, et al.
Veröffentlicht: (2025)
von: Liu, Xiangzhong, et al.
Veröffentlicht: (2025)
Rethinking Practical and Efficient Quantization Calibration for Vision-Language Models
von: Shang, Zhenhao, et al.
Veröffentlicht: (2026)
von: Shang, Zhenhao, et al.
Veröffentlicht: (2026)
ICT: Image-Object Cross-Level Trusted Intervention for Mitigating Object Hallucination in Large Vision-Language Models
von: Chen, Junzhe, et al.
Veröffentlicht: (2024)
von: Chen, Junzhe, et al.
Veröffentlicht: (2024)
Vision-Motion-Reference Alignment for Referring Multi-Object Tracking via Multi-Modal Large Language Models
von: Lv, Weiyi, et al.
Veröffentlicht: (2025)
von: Lv, Weiyi, et al.
Veröffentlicht: (2025)
Confidence Self-Calibration for Multi-Label Class-Incremental Learning
von: Du, Kaile, et al.
Veröffentlicht: (2024)
von: Du, Kaile, et al.
Veröffentlicht: (2024)
Semantic Router: On the Feasibility of Hijacking MLLMs via a Single Adversarial Perturbation
von: Li, Changyue, et al.
Veröffentlicht: (2025)
von: Li, Changyue, et al.
Veröffentlicht: (2025)
A Prediction-as-Perception Framework for 3D Object Detection
von: Zhang, Song, et al.
Veröffentlicht: (2026)
von: Zhang, Song, et al.
Veröffentlicht: (2026)
CLIPose: Category-Level Object Pose Estimation with Pre-trained Vision-Language Knowledge
von: Lin, Xiao, et al.
Veröffentlicht: (2024)
von: Lin, Xiao, et al.
Veröffentlicht: (2024)
Reflectance Prediction-based Knowledge Distillation for Robust 3D Object Detection in Compressed Point Clouds
von: Jing, Hao, et al.
Veröffentlicht: (2025)
von: Jing, Hao, et al.
Veröffentlicht: (2025)
Calibrating Verbalized Confidence with Self-Generated Distractors
von: Wang, Victor, et al.
Veröffentlicht: (2025)
von: Wang, Victor, et al.
Veröffentlicht: (2025)
HiDrop: Hierarchical Vision Token Reduction in MLLMs via Late Injection, Concave Pyramid Pruning, and Early Exit
von: Wu, Hao, et al.
Veröffentlicht: (2026)
von: Wu, Hao, et al.
Veröffentlicht: (2026)
Disrupting Vision-Language Model-Driven Navigation Services via Adversarial Object Fusion
von: Xie, Chunlong, et al.
Veröffentlicht: (2025)
von: Xie, Chunlong, et al.
Veröffentlicht: (2025)
DA-Cal: Towards Cross-Domain Calibration in Semantic Segmentation
von: Li, Wangkai, et al.
Veröffentlicht: (2026)
von: Li, Wangkai, et al.
Veröffentlicht: (2026)
Semantics-aware Motion Retargeting with Vision-Language Models
von: Zhang, Haodong, et al.
Veröffentlicht: (2023)
von: Zhang, Haodong, et al.
Veröffentlicht: (2023)
Semantic Is Enough: Only Semantic Information For NeRF Reconstruction
von: Wang, Ruibo, et al.
Veröffentlicht: (2024)
von: Wang, Ruibo, et al.
Veröffentlicht: (2024)
VORD: Visual Ordinal Calibration for Mitigating Object Hallucinations in Large Vision-Language Models
von: Neo, Dexter, et al.
Veröffentlicht: (2024)
von: Neo, Dexter, et al.
Veröffentlicht: (2024)
Towards Open Environments and Instructions: General Vision-Language Navigation via Fast-Slow Interactive Reasoning
von: Li, Yang, et al.
Veröffentlicht: (2026)
von: Li, Yang, et al.
Veröffentlicht: (2026)
One Perturbation is Enough: On Generating Universal Adversarial Perturbations against Vision-Language Pre-training Models
von: Fang, Hao, et al.
Veröffentlicht: (2024)
von: Fang, Hao, et al.
Veröffentlicht: (2024)
VERA: Explainable Video Anomaly Detection via Verbalized Learning of Vision-Language Models
von: Ye, Muchao, et al.
Veröffentlicht: (2024)
von: Ye, Muchao, et al.
Veröffentlicht: (2024)
Probabilistic Prototype Calibration of Vision-Language Models for Generalized Few-shot Semantic Segmentation
von: Liu, Jie, et al.
Veröffentlicht: (2025)
von: Liu, Jie, et al.
Veröffentlicht: (2025)
UMFC: Unsupervised Multi-Domain Feature Calibration for Vision-Language Models
von: Liang, Jiachen, et al.
Veröffentlicht: (2024)
von: Liang, Jiachen, et al.
Veröffentlicht: (2024)
Training-Free Semantic Multi-Object Tracking with Vision-Language Models
von: Bonat, Laurence, et al.
Veröffentlicht: (2026)
von: Bonat, Laurence, et al.
Veröffentlicht: (2026)
Defending Deepfake via Texture Feature Perturbation
von: Zhang, Xiao, et al.
Veröffentlicht: (2025)
von: Zhang, Xiao, et al.
Veröffentlicht: (2025)
Calibrated and Efficient Sampling-Free Confidence Estimation for LiDAR Scene Semantic Segmentation
von: Miandashti, Hanieh Shojaei, et al.
Veröffentlicht: (2024)
von: Miandashti, Hanieh Shojaei, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Sycophancy in Vision-Language Models: A Systematic Analysis and an Inference-Time Mitigation Framework
von: Zhao, Yunpu, et al.
Veröffentlicht: (2024) -
Decision-Driven Semantic Object Exploration for Legged Robots via Confidence-Calibrated Perception and Topological Subgoal Selection
von: Zhao, Guoyang, et al.
Veröffentlicht: (2025) -
EmoCaliber: Advancing Reliable Visual Emotion Comprehension via Confidence Verbalization and Calibration
von: Wu, Daiqing, et al.
Veröffentlicht: (2025) -
VL-Calibration: Decoupled Confidence Calibration for Large Vision-Language Models Reasoning
von: Xiao, Wenyi, et al.
Veröffentlicht: (2026) -
DA-Ada: Learning Domain-Aware Adapter for Domain Adaptive Object Detection
von: Li, Haochen, et al.
Veröffentlicht: (2024)