Calibrating Uncertainty Quantification of Multi-Modal LLMs using Grounding
Fuente:
arXiv
Saved in:
| Main Authors: | Padhi, Trilok, Kaur, Ramneet, Cobb, Adam D., Acharya, Manoj, Roy, Anirban, Samplawski, Colin, Matejek, Brian, Berenbeim, Alexander M., Bastian, Nathaniel D., Jha, Susmit |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
From Actions to Understanding: Conformal Interpretability of Temporal Concepts in LLM Agents
by: Padhi, Trilok, et al.
Published: (2026)
by: Padhi, Trilok, et al.
Published: (2026)
Scalable Bayesian Low-Rank Adaptation of Large Language Models via Stochastic Variational Subspace Inference
by: Samplawski, Colin, et al.
Published: (2025)
by: Samplawski, Colin, et al.
Published: (2025)
Addressing Uncertainty in LLMs to Enhance Reliability in Generative AI
by: Kaur, Ramneet, et al.
Published: (2024)
by: Kaur, Ramneet, et al.
Published: (2024)
Privacy Preserving In-Context-Learning Framework for Large Language Models
by: Bhusal, Bishnu, et al.
Published: (2025)
by: Bhusal, Bishnu, et al.
Published: (2025)
Breaking Bad: Interpretability-Based Safety Audits of State-of-the-Art LLMs
by: Agarwal, Krishiv, et al.
Published: (2026)
by: Agarwal, Krishiv, et al.
Published: (2026)
Do Diffusion Models Dream of Electric Planes? Discrete and Continuous Simulation-Based Inference for Aircraft Design
by: Ghiglino, Aurelien, et al.
Published: (2026)
by: Ghiglino, Aurelien, et al.
Published: (2026)
Polysemantic Dropout: Conformal OOD Detection for Specialized LLMs
by: Gupta, Ayush, et al.
Published: (2025)
by: Gupta, Ayush, et al.
Published: (2025)
TeleLoRA: Teleporting Model-Specific Alignment Across LLMs
by: Lin, Xiao, et al.
Published: (2025)
by: Lin, Xiao, et al.
Published: (2025)
AGENT: An Aerial Vehicle Generation and Design Tool Using Large Language Models
by: Samplawski, Colin, et al.
Published: (2025)
by: Samplawski, Colin, et al.
Published: (2025)
TOGA: Temporally Grounded Open-Ended Video QA with Weak Supervision
by: Gupta, Ayush, et al.
Published: (2025)
by: Gupta, Ayush, et al.
Published: (2025)
Spatio-Temporal Pruning for Compressed Spiking Large Language Models
by: Jiang, Yi, et al.
Published: (2025)
by: Jiang, Yi, et al.
Published: (2025)
Closed-Loop Neural Activation Control in Vision-Language-Action Models
by: Babu, Abhijith, et al.
Published: (2026)
by: Babu, Abhijith, et al.
Published: (2026)
Taylor-Model Physics-Informed Neural Networks (PINNs) for Ordinary Differential Equations
by: Nagesh, Chandra Kanth, et al.
Published: (2025)
by: Nagesh, Chandra Kanth, et al.
Published: (2025)
Melanoma Detection with Uncertainty Quantification
by: Kim, SangHyuk, et al.
Published: (2024)
by: Kim, SangHyuk, et al.
Published: (2024)
Optimal Abstractions for Verifying Properties of Kolmogorov-Arnold Networks (KANs)
by: Schwartz, Noah, et al.
Published: (2026)
by: Schwartz, Noah, et al.
Published: (2026)
SAFE-NID: Self-Attention with Normalizing-Flow Encodings for Network Intrusion Detection Dataset
by: Matejek, Brian, et al.
Published: (2025)
by: Matejek, Brian, et al.
Published: (2025)
Backpropagation-Free Metropolis-Adjusted Langevin Algorithm
by: Cobb, Adam D., et al.
Published: (2025)
by: Cobb, Adam D., et al.
Published: (2025)
Enhancing Cross-Modal Contextual Congruence for Crowdfunding Success using Knowledge-infused Learning
by: Padhi, Trilok, et al.
Published: (2024)
by: Padhi, Trilok, et al.
Published: (2024)
Hessian-Enhanced Token Attribution (HETA): Interpreting Autoregressive LLMs
by: Pramanik, Vishal, et al.
Published: (2026)
by: Pramanik, Vishal, et al.
Published: (2026)
Just KIDDIN: Knowledge Infusion and Distillation for Detection of INdecent Memes
by: Garg, Rahul, et al.
Published: (2024)
by: Garg, Rahul, et al.
Published: (2024)
Concept-based Analysis of Neural Networks via Vision-Language Models
by: Mangal, Ravi, et al.
Published: (2024)
by: Mangal, Ravi, et al.
Published: (2024)
Models Can Model, But Can't Bind: Structured Grounding in Text-to-Optimization
by: Gao, Zhiqi, et al.
Published: (2026)
by: Gao, Zhiqi, et al.
Published: (2026)
Safety Monitoring for Learning-Enabled Cyber-Physical Systems in Out-of-Distribution Scenarios
by: Lin, Vivian, et al.
Published: (2025)
by: Lin, Vivian, et al.
Published: (2025)
Query2Uncertainty: Robust Uncertainty Quantification and Calibration for 3D Object Detection under Distribution Shift
by: Beemelmanns, Till, et al.
Published: (2026)
by: Beemelmanns, Till, et al.
Published: (2026)
From Reddit to Generative AI: Evaluating Large Language Models for Anxiety Support Fine-tuned on Social Media Data
by: Kursuncu, Ugur, et al.
Published: (2025)
by: Kursuncu, Ugur, et al.
Published: (2025)
Playing Devil's Advocate: Unmasking Toxicity and Vulnerabilities in Large Vision-Language Models
by: Erol, Abdulkadir, et al.
Published: (2025)
by: Erol, Abdulkadir, et al.
Published: (2025)
Calibrated Physics-Informed Uncertainty Quantification
by: Gopakumar, Vignesh, et al.
Published: (2025)
by: Gopakumar, Vignesh, et al.
Published: (2025)
Interpolation of GEDI Biomass Estimates with Calibrated Uncertainty Quantification
by: Young, Robin, et al.
Published: (2026)
by: Young, Robin, et al.
Published: (2026)
Benchmarking LLMs via Uncertainty Quantification
by: Ye, Fanghua, et al.
Published: (2024)
by: Ye, Fanghua, et al.
Published: (2024)
LUQ: Long-text Uncertainty Quantification for LLMs
by: Zhang, Caiqi, et al.
Published: (2024)
by: Zhang, Caiqi, et al.
Published: (2024)
Uncertainty-Aware Attention Heads: Efficient Unsupervised Uncertainty Quantification for LLMs
by: Vazhentsev, Artem, et al.
Published: (2025)
by: Vazhentsev, Artem, et al.
Published: (2025)
Uncertainty Quantification and Confidence Calibration in Large Language Models: A Survey
by: Liu, Xiaoou, et al.
Published: (2025)
by: Liu, Xiaoou, et al.
Published: (2025)
On Subjective Uncertainty Quantification and Calibration in Natural Language Generation
by: Wang, Ziyu, et al.
Published: (2024)
by: Wang, Ziyu, et al.
Published: (2024)
MAQA: Evaluating Uncertainty Quantification in LLMs Regarding Data Uncertainty
by: Yang, Yongjin, et al.
Published: (2024)
by: Yang, Yongjin, et al.
Published: (2024)
Calibration Collapse Under Sycophancy Fine-Tuning: How Reward Hacking Breaks Uncertainty Quantification in LLMs
by: Sahoo, Subramanyam
Published: (2026)
by: Sahoo, Subramanyam
Published: (2026)
OCCUQ: Exploring Efficient Uncertainty Quantification for 3D Occupancy Prediction
by: Heidrich, Severin, et al.
Published: (2025)
by: Heidrich, Severin, et al.
Published: (2025)
TriFusion-AE: Language-Guided Depth and LiDAR Fusion for Robust Point Cloud Processing
by: Neogi, Susmit
Published: (2025)
by: Neogi, Susmit
Published: (2025)
Revisiting Multi-Modal LLM Evaluation
by: Lu, Jian, et al.
Published: (2024)
by: Lu, Jian, et al.
Published: (2024)
From Calibration to Collaboration: LLM Uncertainty Quantification Should Be More Human-Centered
by: Devic, Siddartha, et al.
Published: (2025)
by: Devic, Siddartha, et al.
Published: (2025)
Uncertainty Quantification in Calibration and Simulation of Thermo-Chemical Curing of Epoxy Resins
by: Tröger, Jendrik-Alexander, et al.
Published: (2026)
by: Tröger, Jendrik-Alexander, et al.
Published: (2026)
Similar Items
-
From Actions to Understanding: Conformal Interpretability of Temporal Concepts in LLM Agents
by: Padhi, Trilok, et al.
Published: (2026) -
Scalable Bayesian Low-Rank Adaptation of Large Language Models via Stochastic Variational Subspace Inference
by: Samplawski, Colin, et al.
Published: (2025) -
Addressing Uncertainty in LLMs to Enhance Reliability in Generative AI
by: Kaur, Ramneet, et al.
Published: (2024) -
Privacy Preserving In-Context-Learning Framework for Large Language Models
by: Bhusal, Bishnu, et al.
Published: (2025) -
Breaking Bad: Interpretability-Based Safety Audits of State-of-the-Art LLMs
by: Agarwal, Krishiv, et al.
Published: (2026)