Grounded or Guessing? LVLM Confidence Estimation via Blind-Image Contrastive Ranking
Fuente:
arXiv
Salvato in:
| Autori principali: | Khanmohammadi, Reza, Miahi, Erfan, Kaur, Simerjot, Smiley, Charese H., Brugere, Ivan, Thind, Kundan, Ghassemi, Mohammad M. |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
How Reliable are Confidence Estimators for Large Reasoning Models? A Systematic Benchmark on High-Stakes Domains
di: Khanmohammadi, Reza, et al.
Pubblicazione: (2026)
di: Khanmohammadi, Reza, et al.
Pubblicazione: (2026)
Calibrating LLM Confidence by Probing Perturbed Representation Stability
di: Khanmohammadi, Reza, et al.
Pubblicazione: (2025)
di: Khanmohammadi, Reza, et al.
Pubblicazione: (2025)
The Influence of Biomedical Research on Future Business Funding: Analyzing Scientific Impact and Content in Industrial Investments
di: Khanmohammadi, Reza, et al.
Pubblicazione: (2024)
di: Khanmohammadi, Reza, et al.
Pubblicazione: (2024)
Grounding LLM Reasoning with Knowledge Graphs
di: Amayuelas, Alfonso, et al.
Pubblicazione: (2025)
di: Amayuelas, Alfonso, et al.
Pubblicazione: (2025)
FinQAPT: Empowering Financial Decisions with End-to-End LLM-driven Question Answering Pipeline
di: Singh, Kuldeep, et al.
Pubblicazione: (2024)
di: Singh, Kuldeep, et al.
Pubblicazione: (2024)
Conservative Bias in Large Language Models: Measuring Relation Predictions
di: Aguda, Toyin, et al.
Pubblicazione: (2025)
di: Aguda, Toyin, et al.
Pubblicazione: (2025)
FinNLI: Novel Dataset for Multi-Genre Financial Natural Language Inference Benchmarking
di: Magomere, Jabez, et al.
Pubblicazione: (2025)
di: Magomere, Jabez, et al.
Pubblicazione: (2025)
A Variational Approach for Mitigating Entity Bias in Relation Extraction
di: Mensah, Samuel, et al.
Pubblicazione: (2025)
di: Mensah, Samuel, et al.
Pubblicazione: (2025)
Large Language Models as Financial Data Annotators: A Study on Effectiveness and Efficiency
di: Aguda, Toyin, et al.
Pubblicazione: (2024)
di: Aguda, Toyin, et al.
Pubblicazione: (2024)
Understanding and Exploiting Weight Update Sparsity for Communication-Efficient Distributed RL
di: Miahi, Erfan, et al.
Pubblicazione: (2026)
di: Miahi, Erfan, et al.
Pubblicazione: (2026)
Iterative Prompt Refinement for Radiation Oncology Symptom Extraction Using Teacher-Student Large Language Models
di: Khanmohammadi, Reza, et al.
Pubblicazione: (2024)
di: Khanmohammadi, Reza, et al.
Pubblicazione: (2024)
Distill and Align Decomposition for Enhanced Claim Verification
di: Magomere, Jabez, et al.
Pubblicazione: (2026)
di: Magomere, Jabez, et al.
Pubblicazione: (2026)
Deep FinResearch Bench: Evaluating AI's Ability to Conduct Professional Financial Investment Research
di: Haque, Mirazul, et al.
Pubblicazione: (2026)
di: Haque, Mirazul, et al.
Pubblicazione: (2026)
Hybrid Student-Teacher Large Language Model Refinement for Cancer Toxicity Symptom Extraction
di: Khanmohammadi, Reza, et al.
Pubblicazione: (2024)
di: Khanmohammadi, Reza, et al.
Pubblicazione: (2024)
Reduction in preparatory brain activity preceding gait initiation in individuals with chronic ankle instability: A movement‐related cortical potential study
di: Zivar Beyraghi, et al.
Pubblicazione: (2024)
di: Zivar Beyraghi, et al.
Pubblicazione: (2024)
GenPlanX. Generation of Plans and Execution
di: Borrajo, Daniel, et al.
Pubblicazione: (2025)
di: Borrajo, Daniel, et al.
Pubblicazione: (2025)
AI Analyst: Framework and Comprehensive Evaluation of Large Language Models for Financial Time Series Report Generation
di: Fons, Elizabeth, et al.
Pubblicazione: (2025)
di: Fons, Elizabeth, et al.
Pubblicazione: (2025)
Energy-Efficient Approximate Full Adders Applying Memristive Serial IMPLY Logic For Image Processing
di: Fatemieh, Seyed Erfan, et al.
Pubblicazione: (2024)
di: Fatemieh, Seyed Erfan, et al.
Pubblicazione: (2024)
Energy‐Efficient Approximate Full Adders Applying Memristive Serial IMPLY Logic for Image Processing
di: Seyed Erfan Fatemieh, et al.
Pubblicazione: (2026)
di: Seyed Erfan Fatemieh, et al.
Pubblicazione: (2026)
Luminescent Multifunctional Nanomaterials: Capacitive Removal and Enhanced Detection Efficiency of Heavy Metals Ions for Advanced Water and Wastewater Treatment Application
di: Karim Khanmohammadi Chenab, et al.
Pubblicazione: (2024)
di: Karim Khanmohammadi Chenab, et al.
Pubblicazione: (2024)
Mainstreaming gender in the BOBLME Project
di: Brugere, Cecile
Pubblicazione: (2012)
di: Brugere, Cecile
Pubblicazione: (2012)
La desaparición de la obra
di: Fabienne Brugère
Pubblicazione: (2007)
di: Fabienne Brugère
Pubblicazione: (2007)
MaskCD: Mitigating LVLM Hallucinations by Image Head Masked Contrastive Decoding
di: Deng, Jingyuan, et al.
Pubblicazione: (2025)
di: Deng, Jingyuan, et al.
Pubblicazione: (2025)
Impact of Social Comparison on Fake News Release and News Credibility
di: Esfidani, Mohammad Rahim, et al.
Pubblicazione: (2023)
di: Esfidani, Mohammad Rahim, et al.
Pubblicazione: (2023)
SINDyG: Sparse Identification of Nonlinear Dynamical Systems from Graph-Structured Data, with Applications to Stuart-Landau Oscillator Networks
di: Basiri, Mohammad Amin, et al.
Pubblicazione: (2024)
di: Basiri, Mohammad Amin, et al.
Pubblicazione: (2024)
The Shift From Neo‐Ottomanism to the Century of Türkiye: Domestic, Regional, and Global Implications
di: Mohammad Hadi Khanmohammadi, et al.
Pubblicazione: (2026)
di: Mohammad Hadi Khanmohammadi, et al.
Pubblicazione: (2026)
Robustness of Transformer-Based Fluence Map Prediction Under Clinically Realistic Perturbations
di: Mgboh, Ujunwa, et al.
Pubblicazione: (2026)
di: Mgboh, Ujunwa, et al.
Pubblicazione: (2026)
FluenceFormer: Transformer-Driven Multi-Beam Fluence Map Regression for Radiotherapy Planning
di: Mgboh, Ujunwa, et al.
Pubblicazione: (2025)
di: Mgboh, Ujunwa, et al.
Pubblicazione: (2025)
Energy-Efficient and Fast Memristor-based Serial Multipliers Applicable in Image Processing
di: Fatemieh, Seyed Erfan, et al.
Pubblicazione: (2024)
di: Fatemieh, Seyed Erfan, et al.
Pubblicazione: (2024)
Kestrel: Grounding Self-Refinement for LVLM Hallucination Mitigation
di: Mao, Jiawei, et al.
Pubblicazione: (2026)
di: Mao, Jiawei, et al.
Pubblicazione: (2026)
Utilizing patient data: A tutorial on predicting second cancer with machine learning models
di: Hossein Sadeghi, et al.
Pubblicazione: (2024)
di: Hossein Sadeghi, et al.
Pubblicazione: (2024)
Support-Guessing Decoding Algorithms in the Sum-Rank Metric
di: Jerkovits, Thomas, et al.
Pubblicazione: (2024)
di: Jerkovits, Thomas, et al.
Pubblicazione: (2024)
WorldMemArena: Evaluating Multimodal Agent Memory Through Action-World Interaction
di: Liu, Chengzhi, et al.
Pubblicazione: (2026)
di: Liu, Chengzhi, et al.
Pubblicazione: (2026)
Modernizing Legacy Enterprise Configuration Interfaces: A Bootstrap-Based Approach to Oracle Configurator UI Enhancement
di: Nitin Thind
Pubblicazione: (2026)
di: Nitin Thind
Pubblicazione: (2026)
L-Estimation of Population Quantiles Using Ranked Set Sampling
di: Jozani, Mohammad Jafari, et al.
Pubblicazione: (2026)
di: Jozani, Mohammad Jafari, et al.
Pubblicazione: (2026)
Automated stereotactic radiosurgery planning using a human-in-the-loop reasoning large language model agent
di: Nusrat, Humza, et al.
Pubblicazione: (2025)
di: Nusrat, Humza, et al.
Pubblicazione: (2025)
WhisperPipe: A Resource-Efficient Streaming Architecture for Real-Time Automatic Speech Recognition
di: Ramezani, Erfan, et al.
Pubblicazione: (2026)
di: Ramezani, Erfan, et al.
Pubblicazione: (2026)
Hierarchical Contextual Grounding LVLM: Enhancing Fine-Grained Visual-Language Understanding with Robust Grounding
di: Guo, Leilei, et al.
Pubblicazione: (2025)
di: Guo, Leilei, et al.
Pubblicazione: (2025)
Guess the Age of Photos: An Interactive Web Platform for Historical Image Age Estimation
di: Yucedag, Hasan, et al.
Pubblicazione: (2025)
di: Yucedag, Hasan, et al.
Pubblicazione: (2025)
LVLM-Composer's Explicit Planning for Image Generation
di: Ramsey, Spencer, et al.
Pubblicazione: (2025)
di: Ramsey, Spencer, et al.
Pubblicazione: (2025)
Documenti analoghi
-
How Reliable are Confidence Estimators for Large Reasoning Models? A Systematic Benchmark on High-Stakes Domains
di: Khanmohammadi, Reza, et al.
Pubblicazione: (2026) -
Calibrating LLM Confidence by Probing Perturbed Representation Stability
di: Khanmohammadi, Reza, et al.
Pubblicazione: (2025) -
The Influence of Biomedical Research on Future Business Funding: Analyzing Scientific Impact and Content in Industrial Investments
di: Khanmohammadi, Reza, et al.
Pubblicazione: (2024) -
Grounding LLM Reasoning with Knowledge Graphs
di: Amayuelas, Alfonso, et al.
Pubblicazione: (2025) -
FinQAPT: Empowering Financial Decisions with End-to-End LLM-driven Question Answering Pipeline
di: Singh, Kuldeep, et al.
Pubblicazione: (2024)