Seeing and Reasoning with Confidence: Supercharging Multimodal LLMs with an Uncertainty-Aware Agentic Framework
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhi, Zhuo, Feng, Chen, Daneshmend, Adam, Orlu, Mine, Demosthenous, Andreas, Yin, Lu, Li, Da, Liu, Ziquan, Rodrigues, Miguel R. D. |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Borrowing Treasures from Neighbors: In-Context Learning for Multimodal Learning with Missing Modalities and Data Scarcity
von: Zhi, Zhuo, et al.
Veröffentlicht: (2024)
von: Zhi, Zhuo, et al.
Veröffentlicht: (2024)
HgbNet: predicting hemoglobin level/anemia degree from EHR data
von: Zhi, Zhuo, et al.
Veröffentlicht: (2024)
von: Zhi, Zhuo, et al.
Veröffentlicht: (2024)
Sensory-driven microinterventions for improved health and wellbeing
von: Abdalla, Youssef, et al.
Veröffentlicht: (2025)
von: Abdalla, Youssef, et al.
Veröffentlicht: (2025)
SeeingEye: Agentic Information Flow Unlocks Multimodal Reasoning In Text-only LLMs
von: Zhang, Weijia, et al.
Veröffentlicht: (2025)
von: Zhang, Weijia, et al.
Veröffentlicht: (2025)
Water Pretreatment System for Rapid Preconcentration of Bacteria
von: Panayiota Demosthenous
Veröffentlicht: (2025)
von: Panayiota Demosthenous
Veröffentlicht: (2025)
Breaking disciplinary silos: A global approach to interprofessional education
von: Fraide A. Ganotice, et al.
Veröffentlicht: (2024)
von: Fraide A. Ganotice, et al.
Veröffentlicht: (2024)
PROSAC: Provably Safe Certification for Machine Learning Models under Adversarial Attacks
von: Feng, Chen, et al.
Veröffentlicht: (2024)
von: Feng, Chen, et al.
Veröffentlicht: (2024)
Euclid: Supercharging Multimodal LLMs with Synthetic High-Fidelity Visual Descriptions
von: Zhang, Jiarui, et al.
Veröffentlicht: (2024)
von: Zhang, Jiarui, et al.
Veröffentlicht: (2024)
Confidence-Aware Alignment Makes Reasoning LLMs More Reliable
von: Chen, Kejia, et al.
Veröffentlicht: (2026)
von: Chen, Kejia, et al.
Veröffentlicht: (2026)
To See or To Read: User Behavior Reasoning in Multimodal LLMs
von: Dong, Tianning, et al.
Veröffentlicht: (2025)
von: Dong, Tianning, et al.
Veröffentlicht: (2025)
Can't See the Forest for the Trees: Benchmarking Multimodal Safety Awareness for Multimodal LLMs
von: Wang, Wenxuan, et al.
Veröffentlicht: (2025)
von: Wang, Wenxuan, et al.
Veröffentlicht: (2025)
Detect in Any Scene: An Agentic Framework for Object Detection with Experience-Aware Reasoning
von: Zhang, Wenlun, et al.
Veröffentlicht: (2026)
von: Zhang, Wenlun, et al.
Veröffentlicht: (2026)
A Novel Semi‐Automated Pipeline for Optimizing 3D‐Printed Drug Formulations
von: Youssef Abdalla, et al.
Veröffentlicht: (2025)
von: Youssef Abdalla, et al.
Veröffentlicht: (2025)
Can LLMs Express Their Uncertainty? An Empirical Evaluation of Confidence Elicitation in LLMs
von: Xiong, Miao, et al.
Veröffentlicht: (2023)
von: Xiong, Miao, et al.
Veröffentlicht: (2023)
Multimodal LLMs See Sentiment
von: da Silva, Neemias B., et al.
Veröffentlicht: (2025)
von: da Silva, Neemias B., et al.
Veröffentlicht: (2025)
VideoThinker: Building Agentic VideoLLMs with LLM-Guided Tool Reasoning
von: Li, Chenglin, et al.
Veröffentlicht: (2026)
von: Li, Chenglin, et al.
Veröffentlicht: (2026)
Can Multimodal LLMs See Science Instruction? Benchmarking Pedagogical Reasoning in K-12 Classroom Videos
von: Shen, Yixuan, et al.
Veröffentlicht: (2026)
von: Shen, Yixuan, et al.
Veröffentlicht: (2026)
AutoGluon-Multimodal (AutoMM): Supercharging Multimodal AutoML with Foundation Models
von: Tang, Zhiqiang, et al.
Veröffentlicht: (2024)
von: Tang, Zhiqiang, et al.
Veröffentlicht: (2024)
Impute With Confidence: A Framework for Uncertainty Aware Multivariate Time Series Imputation
von: Weatherhead, Addison, et al.
Veröffentlicht: (2025)
von: Weatherhead, Addison, et al.
Veröffentlicht: (2025)
CER: Confidence Enhanced Reasoning in LLMs
von: Razghandi, Ali, et al.
Veröffentlicht: (2025)
von: Razghandi, Ali, et al.
Veröffentlicht: (2025)
Beyond Seeing: Evaluating Multimodal LLMs on Tool-Enabled Image Perception, Transformation, and Reasoning
von: Guo, Xingang, et al.
Veröffentlicht: (2025)
von: Guo, Xingang, et al.
Veröffentlicht: (2025)
Seeing is Fixing: Cross-Modal Reasoning with Multimodal LLMs for Visual Software Issue Fixing
von: Huang, Kai, et al.
Veröffentlicht: (2025)
von: Huang, Kai, et al.
Veröffentlicht: (2025)
Seeing with You: Perception-Reasoning Coevolution for Multimodal Reasoning
von: Miao, Ziqi, et al.
Veröffentlicht: (2026)
von: Miao, Ziqi, et al.
Veröffentlicht: (2026)
Seeing the Context: Rich Visual Context-Aware Speech Recognition via Multimodal Reasoning
von: Tian, Wenjie, et al.
Veröffentlicht: (2026)
von: Tian, Wenjie, et al.
Veröffentlicht: (2026)
GrACE: A Generative Approach to Better Confidence Elicitation and Efficient Test-Time Scaling in Large Language Models
von: Zhang, Zhaohan, et al.
Veröffentlicht: (2025)
von: Zhang, Zhaohan, et al.
Veröffentlicht: (2025)
Supercharging Federated Intelligence Retrieval
von: Stripelis, Dimitris, et al.
Veröffentlicht: (2026)
von: Stripelis, Dimitris, et al.
Veröffentlicht: (2026)
Reasoning LLMs for User-Aware Multimodal Conversational Agents
von: Rahimi, Hamed, et al.
Veröffentlicht: (2025)
von: Rahimi, Hamed, et al.
Veröffentlicht: (2025)
From On-chain to Macro: Assessing the Importance of Data Source Diversity in Cryptocurrency Market Forecasting
von: Demosthenous, Giorgos, et al.
Veröffentlicht: (2025)
von: Demosthenous, Giorgos, et al.
Veröffentlicht: (2025)
Deep Learning-Based Noninvasive Screening of Type 2 Diabetes with Chest X-ray Images and Electronic Health Records
von: Gundapaneni, Sanjana, et al.
Veröffentlicht: (2024)
von: Gundapaneni, Sanjana, et al.
Veröffentlicht: (2024)
Supercharging Federated Learning with Flower and NVIDIA FLARE
von: Roth, Holger R., et al.
Veröffentlicht: (2024)
von: Roth, Holger R., et al.
Veröffentlicht: (2024)
See Less, See Right: Bi-directional Perceptual Shaping For Multimodal Reasoning
von: Zhang, Shuoshuo, et al.
Veröffentlicht: (2025)
von: Zhang, Shuoshuo, et al.
Veröffentlicht: (2025)
Agentic Confidence Calibration
von: Zhang, Jiaxin, et al.
Veröffentlicht: (2026)
von: Zhang, Jiaxin, et al.
Veröffentlicht: (2026)
Seeing Clearly, Reasoning Confidently: Plug-and-Play Remedies for Vision Language Model Blindness
von: Hu, Xin, et al.
Veröffentlicht: (2026)
von: Hu, Xin, et al.
Veröffentlicht: (2026)
Supercharging Floorplan Localization with Semantic Rays
von: Grader, Yuval, et al.
Veröffentlicht: (2025)
von: Grader, Yuval, et al.
Veröffentlicht: (2025)
Supercharging Graph Transformers with Advective Diffusion
von: Wu, Qitian, et al.
Veröffentlicht: (2023)
von: Wu, Qitian, et al.
Veröffentlicht: (2023)
PaperMind: Benchmarking Agentic Reasoning and Critique over Scientific Papers in Multimodal LLMs
von: Zhao, Yanjun, et al.
Veröffentlicht: (2026)
von: Zhao, Yanjun, et al.
Veröffentlicht: (2026)
Octopus: Agentic Multimodal Reasoning with Six-Capability Orchestration
von: Guo, Yifu, et al.
Veröffentlicht: (2025)
von: Guo, Yifu, et al.
Veröffentlicht: (2025)
Robust Checkpoint Selection for Multimodal LLMs via Agentic Evaluation and Stability-Aware Ranking
von: Xu, Qinwu, et al.
Veröffentlicht: (2026)
von: Xu, Qinwu, et al.
Veröffentlicht: (2026)
Knowledge Boundary Discovery for Large Language Models
von: Wang, Ziquan, et al.
Veröffentlicht: (2026)
von: Wang, Ziquan, et al.
Veröffentlicht: (2026)
Confidence Contours: Uncertainty-Aware Annotation for Medical Semantic Segmentation
von: Ye, Andre, et al.
Veröffentlicht: (2023)
von: Ye, Andre, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
Borrowing Treasures from Neighbors: In-Context Learning for Multimodal Learning with Missing Modalities and Data Scarcity
von: Zhi, Zhuo, et al.
Veröffentlicht: (2024) -
HgbNet: predicting hemoglobin level/anemia degree from EHR data
von: Zhi, Zhuo, et al.
Veröffentlicht: (2024) -
Sensory-driven microinterventions for improved health and wellbeing
von: Abdalla, Youssef, et al.
Veröffentlicht: (2025) -
SeeingEye: Agentic Information Flow Unlocks Multimodal Reasoning In Text-only LLMs
von: Zhang, Weijia, et al.
Veröffentlicht: (2025) -
Water Pretreatment System for Rapid Preconcentration of Bacteria
von: Panayiota Demosthenous
Veröffentlicht: (2025)