CLEAR: Revealing How Noise and Ambiguity Degrade Reliability in LLMs for Medicine
Fuente:
arXiv
Saved in:
| Main Authors: | Guo, Kevin H., Yan, Chao, Baidya, Avinash, Brown, Katherine, Gao, Xiang, Xiong, Juming, Yin, Zhijun, Malin, Bradley A. |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Stop Listening to Me! How Multi-turn Conversations Can Degrade LLM Reliability
by: Guo, Kevin H., et al.
Published: (2026)
by: Guo, Kevin H., et al.
Published: (2026)
It's Not Always Sycophancy: Measuring LLM Conformity as a Function of Epistemic Uncertainty
by: Guo, Kevin H., et al.
Published: (2026)
by: Guo, Kevin H., et al.
Published: (2026)
Learning When to Sample: Confidence-Aware Self-Consistency for Efficient LLM Chain-of-Thought Reasoning
by: Xiong, Juming, et al.
Published: (2026)
by: Xiong, Juming, et al.
Published: (2026)
Vectors Are Not Neutral: Sensitive-Information Inference from Exported LLM Representations in Summarization
by: Liu, Weixin, et al.
Published: (2026)
by: Liu, Weixin, et al.
Published: (2026)
The Behavior Gap: Evaluating Zero-shot LLM Agents in Complex Task-Oriented Dialogs
by: Baidya, Avinash, et al.
Published: (2025)
by: Baidya, Avinash, et al.
Published: (2025)
CLEAR-KGQA: Clarification-Enhanced Ambiguity Resolution for Knowledge Graph Question Answering
by: Wen, Liqiang, et al.
Published: (2025)
by: Wen, Liqiang, et al.
Published: (2025)
CLEAR: Unlocking Generative Potential for Degraded Image Understanding in Unified Multimodal Models
by: Hao, Xiangzhao, et al.
Published: (2026)
by: Hao, Xiangzhao, et al.
Published: (2026)
Blending Human and LLM Expertise to Detect Hallucinations and Omissions in Mental Health Chatbot Responses
by: Hussain, Khizar, et al.
Published: (2026)
by: Hussain, Khizar, et al.
Published: (2026)
Optimizing Domain-Adaptive Self-Supervised Learning for Clinical Voice-Based Disease Classification
by: Liu, Weixin, et al.
Published: (2026)
by: Liu, Weixin, et al.
Published: (2026)
RIMRULE: Improving Tool-Using Language Agents via MDL-Guided Rule Learning
by: Gao, Xiang, et al.
Published: (2025)
by: Gao, Xiang, et al.
Published: (2025)
Do You Know What You Are Talking About? Characterizing Query-Knowledge Relevance For Reliable Retrieval Augmented Generation
by: Li, Zhuohang, et al.
Published: (2024)
by: Li, Zhuohang, et al.
Published: (2024)
Identifiable Object Representations under Spatial Ambiguities
by: Kori, Avinash, et al.
Published: (2025)
by: Kori, Avinash, et al.
Published: (2025)
CLEAR: Contrasting Textual Feedback with Experts and Amateurs for Reasoning
by: Rufail, Andrew, et al.
Published: (2025)
by: Rufail, Andrew, et al.
Published: (2025)
SMOTE-DP: Improving Privacy-Utility Tradeoff with Synthetic Data
by: Zhou, Yan, et al.
Published: (2025)
by: Zhou, Yan, et al.
Published: (2025)
Detecting the Machine: A Comprehensive Benchmark of AI-Generated Text Detectors Across Architectures, Domains, and Adversarial Conditions
by: Baidya, Madhav S., et al.
Published: (2026)
by: Baidya, Madhav S., et al.
Published: (2026)
CLEAR-IR: Clarity-Enhanced Active Reconstruction of Infrared Imagery
by: Shankar, Nathan, et al.
Published: (2025)
by: Shankar, Nathan, et al.
Published: (2025)
PRISM: Differentially Private Synthetic Data with Structure-Aware Budget Allocation for Prediction
by: Asiaee, Amir, et al.
Published: (2026)
by: Asiaee, Amir, et al.
Published: (2026)
Risk-Equalized Differentially Private Synthetic Data: Protecting Outliers by Controlling Record-Level Influence
by: Asiaee, Amir, et al.
Published: (2026)
by: Asiaee, Amir, et al.
Published: (2026)
Catalysts of Conversation: Examining Interaction Dynamics Between Topic Initiators and Commentors in Alzheimer's Disease Online Communities
by: Ni, Congning, et al.
Published: (2024)
by: Ni, Congning, et al.
Published: (2024)
Results for response and reliability-based optimization
by: Baidya, Sanjay
Published: (2026)
by: Baidya, Sanjay
Published: (2026)
Optimal parameters and structural responses for response and reliability based optimization
by: Baidya, Sanjay
Published: (2025)
by: Baidya, Sanjay
Published: (2025)
PassiveQA: A Three-Action Framework for Epistemically Calibrated Question Answering via Supervised Finetuning
by: Baidya, Madhav S
Published: (2026)
by: Baidya, Madhav S
Published: (2026)
PROXIMA: A Reliability Scoring Framework for Proxy Metrics in Online Controlled Experiments
by: Amudala, Avinash
Published: (2026)
by: Amudala, Avinash
Published: (2026)
Revealing the Two-Fold Ambiguity: Tau Momentum Reconstruction and Its Impact on Entanglement Observables
by: Zhou, Xiang, et al.
Published: (2026)
by: Zhou, Xiang, et al.
Published: (2026)
SAC3: Reliable Hallucination Detection in Black-Box Language Models via Semantic-aware Cross-check Consistency
by: Zhang, Jiaxin, et al.
Published: (2023)
by: Zhang, Jiaxin, et al.
Published: (2023)
MHGraphBench: Knowledge Graph-Grounded Benchmarking of Mental Health Knowledge in Large Language Models
by: Liu, Weixin, et al.
Published: (2026)
by: Liu, Weixin, et al.
Published: (2026)
WATER BILLS – ARE THEY CLEAR TO RESIDENTIAL USERS?
by: Rita Martins
Published: (2012)
by: Rita Martins
Published: (2012)
Chemical Literature Exercises and Resources (CLEAR).
by: Hostettler, John D., et al.
Published: (1982)
by: Hostettler, John D., et al.
Published: (1982)
How Close Are We? Limitations and Progress of AI Models in Banff Lesion Scoring
by: Zhu, Yanfan, et al.
Published: (2025)
by: Zhu, Yanfan, et al.
Published: (2025)
VERI-DPO: Evidence-Aware Alignment for Clinical Summarization via Claim Verification and Direct Preference Optimization
by: Liu, Weixin, et al.
Published: (2026)
by: Liu, Weixin, et al.
Published: (2026)
Reliable Explanations or Random Noise? A Reliability Metric for XAI
by: Sengupta, Poushali, et al.
Published: (2026)
by: Sengupta, Poushali, et al.
Published: (2026)
Ambiguity in practice: Benchmarks for the implementation of CCW Protocol V
by: Moyes, Richard, et al.
Published: (2008)
by: Moyes, Richard, et al.
Published: (2008)
Revealing Modular Gradient Noise Imbalance in LLMs: Calibrating Adam via Signal-to-Noise Ratio
by: Wen, Ziqing, et al.
Published: (2026)
by: Wen, Ziqing, et al.
Published: (2026)
CLEAR: Character Unlearning in Textual and Visual Modalities
by: Dontsov, Alexey, et al.
Published: (2024)
by: Dontsov, Alexey, et al.
Published: (2024)
CLEAR: Calibrated Learning for Epistemic and Aleatoric Risk
by: Azizi, Ilia, et al.
Published: (2025)
by: Azizi, Ilia, et al.
Published: (2025)
A Reinforcement Learning Approach to Synthetic Data Generation
by: Espinosa-Dice, Natalia, et al.
Published: (2025)
by: Espinosa-Dice, Natalia, et al.
Published: (2025)
SECNEURON: Reliable and Flexible Abuse Control in Local LLMs via Hybrid Neuron Encryption
by: Wang, Zhiqiang, et al.
Published: (2025)
by: Wang, Zhiqiang, et al.
Published: (2025)
LLMs as Deceptive Agents: How Role-Based Prompting Induces Semantic Ambiguity in Puzzle Tasks
by: Yoo, Seunghyun
Published: (2025)
by: Yoo, Seunghyun
Published: (2025)
Events in Noise-Driven Oscillators: Markov Renewal Processes and the "Unruly" Breakdown of Phase-Reduction Theory
by: Karamchandani, Avinash J.
Published: (2024)
by: Karamchandani, Avinash J.
Published: (2024)
A Blueprint for AI-Driven Software Quality: Integrating LLMs with Established Standards
by: Patil, Avinash
Published: (2025)
by: Patil, Avinash
Published: (2025)
Similar Items
-
Stop Listening to Me! How Multi-turn Conversations Can Degrade LLM Reliability
by: Guo, Kevin H., et al.
Published: (2026) -
It's Not Always Sycophancy: Measuring LLM Conformity as a Function of Epistemic Uncertainty
by: Guo, Kevin H., et al.
Published: (2026) -
Learning When to Sample: Confidence-Aware Self-Consistency for Efficient LLM Chain-of-Thought Reasoning
by: Xiong, Juming, et al.
Published: (2026) -
Vectors Are Not Neutral: Sensitive-Information Inference from Exported LLM Representations in Summarization
by: Liu, Weixin, et al.
Published: (2026) -
The Behavior Gap: Evaluating Zero-shot LLM Agents in Complex Task-Oriented Dialogs
by: Baidya, Avinash, et al.
Published: (2025)