The Confidence Paradox: Can LLM Know When It's Wrong
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Tripathi, Sahil, Nafis, Md Tabrez, Hussain, Imran, Gao, Jiechao |
|---|---|
| Format: | Preprint |
| Publié: |
2025
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
A Study of Hybrid and Evolutionary Metaheuristics for Single Hidden Layer Feedforward Neural Network Architecture
par: Kashyap, Gautam Siddharth, et autres
Publié: (2025)
par: Kashyap, Gautam Siddharth, et autres
Publié: (2025)
Grading and Anomaly Detection for Automated Retinal Image Analysis using Deep Learning
par: Malik, Syed Mohd Faisal, et autres
Publié: (2024)
par: Malik, Syed Mohd Faisal, et autres
Publié: (2024)
Can We Predict Your Next Move Without Breaking Your Privacy?
par: Soni, Arpita, et autres
Publié: (2025)
par: Soni, Arpita, et autres
Publié: (2025)
When Right Meets Wrong: Bilateral Context Conditioning with Reward-Confidence Correction for GRPO
par: Li, Yu, et autres
Publié: (2026)
par: Li, Yu, et autres
Publié: (2026)
From Text to Transformation: A Comprehensive Review of Large Language Models' Versatility
par: Kaur, Pravneet, et autres
Publié: (2024)
par: Kaur, Pravneet, et autres
Publié: (2024)
KnowRL: Teaching Language Models to Know What They Know
par: Kale, Sahil, et autres
Publié: (2025)
par: Kale, Sahil, et autres
Publié: (2025)
Does RAG Know When Retrieval Is Wrong? Diagnosing Context Compliance under Knowledge Conflict
par: Chen, Yihang, et autres
Publié: (2026)
par: Chen, Yihang, et autres
Publié: (2026)
FedMetaMed: Federated Meta-Learning for Personalized Medication in Distributed Healthcare Systems
par: Gao, Jiechao, et autres
Publié: (2024)
par: Gao, Jiechao, et autres
Publié: (2024)
Do Small Language Models Know When They're Wrong? Confidence-Based Cascade Scoring for Educational Assessment
par: Burleigh, Tyler
Publié: (2026)
par: Burleigh, Tyler
Publié: (2026)
Knowing When to Abstain: Medical LLMs Under Clinical Uncertainty
par: Machcha, Sravanthi, et autres
Publié: (2026)
par: Machcha, Sravanthi, et autres
Publié: (2026)
Multiple Choice Questions: Reasoning Makes Large Language Models (LLMs) More Self-Confident, Especially When They are Wrong
par: Fu, Tairan, et autres
Publié: (2025)
par: Fu, Tairan, et autres
Publié: (2025)
CaRT: Teaching LLM Agents to Know When They Know Enough
par: Liu, Grace, et autres
Publié: (2025)
par: Liu, Grace, et autres
Publié: (2025)
When Models Know When They Do Not Know: Calibration, Cascading, and Cleaning
par: Hao, Chenjie, et autres
Publié: (2026)
par: Hao, Chenjie, et autres
Publié: (2026)
Why and When LLM-Based Assistants Can Go Wrong: Investigating the Effectiveness of Prompt-Based Interactions for Software Help-Seeking
par: Khurana, Anjali, et autres
Publié: (2024)
par: Khurana, Anjali, et autres
Publié: (2024)
Know When to Explore: Difficulty-Aware Certainty as a Guide for LLM Reinforcement Learning
par: Li, Ang, et autres
Publié: (2025)
par: Li, Ang, et autres
Publié: (2025)
Federated Neural Architecture Search with Model-Agnostic Meta Learning
par: Huang, Xinyuan, et autres
Publié: (2025)
par: Huang, Xinyuan, et autres
Publié: (2025)
Anchorless Diversification for Parallel LLM Ideation
par: Ibrahim, Fares Nabil, et autres
Publié: (2026)
par: Ibrahim, Fares Nabil, et autres
Publié: (2026)
When the Majority is Wrong: Modeling Annotator Disagreement for Subjective Tasks
par: Fleisig, Eve, et autres
Publié: (2023)
par: Fleisig, Eve, et autres
Publié: (2023)
Stable but Wrong: When More Data Degrades Scientific Conclusions
par: Zhang, Zhipeng, et autres
Publié: (2026)
par: Zhang, Zhipeng, et autres
Publié: (2026)
Can AI Assistants Know What They Don't Know?
par: Cheng, Qinyuan, et autres
Publié: (2024)
par: Cheng, Qinyuan, et autres
Publié: (2024)
Beyond Words: Multimodal LLM Knows When to Speak
par: Liao, Zikai, et autres
Publié: (2025)
par: Liao, Zikai, et autres
Publié: (2025)
Do Retrieval Augmented Language Models Know When They Don't Know?
par: Zhou, Youchao, et autres
Publié: (2025)
par: Zhou, Youchao, et autres
Publié: (2025)
Towards Agents That Know When They Don't Know: Uncertainty as a Control Signal for Structured Reasoning
par: Stoisser, Josefa Lia, et autres
Publié: (2025)
par: Stoisser, Josefa Lia, et autres
Publié: (2025)
Epistemic Artificial Intelligence is Essential for Machine Learning Models to Truly 'Know When They Do Not Know'
par: Manchingal, Shireen Kudukkil, et autres
Publié: (2025)
par: Manchingal, Shireen Kudukkil, et autres
Publié: (2025)
The First Token Knows: Single-Decode Confidence for Hallucination Detection
par: Gabriel, Mina
Publié: (2026)
par: Gabriel, Mina
Publié: (2026)
Can Unconfident LLM Annotations Be Used for Confident Conclusions?
par: Gligorić, Kristina, et autres
Publié: (2024)
par: Gligorić, Kristina, et autres
Publié: (2024)
CollabStory: Multi-LLM Collaborative Story Generation and Authorship Analysis
par: Venkatraman, Saranya, et autres
Publié: (2024)
par: Venkatraman, Saranya, et autres
Publié: (2024)
OffTopicEval: When Large Language Models Enter the Wrong Chat, Almost Always!
par: Lei, Jingdi, et autres
Publié: (2025)
par: Lei, Jingdi, et autres
Publié: (2025)
Reasoning about Uncertainty: Do Reasoning Models Know When They Don't Know?
par: Mei, Zhiting, et autres
Publié: (2025)
par: Mei, Zhiting, et autres
Publié: (2025)
Know When You're Wrong: Aligning Confidence with Correctness for LLM Error Detection
par: Xiaohu, Xie, et autres
Publié: (2026)
par: Xiaohu, Xie, et autres
Publié: (2026)
Not Wrong, But Untrue: LLM Overconfidence in Document-Based Queries
par: Hagar, Nick, et autres
Publié: (2025)
par: Hagar, Nick, et autres
Publié: (2025)
What's Wrong? Refining Meeting Summaries with LLM Feedback
par: Kirstein, Frederic, et autres
Publié: (2024)
par: Kirstein, Frederic, et autres
Publié: (2024)
Epistemic Deep Learning: Enabling Machine Learning Models to Know When They Do Not Know
par: Manchingal, Shireen Kudukkil
Publié: (2025)
par: Manchingal, Shireen Kudukkil
Publié: (2025)
Right Prediction, Wrong Reasoning: Uncovering LLM Misalignment in RA Disease Diagnosis
par: Maharana, Umakanta, et autres
Publié: (2025)
par: Maharana, Umakanta, et autres
Publié: (2025)
When Safety Blocks Sense: Measuring Semantic Confusion in LLM Refusals
par: Anonto, Riad Ahmed, et autres
Publié: (2025)
par: Anonto, Riad Ahmed, et autres
Publié: (2025)
S2D-ALIGN: Shallow-to-Deep Auxiliary Learning for Anatomically-Grounded Radiology Report Generation
par: Gao, Jiechao, et autres
Publié: (2025)
par: Gao, Jiechao, et autres
Publié: (2025)
When Search Goes Wrong: Red-Teaming Web-Augmented Large Language Models
par: Ou, Haoran, et autres
Publié: (2025)
par: Ou, Haoran, et autres
Publié: (2025)
When Learning Rates Go Wrong: Early Structural Signals in PPO Actor-Critic
par: Fernández-Hernández, Alberto, et autres
Publié: (2026)
par: Fernández-Hernández, Alberto, et autres
Publié: (2026)
What Models Know, How Well They Know It: Knowledge-Weighted Fine-Tuning for Learning When to Say "I Don't Know"
par: Lee, Joosung, et autres
Publié: (2026)
par: Lee, Joosung, et autres
Publié: (2026)
Does Your Reasoning Model Implicitly Know When to Stop Thinking?
par: Huang, Zixuan, et autres
Publié: (2026)
par: Huang, Zixuan, et autres
Publié: (2026)
Documents similaires
-
A Study of Hybrid and Evolutionary Metaheuristics for Single Hidden Layer Feedforward Neural Network Architecture
par: Kashyap, Gautam Siddharth, et autres
Publié: (2025) -
Grading and Anomaly Detection for Automated Retinal Image Analysis using Deep Learning
par: Malik, Syed Mohd Faisal, et autres
Publié: (2024) -
Can We Predict Your Next Move Without Breaking Your Privacy?
par: Soni, Arpita, et autres
Publié: (2025) -
When Right Meets Wrong: Bilateral Context Conditioning with Reward-Confidence Correction for GRPO
par: Li, Yu, et autres
Publié: (2026) -
From Text to Transformation: A Comprehensive Review of Large Language Models' Versatility
par: Kaur, Pravneet, et autres
Publié: (2024)