Large Language Models Must Be Taught to Know What They Don't Know
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Kapoor, Sanyam, Gruver, Nate, Roberts, Manley, Collins, Katherine, Pal, Arka, Bhatt, Umang, Weller, Adrian, Dooley, Samuel, Goldblum, Micah, Wilson, Andrew Gordon |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Knowing What You Know Is Not Enough: Large Language Model Confidences Don't Align With Their Actions
von: Pal, Arka, et al.
Veröffentlicht: (2025)
von: Pal, Arka, et al.
Veröffentlicht: (2025)
The Lie Derivative for Measuring Learned Equivariance
von: Gruver, Nate, et al.
Veröffentlicht: (2022)
von: Gruver, Nate, et al.
Veröffentlicht: (2022)
Can AI Assistants Know What They Don't Know?
von: Cheng, Qinyuan, et al.
Veröffentlicht: (2024)
von: Cheng, Qinyuan, et al.
Veröffentlicht: (2024)
Imagining What We Don't Know
von: Samuels, Lisa
Veröffentlicht: (2026)
von: Samuels, Lisa
Veröffentlicht: (2026)
Imagining What We Don't Know
von: Samuels, Lisa
Veröffentlicht: (2026)
von: Samuels, Lisa
Veröffentlicht: (2026)
World Models That Know When They Don't Know - Controllable Video Generation with Calibrated Uncertainty
von: Mei, Zhiting, et al.
Veröffentlicht: (2025)
von: Mei, Zhiting, et al.
Veröffentlicht: (2025)
When Should We Orchestrate Multiple Agents?
von: Bhatt, Umang, et al.
Veröffentlicht: (2025)
von: Bhatt, Umang, et al.
Veröffentlicht: (2025)
Visually Dehallucinative Instruction Generation: Know What You Don't Know
von: Cha, Sungguk, et al.
Veröffentlicht: (2024)
von: Cha, Sungguk, et al.
Veröffentlicht: (2024)
What We Know and What We Don't Know About the Function of γδ T Cells
von: Immo Prinz, et al.
Veröffentlicht: (2025)
von: Immo Prinz, et al.
Veröffentlicht: (2025)
Experts Don't Cheat: Learning What You Don't Know By Predicting Pairs
von: Johnson, Daniel D., et al.
Veröffentlicht: (2024)
von: Johnson, Daniel D., et al.
Veröffentlicht: (2024)
Bayesian Mixture-of-Experts: Towards Making LLMs Know What They Don't Know
von: Li, Albus Yizhuo
Veröffentlicht: (2025)
von: Li, Albus Yizhuo
Veröffentlicht: (2025)
Know What You Don't Know: Uncertainty Calibration of Process Reward Models
von: Park, Young-Jin, et al.
Veröffentlicht: (2025)
von: Park, Young-Jin, et al.
Veröffentlicht: (2025)
Know What You Don't Know: Selective Prediction for Early Exit DNNs
von: Bajpai, Divya Jyoti, et al.
Veröffentlicht: (2025)
von: Bajpai, Divya Jyoti, et al.
Veröffentlicht: (2025)
Pulsed‐Field Ablation: What We Know and What We Don't
von: Sanghamitra Mohanty, et al.
Veröffentlicht: (2025)
von: Sanghamitra Mohanty, et al.
Veröffentlicht: (2025)
Do We Know What LLMs Don't Know? A Study of Consistency in Knowledge Probing
von: Zhao, Raoyuan, et al.
Veröffentlicht: (2025)
von: Zhao, Raoyuan, et al.
Veröffentlicht: (2025)
EpiCaR: Knowing What You Don't Know Matters for Better Reasoning in LLMs
von: Yeom, Jewon, et al.
Veröffentlicht: (2026)
von: Yeom, Jewon, et al.
Veröffentlicht: (2026)
What Librarians Still Don't Know about Free Software
von: Chudnov, Daniel
Veröffentlicht: (2009)
von: Chudnov, Daniel
Veröffentlicht: (2009)
What Models Know, How Well They Know It: Knowledge-Weighted Fine-Tuning for Learning When to Say "I Don't Know"
von: Lee, Joosung, et al.
Veröffentlicht: (2026)
von: Lee, Joosung, et al.
Veröffentlicht: (2026)
Do Retrieval Augmented Language Models Know When They Don't Know?
von: Zhou, Youchao, et al.
Veröffentlicht: (2025)
von: Zhou, Youchao, et al.
Veröffentlicht: (2025)
Can Molecular Foundation Models Know What They Don't Know? A Simple Remedy with Preference Optimization
von: He, Langzhou, et al.
Veröffentlicht: (2025)
von: He, Langzhou, et al.
Veröffentlicht: (2025)
vTune: Verifiable Fine-Tuning for LLMs Through Backdooring
von: Zhang, Eva, et al.
Veröffentlicht: (2024)
von: Zhang, Eva, et al.
Veröffentlicht: (2024)
Dynamic Delayed Tree Expansion For Improved Multi-Path Speculative Decoding
von: Thomas, Rahul, et al.
Veröffentlicht: (2026)
von: Thomas, Rahul, et al.
Veröffentlicht: (2026)
Coding Agents Don't Know When to Act
von: Gloaguen, Thibaud, et al.
Veröffentlicht: (2026)
von: Gloaguen, Thibaud, et al.
Veröffentlicht: (2026)
I Know What I Don't Know: Latent Posterior Factor Models for Multi-Evidence Probabilistic Reasoning
von: Alege, Aliyu Agboola
Veröffentlicht: (2026)
von: Alege, Aliyu Agboola
Veröffentlicht: (2026)
Reasoning about Uncertainty: Do Reasoning Models Know When They Don't Know?
von: Mei, Zhiting, et al.
Veröffentlicht: (2025)
von: Mei, Zhiting, et al.
Veröffentlicht: (2025)
Smaug: Fixing Failure Modes of Preference Optimisation with DPO-Positive
von: Pal, Arka, et al.
Veröffentlicht: (2024)
von: Pal, Arka, et al.
Veröffentlicht: (2024)
Why Masking Diffusion Works: Condition on the Jump Schedule for Improved Discrete Diffusion
von: Amin, Alan N., et al.
Veröffentlicht: (2025)
von: Amin, Alan N., et al.
Veröffentlicht: (2025)
When You Don't Know the Answer, Say So
Veröffentlicht: (2024)
Veröffentlicht: (2024)
Ventricular Tachycardia Ablation in LVAD Recipients: What We Know, What We Still Don't
von: Claire E. Brady, et al.
Veröffentlicht: (2025)
von: Claire E. Brady, et al.
Veröffentlicht: (2025)
Do LLMs Really Know What They Don't Know? Internal States Mainly Reflect Knowledge Recall Rather Than Truthfulness
von: Cheang, Chi Seng, et al.
Veröffentlicht: (2025)
von: Cheang, Chi Seng, et al.
Veröffentlicht: (2025)
Physicists Don't Know What They're Talking About When They Say 'Order'
von: Arafat Gaspar Jiménez Gaistardo
Veröffentlicht: (2025)
von: Arafat Gaspar Jiménez Gaistardo
Veröffentlicht: (2025)
Bonds on the Ballot: What Voters (Don't) Know About Debt Financing and Why It Matters
von: Shanna Pearson‐Merkowitz, et al.
Veröffentlicht: (2026)
von: Shanna Pearson‐Merkowitz, et al.
Veröffentlicht: (2026)
What Language Models Know But Don't Say: Non-Generative Prior Extraction for Generalization
von: Rezaeimanesh, Sara, et al.
Veröffentlicht: (2026)
von: Rezaeimanesh, Sara, et al.
Veröffentlicht: (2026)
Fine-Tuned LLMs Know They Don't Know: A Parameter-Efficient Approach to Recovering Honesty
von: Shi, Zeyu, et al.
Veröffentlicht: (2025)
von: Shi, Zeyu, et al.
Veröffentlicht: (2025)
Towards Agents That Know When They Don't Know: Uncertainty as a Control Signal for Structured Reasoning
von: Stoisser, Josefa Lia, et al.
Veröffentlicht: (2025)
von: Stoisser, Josefa Lia, et al.
Veröffentlicht: (2025)
Privacy-Preserving Mechanisms Enable Cheap Verifiable Inference of LLMs
von: Pal, Arka, et al.
Veröffentlicht: (2026)
von: Pal, Arka, et al.
Veröffentlicht: (2026)
Large Language Models Are Zero-Shot Time Series Forecasters
von: Gruver, Nate, et al.
Veröffentlicht: (2023)
von: Gruver, Nate, et al.
Veröffentlicht: (2023)
I Don't Know: Explicit Modeling of Uncertainty with an [IDK] Token
von: Cohen, Roi, et al.
Veröffentlicht: (2024)
von: Cohen, Roi, et al.
Veröffentlicht: (2024)
First, Learn What You Don't Know: Active Information Gathering for Driving at the Limits of Handling
von: Davydov, Alexander, et al.
Veröffentlicht: (2024)
von: Davydov, Alexander, et al.
Veröffentlicht: (2024)
Do Large Language Models Know What They Don't Know? Kalshibench: A New Benchmark for Evaluating Epistemic Calibration via Prediction Markets
von: Nel, Lukas
Veröffentlicht: (2025)
von: Nel, Lukas
Veröffentlicht: (2025)
Ähnliche Einträge
-
Knowing What You Know Is Not Enough: Large Language Model Confidences Don't Align With Their Actions
von: Pal, Arka, et al.
Veröffentlicht: (2025) -
The Lie Derivative for Measuring Learned Equivariance
von: Gruver, Nate, et al.
Veröffentlicht: (2022) -
Can AI Assistants Know What They Don't Know?
von: Cheng, Qinyuan, et al.
Veröffentlicht: (2024) -
Imagining What We Don't Know
von: Samuels, Lisa
Veröffentlicht: (2026) -
Imagining What We Don't Know
von: Samuels, Lisa
Veröffentlicht: (2026)