What Models Know, How Well They Know It: Knowledge-Weighted Fine-Tuning for Learning When to Say "I Don't Know"
Fuente:
arXiv
Salvato in:
| Autori principali: | Lee, Joosung, Jo, Hwiyeol, Ko, Donghyeon, Chae, Kyubyung, Park, Cheonbok, Kim, Jeonghoon |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Enhancing Hallucination Detection via Future Context
di: Lee, Joosung, et al.
Pubblicazione: (2025)
di: Lee, Joosung, et al.
Pubblicazione: (2025)
When You Don't Know the Answer, Say So
Pubblicazione: (2024)
Pubblicazione: (2024)
Visually Dehallucinative Instruction Generation: Know What You Don't Know
di: Cha, Sungguk, et al.
Pubblicazione: (2024)
di: Cha, Sungguk, et al.
Pubblicazione: (2024)
Can AI Assistants Know What They Don't Know?
di: Cheng, Qinyuan, et al.
Pubblicazione: (2024)
di: Cheng, Qinyuan, et al.
Pubblicazione: (2024)
R-Tuning: Instructing Large Language Models to Say `I Don't Know'
di: Zhang, Hanning, et al.
Pubblicazione: (2023)
di: Zhang, Hanning, et al.
Pubblicazione: (2023)
Physicists Don't Know What They're Talking About When They Say 'Order'
di: Arafat Gaspar Jiménez Gaistardo
Pubblicazione: (2025)
di: Arafat Gaspar Jiménez Gaistardo
Pubblicazione: (2025)
EpiCaR: Knowing What You Don't Know Matters for Better Reasoning in LLMs
di: Yeom, Jewon, et al.
Pubblicazione: (2026)
di: Yeom, Jewon, et al.
Pubblicazione: (2026)
Know What You Don't Know: Uncertainty Calibration of Process Reward Models
di: Park, Young-Jin, et al.
Pubblicazione: (2025)
di: Park, Young-Jin, et al.
Pubblicazione: (2025)
Fine-Tuning Language Models to Know What They Know
di: Park, Sangjun, et al.
Pubblicazione: (2026)
di: Park, Sangjun, et al.
Pubblicazione: (2026)
Do Retrieval Augmented Language Models Know When They Don't Know?
di: Zhou, Youchao, et al.
Pubblicazione: (2025)
di: Zhou, Youchao, et al.
Pubblicazione: (2025)
Imagining What We Don't Know
di: Samuels, Lisa
Pubblicazione: (2026)
di: Samuels, Lisa
Pubblicazione: (2026)
Imagining What We Don't Know
di: Samuels, Lisa
Pubblicazione: (2026)
di: Samuels, Lisa
Pubblicazione: (2026)
Fine-Tuned LLMs Know They Don't Know: A Parameter-Efficient Approach to Recovering Honesty
di: Shi, Zeyu, et al.
Pubblicazione: (2025)
di: Shi, Zeyu, et al.
Pubblicazione: (2025)
Honest AI: Fine-Tuning "Small" Language Models to Say "I Don't Know", and Reducing Hallucination in RAG
di: Chen, Xinxi, et al.
Pubblicazione: (2024)
di: Chen, Xinxi, et al.
Pubblicazione: (2024)
Large Language Models Must Be Taught to Know What They Don't Know
di: Kapoor, Sanyam, et al.
Pubblicazione: (2024)
di: Kapoor, Sanyam, et al.
Pubblicazione: (2024)
When Robots Should Say "I Don't Know": Benchmarking Abstention in Embodied Question Answering
di: Wu, Tao, et al.
Pubblicazione: (2025)
di: Wu, Tao, et al.
Pubblicazione: (2025)
Do We Know What LLMs Don't Know? A Study of Consistency in Knowledge Probing
di: Zhao, Raoyuan, et al.
Pubblicazione: (2025)
di: Zhao, Raoyuan, et al.
Pubblicazione: (2025)
Reasoning about Uncertainty: Do Reasoning Models Know When They Don't Know?
di: Mei, Zhiting, et al.
Pubblicazione: (2025)
di: Mei, Zhiting, et al.
Pubblicazione: (2025)
Coding Agents Don't Know When to Act
di: Gloaguen, Thibaud, et al.
Pubblicazione: (2026)
di: Gloaguen, Thibaud, et al.
Pubblicazione: (2026)
Decomposed Prompting Does Not Fix Knowledge Gaps, But Helps Models Say "I Don't Know"
di: Madhwal, Dhruv, et al.
Pubblicazione: (2026)
di: Madhwal, Dhruv, et al.
Pubblicazione: (2026)
What We Know and What We Don't Know About the Function of γδ T Cells
di: Immo Prinz, et al.
Pubblicazione: (2025)
di: Immo Prinz, et al.
Pubblicazione: (2025)
Bayesian Mixture-of-Experts: Towards Making LLMs Know What They Don't Know
di: Li, Albus Yizhuo
Pubblicazione: (2025)
di: Li, Albus Yizhuo
Pubblicazione: (2025)
Know What You Don't Know: Selective Prediction for Early Exit DNNs
di: Bajpai, Divya Jyoti, et al.
Pubblicazione: (2025)
di: Bajpai, Divya Jyoti, et al.
Pubblicazione: (2025)
What Language Models Know But Don't Say: Non-Generative Prior Extraction for Generalization
di: Rezaeimanesh, Sara, et al.
Pubblicazione: (2026)
di: Rezaeimanesh, Sara, et al.
Pubblicazione: (2026)
World Models That Know When They Don't Know - Controllable Video Generation with Calibrated Uncertainty
di: Mei, Zhiting, et al.
Pubblicazione: (2025)
di: Mei, Zhiting, et al.
Pubblicazione: (2025)
I Know What I Don't Know: Latent Posterior Factor Models for Multi-Evidence Probabilistic Reasoning
di: Alege, Aliyu Agboola
Pubblicazione: (2026)
di: Alege, Aliyu Agboola
Pubblicazione: (2026)
How to Computerize Your Serials and Periodicals When You Don't Know How
di: Matthews, Mary, et al.
Pubblicazione: (1970)
di: Matthews, Mary, et al.
Pubblicazione: (1970)
Towards Agents That Know When They Don't Know: Uncertainty as a Control Signal for Structured Reasoning
di: Stoisser, Josefa Lia, et al.
Pubblicazione: (2025)
di: Stoisser, Josefa Lia, et al.
Pubblicazione: (2025)
Experts Don't Cheat: Learning What You Don't Know By Predicting Pairs
di: Johnson, Daniel D., et al.
Pubblicazione: (2024)
di: Johnson, Daniel D., et al.
Pubblicazione: (2024)
Knowing What You Know Is Not Enough: Large Language Model Confidences Don't Align With Their Actions
di: Pal, Arka, et al.
Pubblicazione: (2025)
di: Pal, Arka, et al.
Pubblicazione: (2025)
Can Molecular Foundation Models Know What They Don't Know? A Simple Remedy with Preference Optimization
di: He, Langzhou, et al.
Pubblicazione: (2025)
di: He, Langzhou, et al.
Pubblicazione: (2025)
Do LLMs Really Know What They Don't Know? Internal States Mainly Reflect Knowledge Recall Rather Than Truthfulness
di: Cheang, Chi Seng, et al.
Pubblicazione: (2025)
di: Cheang, Chi Seng, et al.
Pubblicazione: (2025)
What Librarians Still Don't Know about Free Software
di: Chudnov, Daniel
Pubblicazione: (2009)
di: Chudnov, Daniel
Pubblicazione: (2009)
Pulsed‐Field Ablation: What We Know and What We Don't
di: Sanghamitra Mohanty, et al.
Pubblicazione: (2025)
di: Sanghamitra Mohanty, et al.
Pubblicazione: (2025)
Knowing You Don't Know: Learning When to Continue Search in Multi-round RAG through Self-Practicing
di: Yang, Diji, et al.
Pubblicazione: (2025)
di: Yang, Diji, et al.
Pubblicazione: (2025)
"Knowing When You Don't Know": A Multilingual Relevance Assessment Dataset for Robust Retrieval-Augmented Generation
di: Thakur, Nandan, et al.
Pubblicazione: (2023)
di: Thakur, Nandan, et al.
Pubblicazione: (2023)
Large Language Models Know What To Say But Not When To Speak
di: Umair, Muhammad, et al.
Pubblicazione: (2024)
di: Umair, Muhammad, et al.
Pubblicazione: (2024)
Cross-lingual Collapse: How Language-Centric Foundation Models Shape Reasoning in Large Language Models
di: Park, Cheonbok, et al.
Pubblicazione: (2025)
di: Park, Cheonbok, et al.
Pubblicazione: (2025)
I Don't Know: Explicit Modeling of Uncertainty with an [IDK] Token
di: Cohen, Roi, et al.
Pubblicazione: (2024)
di: Cohen, Roi, et al.
Pubblicazione: (2024)
NanoKnow: How to Know What Your Language Model Knows
di: Gu, Lingwei, et al.
Pubblicazione: (2026)
di: Gu, Lingwei, et al.
Pubblicazione: (2026)
Documenti analoghi
-
Enhancing Hallucination Detection via Future Context
di: Lee, Joosung, et al.
Pubblicazione: (2025) -
When You Don't Know the Answer, Say So
Pubblicazione: (2024) -
Visually Dehallucinative Instruction Generation: Know What You Don't Know
di: Cha, Sungguk, et al.
Pubblicazione: (2024) -
Can AI Assistants Know What They Don't Know?
di: Cheng, Qinyuan, et al.
Pubblicazione: (2024) -
R-Tuning: Instructing Large Language Models to Say `I Don't Know'
di: Zhang, Hanning, et al.
Pubblicazione: (2023)