Do LLMs Know When to NOT Answer? Investigating Abstention Abilities of Large Language Models
Fuente:
arXiv
Salvato in:
| Autori principali: | Madhusudhan, Nishanth, Madhusudhan, Sathwik Tejaswi, Yadav, Vikas, Hashemi, Masoud |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Knowing When Not to Answer: Evaluating Abstention in Multimodal Reasoning Systems
di: Madhusudhan, Nishanth, et al.
Pubblicazione: (2026)
di: Madhusudhan, Nishanth, et al.
Pubblicazione: (2026)
Revitalizing Saturated Benchmarks: A Weighted Metric Approach for Differentiating Large Language Model Performance
di: Etzine, Bryan, et al.
Pubblicazione: (2025)
di: Etzine, Bryan, et al.
Pubblicazione: (2025)
Auto-Cypher: Improving LLMs on Cypher generation via LLM-supervised generation-verification framework
di: Tiwari, Aman, et al.
Pubblicazione: (2024)
di: Tiwari, Aman, et al.
Pubblicazione: (2024)
M2Lingual: Enhancing Multilingual, Multi-Turn Instruction Alignment in Large Language Models
di: Maheshwary, Rishabh, et al.
Pubblicazione: (2024)
di: Maheshwary, Rishabh, et al.
Pubblicazione: (2024)
Curry-DPO: Enhancing Alignment using Curriculum Learning & Ranked Preferences
di: Pattnaik, Pulkit, et al.
Pubblicazione: (2024)
di: Pattnaik, Pulkit, et al.
Pubblicazione: (2024)
Augmenting LLM Reasoning with Dynamic Notes Writing for Complex QA
di: Maheshwary, Rishabh, et al.
Pubblicazione: (2025)
di: Maheshwary, Rishabh, et al.
Pubblicazione: (2025)
DNR Bench: Benchmarking Over-Reasoning in Reasoning LLMs
di: Hashemi, Masoud, et al.
Pubblicazione: (2025)
di: Hashemi, Masoud, et al.
Pubblicazione: (2025)
Layer-Wise Quantization: A Pragmatic and Effective Method for Quantizing LLMs Beyond Integer Bit-Levels
di: Dumitru, Razvan-Gabriel, et al.
Pubblicazione: (2024)
di: Dumitru, Razvan-Gabriel, et al.
Pubblicazione: (2024)
Grammar Search for Multi-Agent Systems
di: Singh, Mayank, et al.
Pubblicazione: (2025)
di: Singh, Mayank, et al.
Pubblicazione: (2025)
DeepSRGM -- Sequence Classification and Ranking in Indian Classical Music with Deep Learning
di: Madhusudhan, Sathwik Tejaswi, et al.
Pubblicazione: (2024)
di: Madhusudhan, Sathwik Tejaswi, et al.
Pubblicazione: (2024)
Knowing When Not to Answer: Abstention-Aware Scientific Reasoning
di: Abdaljalil, Samir, et al.
Pubblicazione: (2026)
di: Abdaljalil, Samir, et al.
Pubblicazione: (2026)
BigCharts-R1: Enhanced Chart Reasoning with Visual Reinforcement Finetuning
di: Masry, Ahmed, et al.
Pubblicazione: (2025)
di: Masry, Ahmed, et al.
Pubblicazione: (2025)
Know Your Limits: A Survey of Abstention in Large Language Models
di: Wen, Bingbing, et al.
Pubblicazione: (2024)
di: Wen, Bingbing, et al.
Pubblicazione: (2024)
Cats Confuse Reasoning LLM: Query Agnostic Adversarial Triggers for Reasoning Models
di: Rajeev, Meghana, et al.
Pubblicazione: (2025)
di: Rajeev, Meghana, et al.
Pubblicazione: (2025)
Seeing Isn't Knowing: Do VLMs Know When Not to Answer Spatial Questions (and Why)?
di: Zhang, Yue, et al.
Pubblicazione: (2026)
di: Zhang, Yue, et al.
Pubblicazione: (2026)
Explicit Diversity Conditions for Effective Question Answer Generation with Large Language Models
di: Yadav, Vikas, et al.
Pubblicazione: (2024)
di: Yadav, Vikas, et al.
Pubblicazione: (2024)
Prompting with Phonemes: Enhancing LLMs' Multilinguality for Non-Latin Script Languages
di: Nguyen, Hoang H, et al.
Pubblicazione: (2024)
di: Nguyen, Hoang H, et al.
Pubblicazione: (2024)
When Not to Answer: Evaluating Prompts on GPT Models for Effective Abstention in Unanswerable Math Word Problems
di: Saadat, Asir, et al.
Pubblicazione: (2024)
di: Saadat, Asir, et al.
Pubblicazione: (2024)
Do Retrieval Augmented Language Models Know When They Don't Know?
di: Zhou, Youchao, et al.
Pubblicazione: (2025)
di: Zhou, Youchao, et al.
Pubblicazione: (2025)
ColMate: Contrastive Late Interaction and Masked Text for Multimodal Document Retrieval
di: Masry, Ahmed, et al.
Pubblicazione: (2025)
di: Masry, Ahmed, et al.
Pubblicazione: (2025)
Second Guess: Detecting Uncertainty Through Abstention and Answer Stability in Small Language Models
di: Aravindan, Ashwath Vaithinathan, et al.
Pubblicazione: (2026)
di: Aravindan, Ashwath Vaithinathan, et al.
Pubblicazione: (2026)
Answering the Unanswerable Is to Err Knowingly: Analyzing and Mitigating Abstention Failures in Large Reasoning Models
di: Liu, Yi, et al.
Pubblicazione: (2025)
di: Liu, Yi, et al.
Pubblicazione: (2025)
Do Large Language Models Know Conflict? Investigating Parametric vs. Non-Parametric Knowledge of LLMs for Conflict Forecasting
di: Nemkova, Apollinaire Poli, et al.
Pubblicazione: (2025)
di: Nemkova, Apollinaire Poli, et al.
Pubblicazione: (2025)
When to Speak, When to Abstain: Contrastive Decoding with Abstention
di: Kim, Hyuhng Joon, et al.
Pubblicazione: (2024)
di: Kim, Hyuhng Joon, et al.
Pubblicazione: (2024)
Large Language Models Know What To Say But Not When To Speak
di: Umair, Muhammad, et al.
Pubblicazione: (2024)
di: Umair, Muhammad, et al.
Pubblicazione: (2024)
Do Language Models Know When They're Hallucinating References?
di: Agrawal, Ayush, et al.
Pubblicazione: (2023)
di: Agrawal, Ayush, et al.
Pubblicazione: (2023)
Do Large Language Models Have Compositional Ability? An Investigation into Limitations and Scalability
di: Xu, Zhuoyan, et al.
Pubblicazione: (2024)
di: Xu, Zhuoyan, et al.
Pubblicazione: (2024)
Do Large Language Models Know How Much They Know?
di: Prato, Gabriele, et al.
Pubblicazione: (2025)
di: Prato, Gabriele, et al.
Pubblicazione: (2025)
Diffusion Language Models Know the Answer Before Decoding
di: Li, Pengxiang, et al.
Pubblicazione: (2025)
di: Li, Pengxiang, et al.
Pubblicazione: (2025)
Do LLMs Know about Hallucination? An Empirical Investigation of LLM's Hidden States
di: Duan, Hanyu, et al.
Pubblicazione: (2024)
di: Duan, Hanyu, et al.
Pubblicazione: (2024)
KnowGuard: Knowledge-Driven Abstention for Multi-Round Clinical Reasoning
di: Dang, Xilin, et al.
Pubblicazione: (2025)
di: Dang, Xilin, et al.
Pubblicazione: (2025)
Do Large Language Models Know What They Are Capable Of?
di: Barkan, Casey O., et al.
Pubblicazione: (2025)
di: Barkan, Casey O., et al.
Pubblicazione: (2025)
Reasoning about Uncertainty: Do Reasoning Models Know When They Don't Know?
di: Mei, Zhiting, et al.
Pubblicazione: (2025)
di: Mei, Zhiting, et al.
Pubblicazione: (2025)
Hallucination as Commitment Failure: Larger LLMs Misfire Despite Knowing the Answer
di: Yeom, Jewon, et al.
Pubblicazione: (2026)
di: Yeom, Jewon, et al.
Pubblicazione: (2026)
AprielGuard
di: Kasundra, Jaykumar, et al.
Pubblicazione: (2025)
di: Kasundra, Jaykumar, et al.
Pubblicazione: (2025)
Knowing When to Ask -- Bridging Large Language Models and Data
di: Radhakrishnan, Prashanth, et al.
Pubblicazione: (2024)
di: Radhakrishnan, Prashanth, et al.
Pubblicazione: (2024)
Large Language Models Often Know When They Are Being Evaluated
di: Needham, Joe, et al.
Pubblicazione: (2025)
di: Needham, Joe, et al.
Pubblicazione: (2025)
Do Language Models Know Theo Has a Wife? Investigating the Proviso Problem
di: Azin, Tara, et al.
Pubblicazione: (2026)
di: Azin, Tara, et al.
Pubblicazione: (2026)
DataAgent: Evaluating Large Language Models' Ability to Answer Zero-Shot, Natural Language Queries
di: Mishra, Manit, et al.
Pubblicazione: (2024)
di: Mishra, Manit, et al.
Pubblicazione: (2024)
Do LLMs Know What Is Private Internally? Probing and Steering Contextual Privacy Norms in Large Language Model Representations
di: Wang, Haoran, et al.
Pubblicazione: (2026)
di: Wang, Haoran, et al.
Pubblicazione: (2026)
Documenti analoghi
-
Knowing When Not to Answer: Evaluating Abstention in Multimodal Reasoning Systems
di: Madhusudhan, Nishanth, et al.
Pubblicazione: (2026) -
Revitalizing Saturated Benchmarks: A Weighted Metric Approach for Differentiating Large Language Model Performance
di: Etzine, Bryan, et al.
Pubblicazione: (2025) -
Auto-Cypher: Improving LLMs on Cypher generation via LLM-supervised generation-verification framework
di: Tiwari, Aman, et al.
Pubblicazione: (2024) -
M2Lingual: Enhancing Multilingual, Multi-Turn Instruction Alignment in Large Language Models
di: Maheshwary, Rishabh, et al.
Pubblicazione: (2024) -
Curry-DPO: Enhancing Alignment using Curriculum Learning & Ranked Preferences
di: Pattnaik, Pulkit, et al.
Pubblicazione: (2024)