Contextual Candor: Enhancing LLM Trustworthiness Through Hierarchical Unanswerability Detection
Fuente:
arXiv
Salvato in:
| Autori principali: | Robinson, Steven, Rivera, Antonio Carlos |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Harnessing RLHF for Robust Unanswerability Recognition and Trustworthy Response Generation in LLMs
di: Lin, Shuyuan, et al.
Pubblicazione: (2025)
di: Lin, Shuyuan, et al.
Pubblicazione: (2025)
TUBench: Benchmarking Large Vision-Language Models on Trustworthiness with Unanswerable Questions
di: He, Xingwei, et al.
Pubblicazione: (2024)
di: He, Xingwei, et al.
Pubblicazione: (2024)
Coal Mining Question Answering with LLMs
di: Rivera, Antonio Carlos, et al.
Pubblicazione: (2024)
di: Rivera, Antonio Carlos, et al.
Pubblicazione: (2024)
Towards Unbiased Evaluation of Detecting Unanswerable Questions in EHRSQL
di: Yang, Yongjin, et al.
Pubblicazione: (2024)
di: Yang, Yongjin, et al.
Pubblicazione: (2024)
Unanswerability Evaluation for Retrieval Augmented Generation
di: Peng, Xiangyu, et al.
Pubblicazione: (2024)
di: Peng, Xiangyu, et al.
Pubblicazione: (2024)
Query Carefully: Detecting the Unanswerables in Text-to-SQL Tasks
di: Saxer, Jasmin, et al.
Pubblicazione: (2025)
di: Saxer, Jasmin, et al.
Pubblicazione: (2025)
FactGuard: Leveraging Multi-Agent Systems to Generate Answerable and Unanswerable Questions for Enhanced Long-Context LLM Extraction
di: Zhang, Qian-Wen, et al.
Pubblicazione: (2025)
di: Zhang, Qian-Wen, et al.
Pubblicazione: (2025)
Drawing the Line: Enhancing Trustworthiness of MLLMs Through the Power of Refusal
di: Wang, Yuhao, et al.
Pubblicazione: (2024)
di: Wang, Yuhao, et al.
Pubblicazione: (2024)
UAQFact: Evaluating Factual Knowledge Utilization of LLMs on Unanswerable Questions
di: Tan, Chuanyuan, et al.
Pubblicazione: (2025)
di: Tan, Chuanyuan, et al.
Pubblicazione: (2025)
Towards Reliable and Factual Response Generation: Detecting Unanswerable Questions in Information-Seeking Conversations
di: Łajewska, Weronika, et al.
Pubblicazione: (2024)
di: Łajewska, Weronika, et al.
Pubblicazione: (2024)
I Could've Asked That: Reformulating Unanswerable Questions
di: Zhao, Wenting, et al.
Pubblicazione: (2024)
di: Zhao, Wenting, et al.
Pubblicazione: (2024)
Benchmarking Hallucination in Large Language Models based on Unanswerable Math Word Problem
di: Sun, Yuhong, et al.
Pubblicazione: (2024)
di: Sun, Yuhong, et al.
Pubblicazione: (2024)
TreeCut: A Synthetic Unanswerable Math Word Problem Dataset for LLM Hallucination Evaluation
di: Ouyang, Jialin
Pubblicazione: (2025)
di: Ouyang, Jialin
Pubblicazione: (2025)
TrustLLM: Trustworthiness in Large Language Models
di: Huang, Yue, et al.
Pubblicazione: (2024)
di: Huang, Yue, et al.
Pubblicazione: (2024)
Learning to Contextualize Web Pages for Enhanced Decision Making by LLM Agents
di: Lee, Dongjun, et al.
Pubblicazione: (2025)
di: Lee, Dongjun, et al.
Pubblicazione: (2025)
Trustworthy Reasoning: Evaluating and Enhancing Factual Accuracy in LLM Intermediate Thought Processes
di: Jiao, Rui, et al.
Pubblicazione: (2025)
di: Jiao, Rui, et al.
Pubblicazione: (2025)
Iterative Repair with Weak Verifiers for Few-shot Transfer in KBQA with Unanswerability
di: Sawhney, Riya, et al.
Pubblicazione: (2024)
di: Sawhney, Riya, et al.
Pubblicazione: (2024)
Suicidal Comment Tree Dataset: Enhancing Risk Assessment and Prediction Through Contextual Analysis
di: Li, Jun, et al.
Pubblicazione: (2025)
di: Li, Jun, et al.
Pubblicazione: (2025)
CATCH: A Controllable Theme Detection Framework with Contextualized Clustering and Hierarchical Generation
di: Ke, Rui, et al.
Pubblicazione: (2025)
di: Ke, Rui, et al.
Pubblicazione: (2025)
PRACTIQ: A Practical Conversational Text-to-SQL dataset with Ambiguous and Unanswerable Queries
di: Dong, Mingwen, et al.
Pubblicazione: (2024)
di: Dong, Mingwen, et al.
Pubblicazione: (2024)
RetinaQA: A Robust Knowledge Base Question Answering Model for both Answerable and Unanswerable Questions
di: Faldu, Prayushi, et al.
Pubblicazione: (2024)
di: Faldu, Prayushi, et al.
Pubblicazione: (2024)
CLARITY: A Framework and Benchmark for Conversational Language Ambiguity and Unanswerability in Interactive NL2SQL Systems
di: Sarwar, Tabinda, et al.
Pubblicazione: (2026)
di: Sarwar, Tabinda, et al.
Pubblicazione: (2026)
Establishing Trustworthy LLM Evaluation via Shortcut Neuron Analysis
di: Zhu, Kejian, et al.
Pubblicazione: (2025)
di: Zhu, Kejian, et al.
Pubblicazione: (2025)
TrustScore: Reference-Free Evaluation of LLM Response Trustworthiness
di: Zheng, Danna, et al.
Pubblicazione: (2024)
di: Zheng, Danna, et al.
Pubblicazione: (2024)
PEACH: Pretrained-embedding Explanation Across Contextual and Hierarchical Structure
di: Cao, Feiqi, et al.
Pubblicazione: (2024)
di: Cao, Feiqi, et al.
Pubblicazione: (2024)
Answering the Unanswerable Is to Err Knowingly: Analyzing and Mitigating Abstention Failures in Large Reasoning Models
di: Liu, Yi, et al.
Pubblicazione: (2025)
di: Liu, Yi, et al.
Pubblicazione: (2025)
Investigating Retrieval-Augmented Generation Systems on Unanswerable, Uncheatable, Realistic, Multi-hop Queries
di: Liu, Gabrielle Kaili-May, et al.
Pubblicazione: (2025)
di: Liu, Gabrielle Kaili-May, et al.
Pubblicazione: (2025)
Hierarchical Chain-of-Thought Prompting: Enhancing LLM Reasoning Performance and Efficiency
di: Huang, Xingshuai, et al.
Pubblicazione: (2026)
di: Huang, Xingshuai, et al.
Pubblicazione: (2026)
TrustRAG: Enhancing Robustness and Trustworthiness in Retrieval-Augmented Generation
di: Zhou, Huichi, et al.
Pubblicazione: (2025)
di: Zhou, Huichi, et al.
Pubblicazione: (2025)
Trustworthy AI for Medicine: Continuous Hallucination Detection and Elimination with CHECK
di: Garcia-Fernandez, Carlos, et al.
Pubblicazione: (2025)
di: Garcia-Fernandez, Carlos, et al.
Pubblicazione: (2025)
When Not to Answer: Evaluating Prompts on GPT Models for Effective Abstention in Unanswerable Math Word Problems
di: Saadat, Asir, et al.
Pubblicazione: (2024)
di: Saadat, Asir, et al.
Pubblicazione: (2024)
FIRST: Teach A Reliable Large Language Model Through Efficient Trustworthy Distillation
di: Shum, KaShun, et al.
Pubblicazione: (2024)
di: Shum, KaShun, et al.
Pubblicazione: (2024)
Towards Trustworthy Multimodal Moderation via Policy-Aligned Reasoning and Hierarchical Labeling
di: Li, Anqi, et al.
Pubblicazione: (2025)
di: Li, Anqi, et al.
Pubblicazione: (2025)
Hierarchical Contextual Manifold Alignment for Structuring Latent Representations in Large Language Models
di: Dong, Meiquan, et al.
Pubblicazione: (2025)
di: Dong, Meiquan, et al.
Pubblicazione: (2025)
Bridging Context Gaps: Enhancing Comprehension in Long-Form Social Conversations Through Contextualized Excerpts
di: Mohanty, Shrestha, et al.
Pubblicazione: (2024)
di: Mohanty, Shrestha, et al.
Pubblicazione: (2024)
ClickGuard: A Trustworthy Adaptive Fusion Framework for Clickbait Detection
di: Dhiman, Chhavi, et al.
Pubblicazione: (2026)
di: Dhiman, Chhavi, et al.
Pubblicazione: (2026)
Advancing AI Trustworthiness Through Patient Simulation: Risk Assessment of Conversational Agents for Antidepressant Selection
di: Shawon, Md Tanvir Rouf, et al.
Pubblicazione: (2026)
di: Shawon, Md Tanvir Rouf, et al.
Pubblicazione: (2026)
CiteLLM: An Agentic Platform for Trustworthy Scientific Reference Discovery
di: Hong, Mengze, et al.
Pubblicazione: (2026)
di: Hong, Mengze, et al.
Pubblicazione: (2026)
Enhancing Talent Employment Insights Through Feature Extraction with LLM Finetuning
di: Thakrar, Karishma, et al.
Pubblicazione: (2025)
di: Thakrar, Karishma, et al.
Pubblicazione: (2025)
JT-Safe: Intrinsically Enhancing the Safety and Trustworthiness of LLMs
di: Feng, Junlan, et al.
Pubblicazione: (2025)
di: Feng, Junlan, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Harnessing RLHF for Robust Unanswerability Recognition and Trustworthy Response Generation in LLMs
di: Lin, Shuyuan, et al.
Pubblicazione: (2025) -
TUBench: Benchmarking Large Vision-Language Models on Trustworthiness with Unanswerable Questions
di: He, Xingwei, et al.
Pubblicazione: (2024) -
Coal Mining Question Answering with LLMs
di: Rivera, Antonio Carlos, et al.
Pubblicazione: (2024) -
Towards Unbiased Evaluation of Detecting Unanswerable Questions in EHRSQL
di: Yang, Yongjin, et al.
Pubblicazione: (2024) -
Unanswerability Evaluation for Retrieval Augmented Generation
di: Peng, Xiangyu, et al.
Pubblicazione: (2024)