Do LLMs Understand Ambiguity in Text? A Case Study in Open-world Question Answering
Fuente:
arXiv
Guardado en:
| Autores principales: | Keluskar, Aryan, Bhattacharjee, Amrita, Liu, Huan |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Mind the Ambiguity: Aleatoric Uncertainty Quantification in LLMs for Safe Medical Question Answering
por: Liu, Yaokun, et al.
Publicado: (2026)
por: Liu, Yaokun, et al.
Publicado: (2026)
Zero-shot LLM-guided Counterfactual Generation: A Case Study on NLP Model Evaluation
por: Bhattacharjee, Amrita, et al.
Publicado: (2024)
por: Bhattacharjee, Amrita, et al.
Publicado: (2024)
EAGLE: A Domain Generalization Framework for AI-generated Text Detection
por: Bhattacharjee, Amrita, et al.
Publicado: (2024)
por: Bhattacharjee, Amrita, et al.
Publicado: (2024)
Towards LLM-guided Causal Explainability for Black-box Text Classifiers
por: Bhattacharjee, Amrita, et al.
Publicado: (2023)
por: Bhattacharjee, Amrita, et al.
Publicado: (2023)
Evaluating Adaptive Personalization of Educational Readings with Simulated Learners
por: Woo, Ryan T., et al.
Publicado: (2026)
por: Woo, Ryan T., et al.
Publicado: (2026)
Adversarial Text Purification: A Large Language Model Approach for Defense
por: Moraffah, Raha, et al.
Publicado: (2024)
por: Moraffah, Raha, et al.
Publicado: (2024)
QPaug: Question and Passage Augmentation for Open-Domain Question Answering of LLMs
por: Kim, Minsang, et al.
Publicado: (2024)
por: Kim, Minsang, et al.
Publicado: (2024)
A$^2$Search: Ambiguity-Aware Question Answering with Reinforcement Learning
por: Zhang, Fengji, et al.
Publicado: (2025)
por: Zhang, Fengji, et al.
Publicado: (2025)
Denoising Table-Text Retrieval for Open-Domain Question Answering
por: Kang, Deokhyung, et al.
Publicado: (2024)
por: Kang, Deokhyung, et al.
Publicado: (2024)
Hybrid Graphs for Table-and-Text based Question Answering using LLMs
por: Agarwal, Ankush, et al.
Publicado: (2025)
por: Agarwal, Ankush, et al.
Publicado: (2025)
Towards Interpretable Hate Speech Detection using Large Language Model-extracted Rationales
por: Nirmal, Ayushi, et al.
Publicado: (2024)
por: Nirmal, Ayushi, et al.
Publicado: (2024)
Sparse Neurons Carry Strong Signals of Question Ambiguity in LLMs
por: Zhang, Zhuoxuan, et al.
Publicado: (2025)
por: Zhang, Zhuoxuan, et al.
Publicado: (2025)
Open Domain Question Answering with Conflicting Contexts
por: Liu, Siyi, et al.
Publicado: (2024)
por: Liu, Siyi, et al.
Publicado: (2024)
HPE:Answering Complex Questions over Text by Hybrid Question Parsing and Execution
por: Liu, Ye, et al.
Publicado: (2023)
por: Liu, Ye, et al.
Publicado: (2023)
The Essence of Contextual Understanding in Theory of Mind: A Study on Question Answering with Story Characters
por: Zhou, Chulun, et al.
Publicado: (2025)
por: Zhou, Chulun, et al.
Publicado: (2025)
Understanding Network Behaviors through Natural Language Question-Answering
por: Xing, Mingzhe, et al.
Publicado: (2025)
por: Xing, Mingzhe, et al.
Publicado: (2025)
CBR-RAG: Case-Based Reasoning for Retrieval Augmented Generation in LLMs for Legal Question Answering
por: Wiratunga, Nirmalie, et al.
Publicado: (2024)
por: Wiratunga, Nirmalie, et al.
Publicado: (2024)
Automatic Evaluation of Healthcare LLMs Beyond Question-Answering
por: Arias-Duart, Anna, et al.
Publicado: (2025)
por: Arias-Duart, Anna, et al.
Publicado: (2025)
Harnessing Artificial Intelligence to Combat Online Hate: Exploring the Challenges and Opportunities of Large Language Models in Hate Speech Detection
por: Kumarage, Tharindu, et al.
Publicado: (2024)
por: Kumarage, Tharindu, et al.
Publicado: (2024)
A Dataset of Open-Domain Question Answering with Multiple-Span Answers
por: Luo, Zhiyi, et al.
Publicado: (2024)
por: Luo, Zhiyi, et al.
Publicado: (2024)
Fine-Tuning LLMs for Reliable Medical Question-Answering Services
por: Anaissi, Ali, et al.
Publicado: (2024)
por: Anaissi, Ali, et al.
Publicado: (2024)
Knowledgeable Preference Alignment for LLMs in Domain-specific Question Answering
por: Zhang, Yichi, et al.
Publicado: (2023)
por: Zhang, Yichi, et al.
Publicado: (2023)
Text to Query Plans for Question Answering on Large Tables
por: Zhang, Yipeng, et al.
Publicado: (2025)
por: Zhang, Yipeng, et al.
Publicado: (2025)
BanglaQuAD: A Bengali Open-domain Question Answering Dataset
por: Rony, Md Rashad Al Hasan, et al.
Publicado: (2024)
por: Rony, Md Rashad Al Hasan, et al.
Publicado: (2024)
O$^2$-Searcher: A Searching-based Agent Model for Open-Domain Open-Ended Question Answering
por: Mei, Jianbiao, et al.
Publicado: (2025)
por: Mei, Jianbiao, et al.
Publicado: (2025)
Enhancing Large Language Models with Pseudo- and Multisource- Knowledge Graphs for Open-ended Question Answering
por: Liu, Jiaxiang, et al.
Publicado: (2024)
por: Liu, Jiaxiang, et al.
Publicado: (2024)
FinTextQA: A Dataset for Long-form Financial Question Answering
por: Chen, Jian, et al.
Publicado: (2024)
por: Chen, Jian, et al.
Publicado: (2024)
60 Data Points are Sufficient to Fine-Tune LLMs for Question-Answering
por: Ye, Junjie, et al.
Publicado: (2024)
por: Ye, Junjie, et al.
Publicado: (2024)
Augmenting Black-box LLMs with Medical Textbooks for Biomedical Question Answering
por: Wang, Yubo, et al.
Publicado: (2023)
por: Wang, Yubo, et al.
Publicado: (2023)
Question Answering on Patient Medical Records with Private Fine-Tuned LLMs
por: Kothari, Sara, et al.
Publicado: (2025)
por: Kothari, Sara, et al.
Publicado: (2025)
Training LLMs with Reinforcement Learning for Intent-Aware Personalized Question Answering
por: Amirizaniani, Maryam, et al.
Publicado: (2026)
por: Amirizaniani, Maryam, et al.
Publicado: (2026)
MedEthicsQA: A Comprehensive Question Answering Benchmark for Medical Ethics Evaluation of LLMs
por: Wei, Jianhui, et al.
Publicado: (2025)
por: Wei, Jianhui, et al.
Publicado: (2025)
To Generate or to Retrieve? On the Effectiveness of Artificial Contexts for Medical Open-Domain Question Answering
por: Frisoni, Giacomo, et al.
Publicado: (2024)
por: Frisoni, Giacomo, et al.
Publicado: (2024)
Towards Better Generalization in Open-Domain Question Answering by Mitigating Context Memorization
por: Zhang, Zixuan, et al.
Publicado: (2024)
por: Zhang, Zixuan, et al.
Publicado: (2024)
Question Answering with LLMs and Learning from Answer Sets
por: Borroto, Manuel, et al.
Publicado: (2025)
por: Borroto, Manuel, et al.
Publicado: (2025)
I've got the "Answer"! Interpretation of LLMs Hidden States in Question Answering
por: Goloviznina, Valeriya, et al.
Publicado: (2024)
por: Goloviznina, Valeriya, et al.
Publicado: (2024)
Towards Inference-time Category-wise Safety Steering for Large Language Models
por: Bhattacharjee, Amrita, et al.
Publicado: (2024)
por: Bhattacharjee, Amrita, et al.
Publicado: (2024)
Generative Data Augmentation using LLMs improves Distributional Robustness in Question Answering
por: Chowdhury, Arijit Ghosh, et al.
Publicado: (2023)
por: Chowdhury, Arijit Ghosh, et al.
Publicado: (2023)
Conv-CoA: Improving Open-domain Question Answering in Large Language Models via Conversational Chain-of-Action
por: Pan, Zhenyu, et al.
Publicado: (2024)
por: Pan, Zhenyu, et al.
Publicado: (2024)
RepLiQA: A Question-Answering Dataset for Benchmarking LLMs on Unseen Reference Content
por: Monteiro, Joao, et al.
Publicado: (2024)
por: Monteiro, Joao, et al.
Publicado: (2024)
Ejemplares similares
-
Mind the Ambiguity: Aleatoric Uncertainty Quantification in LLMs for Safe Medical Question Answering
por: Liu, Yaokun, et al.
Publicado: (2026) -
Zero-shot LLM-guided Counterfactual Generation: A Case Study on NLP Model Evaluation
por: Bhattacharjee, Amrita, et al.
Publicado: (2024) -
EAGLE: A Domain Generalization Framework for AI-generated Text Detection
por: Bhattacharjee, Amrita, et al.
Publicado: (2024) -
Towards LLM-guided Causal Explainability for Black-box Text Classifiers
por: Bhattacharjee, Amrita, et al.
Publicado: (2023) -
Evaluating Adaptive Personalization of Educational Readings with Simulated Learners
por: Woo, Ryan T., et al.
Publicado: (2026)