ConSens: Assessing context grounding in open-book question answering
Fuente:
arXiv
Saved in:
| Main Authors: | Vankov, Ivan, Ivanov, Matyo, Correia, Adriana, Botev, Victor |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
The illusion of a perfect metric: Why evaluating AI's words is harder than it looks
by: Oliva, Maria Paz, et al.
Published: (2025)
by: Oliva, Maria Paz, et al.
Published: (2025)
Agribot: agriculture-specific question answer system
by: Jain, Naman, et al.
Published: (2025)
by: Jain, Naman, et al.
Published: (2025)
Evaluating Embedding Frameworks for Scientific Domain
by: Ahmed, Nouman, et al.
Published: (2025)
by: Ahmed, Nouman, et al.
Published: (2025)
ChatGPT for automated grading of short answer questions in mechanical ventilation
by: Jade, Tejas, et al.
Published: (2025)
by: Jade, Tejas, et al.
Published: (2025)
Why does in-context learning fail sometimes? Evaluating in-context learning on open and closed questions
by: Li, Xiang, et al.
Published: (2024)
by: Li, Xiang, et al.
Published: (2024)
Multi-step retrieval and reasoning improves radiology question answering with large language models
by: Wind, Sebastian, et al.
Published: (2025)
by: Wind, Sebastian, et al.
Published: (2025)
What's the plan? Metrics for implicit planning in LLMs and their application to rhyme generation and question answering
by: Maar, Jim, et al.
Published: (2026)
by: Maar, Jim, et al.
Published: (2026)
CaLMQA: Exploring culturally specific long-form question answering across 23 languages
by: Arora, Shane, et al.
Published: (2024)
by: Arora, Shane, et al.
Published: (2024)
Retrieval Enhanced Feedback via In-context Neural Error-book
by: Hyun, Jongyeop, et al.
Published: (2025)
by: Hyun, Jongyeop, et al.
Published: (2025)
Extracting memorized pieces of (copyrighted) books from open-weight language models
by: Cooper, A. Feder, et al.
Published: (2025)
by: Cooper, A. Feder, et al.
Published: (2025)
TANQ: An open domain dataset of table answered questions
by: Akhtar, Mubashara, et al.
Published: (2024)
by: Akhtar, Mubashara, et al.
Published: (2024)
FusionMind -- Improving question and answering with external context fusion
by: Verma, Shreyas, et al.
Published: (2023)
by: Verma, Shreyas, et al.
Published: (2023)
When an LLM is apprehensive about its answers -- and when its uncertainty is justified
by: Sychev, Petr, et al.
Published: (2025)
by: Sychev, Petr, et al.
Published: (2025)
Question answering system of bridge design specification based on large language model
by: Zhang, Leye, et al.
Published: (2024)
by: Zhang, Leye, et al.
Published: (2024)
ConFu: Contemplate the Future for Better Speculative Sampling
by: Qin, Zongyue, et al.
Published: (2026)
by: Qin, Zongyue, et al.
Published: (2026)
SenTSR-Bench: Thinking with Injected Knowledge for Time-Series Reasoning
by: He, Zelin, et al.
Published: (2026)
by: He, Zelin, et al.
Published: (2026)
Navigating the Knowledge Sea: Planet-scale answer retrieval using LLMs
by: Sarkar, Dipankar
Published: (2024)
by: Sarkar, Dipankar
Published: (2024)
Asking an AI for salary negotiation advice is a matter of concern: Controlled experimental perturbation of ChatGPT for protected and non-protected group discrimination on a contextual task with no clear ground truth answers
by: Geiger, R. Stuart, et al.
Published: (2024)
by: Geiger, R. Stuart, et al.
Published: (2024)
Extracting books from production language models
by: Ahmed, Ahmed, et al.
Published: (2026)
by: Ahmed, Ahmed, et al.
Published: (2026)
The broader spectrum of in-context learning
by: Lampinen, Andrew Kyle, et al.
Published: (2024)
by: Lampinen, Andrew Kyle, et al.
Published: (2024)
Transformers are Universal In-context Learners
by: Furuya, Takashi, et al.
Published: (2024)
by: Furuya, Takashi, et al.
Published: (2024)
The benefits of query-based KGQA systems for complex and temporal questions in LLM era
by: Alekseev, Artem, et al.
Published: (2025)
by: Alekseev, Artem, et al.
Published: (2025)
CausalLM is not optimal for in-context learning
by: Ding, Nan, et al.
Published: (2023)
by: Ding, Nan, et al.
Published: (2023)
In-context Learning in Presence of Spurious Correlations
by: Harutyunyan, Hrayr, et al.
Published: (2024)
by: Harutyunyan, Hrayr, et al.
Published: (2024)
Guideline Learning for In-context Information Extraction
by: Pang, Chaoxu, et al.
Published: (2023)
by: Pang, Chaoxu, et al.
Published: (2023)
In-context Learning and Gradient Descent Revisited
by: Deutch, Gilad, et al.
Published: (2023)
by: Deutch, Gilad, et al.
Published: (2023)
Re-examining learning linear functions in context
by: Naim, Omar, et al.
Published: (2024)
by: Naim, Omar, et al.
Published: (2024)
If generative AI is the answer, what is the question?
by: Tewari, Ambuj
Published: (2025)
by: Tewari, Ambuj
Published: (2025)
Few shot chain-of-thought driven reasoning to prompt LLMs for open ended medical question answering
by: Nachane, Saeel Sandeep, et al.
Published: (2024)
by: Nachane, Saeel Sandeep, et al.
Published: (2024)
Efficacy of Synthetic Data as a Benchmark
by: Maheshwari, Gaurav, et al.
Published: (2024)
by: Maheshwari, Gaurav, et al.
Published: (2024)
Sparse Attention across Multiple-context KV Cache
by: Cao, Ziyi, et al.
Published: (2025)
by: Cao, Ziyi, et al.
Published: (2025)
Learning without training: The implicit dynamics of in-context learning
by: Dherin, Benoit, et al.
Published: (2025)
by: Dherin, Benoit, et al.
Published: (2025)
Long-context Reference-based MT Quality Estimation
by: Haq, Sami Ul, et al.
Published: (2025)
by: Haq, Sami Ul, et al.
Published: (2025)
RNNs are not Transformers (Yet): The Key Bottleneck on In-context Retrieval
by: Wen, Kaiyue, et al.
Published: (2024)
by: Wen, Kaiyue, et al.
Published: (2024)
Linking In-context Learning in Transformers to Human Episodic Memory
by: Ji-An, Li, et al.
Published: (2024)
by: Ji-An, Li, et al.
Published: (2024)
Text clustering applied to data augmentation in legal contexts
by: Freitas, Lucas José Gonçalves, et al.
Published: (2024)
by: Freitas, Lucas José Gonçalves, et al.
Published: (2024)
What is Wrong with Perplexity for Long-context Language Modeling?
by: Fang, Lizhe, et al.
Published: (2024)
by: Fang, Lizhe, et al.
Published: (2024)
Curse of High Dimensionality Issue in Transformer for Long-context Modeling
by: Zhang, Shuhai, et al.
Published: (2025)
by: Zhang, Shuhai, et al.
Published: (2025)
Diversity Enhances an LLM's Performance in RAG and Long-context Task
by: Wang, Zhichao, et al.
Published: (2025)
by: Wang, Zhichao, et al.
Published: (2025)
Graph-based Molecular In-context Learning Grounded on Morgan Fingerprints
by: Al-Lawati, Ali, et al.
Published: (2025)
by: Al-Lawati, Ali, et al.
Published: (2025)
Similar Items
-
The illusion of a perfect metric: Why evaluating AI's words is harder than it looks
by: Oliva, Maria Paz, et al.
Published: (2025) -
Agribot: agriculture-specific question answer system
by: Jain, Naman, et al.
Published: (2025) -
Evaluating Embedding Frameworks for Scientific Domain
by: Ahmed, Nouman, et al.
Published: (2025) -
ChatGPT for automated grading of short answer questions in mechanical ventilation
by: Jade, Tejas, et al.
Published: (2025) -
Why does in-context learning fail sometimes? Evaluating in-context learning on open and closed questions
by: Li, Xiang, et al.
Published: (2024)