Fine-Grained Detection of Context-Grounded Hallucinations Using LLMs
Fuente:
arXiv
Salvato in:
| Autori principali: | Peisakhovsky, Yehonatan, Gekhman, Zorik, Mass, Yosi, Ein-Dor, Liat, Reichart, Roi |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Multi-Domain Explainability of Preferences
di: Calderon, Nitay, et al.
Pubblicazione: (2025)
di: Calderon, Nitay, et al.
Pubblicazione: (2025)
Does Fine-Tuning LLMs on New Knowledge Encourage Hallucinations?
di: Gekhman, Zorik, et al.
Pubblicazione: (2024)
di: Gekhman, Zorik, et al.
Pubblicazione: (2024)
Can LLMs Learn Macroeconomic Narratives from Social Media?
di: Gueta, Almog, et al.
Pubblicazione: (2024)
di: Gueta, Almog, et al.
Pubblicazione: (2024)
Thinking to Recall: How Reasoning Unlocks Parametric Knowledge in LLMs
di: Gekhman, Zorik, et al.
Pubblicazione: (2026)
di: Gekhman, Zorik, et al.
Pubblicazione: (2026)
LLMs Know More Than They Show: On the Intrinsic Representation of LLM Hallucinations
di: Orgad, Hadas, et al.
Pubblicazione: (2024)
di: Orgad, Hadas, et al.
Pubblicazione: (2024)
NL-Eye: Abductive NLI for Images
di: Ventura, Mor, et al.
Pubblicazione: (2024)
di: Ventura, Mor, et al.
Pubblicazione: (2024)
Inside-Out: Hidden Factual Knowledge in LLMs
di: Gekhman, Zorik, et al.
Pubblicazione: (2025)
di: Gekhman, Zorik, et al.
Pubblicazione: (2025)
HACK: Hallucinations Along Certainty and Knowledge Axes
di: Simhi, Adi, et al.
Pubblicazione: (2025)
di: Simhi, Adi, et al.
Pubblicazione: (2025)
Measuring the Robustness of NLP Models to Domain Shifts
di: Calderon, Nitay, et al.
Pubblicazione: (2023)
di: Calderon, Nitay, et al.
Pubblicazione: (2023)
Why Fine-Tuning Encourages Hallucinations and How to Fix It
di: Kaplan, Guy, et al.
Pubblicazione: (2026)
di: Kaplan, Guy, et al.
Pubblicazione: (2026)
WildIFEval: Instruction Following in the Wild
di: Lior, Gili, et al.
Pubblicazione: (2025)
di: Lior, Gili, et al.
Pubblicazione: (2025)
Masked by Consensus: Disentangling Privileged Knowledge in LLM Correctness
di: Ashuach, Tomer, et al.
Pubblicazione: (2026)
di: Ashuach, Tomer, et al.
Pubblicazione: (2026)
Confidence Improves Self-Consistency in LLMs
di: Taubenfeld, Amir, et al.
Pubblicazione: (2025)
di: Taubenfeld, Amir, et al.
Pubblicazione: (2025)
On Behalf of the Stakeholders: Trends in NLP Model Interpretability in the Era of LLMs
di: Calderon, Nitay, et al.
Pubblicazione: (2024)
di: Calderon, Nitay, et al.
Pubblicazione: (2024)
Empty Shelves or Lost Keys? Recall Is the Bottleneck for Parametric Factuality
di: Calderon, Nitay, et al.
Pubblicazione: (2026)
di: Calderon, Nitay, et al.
Pubblicazione: (2026)
Label-Efficient Model Selection for Text Generation
di: Ashury-Tahan, Shir, et al.
Pubblicazione: (2024)
di: Ashury-Tahan, Shir, et al.
Pubblicazione: (2024)
Are LLMs Better than Reported? Detecting Label Errors and Mitigating Their Effect on Model Performance
di: Nahum, Omer, et al.
Pubblicazione: (2024)
di: Nahum, Omer, et al.
Pubblicazione: (2024)
Leveraging NTPs for Efficient Hallucination Detection in VLMs
di: Azachi, Ofir, et al.
Pubblicazione: (2025)
di: Azachi, Ofir, et al.
Pubblicazione: (2025)
The Alternative Annotator Test for LLM-as-a-Judge: How to Statistically Justify Replacing Human Annotators with LLMs
di: Calderon, Nitay, et al.
Pubblicazione: (2025)
di: Calderon, Nitay, et al.
Pubblicazione: (2025)
Evaluating Alignment of Behavioral Dispositions in LLMs
di: Taubenfeld, Amir, et al.
Pubblicazione: (2026)
di: Taubenfeld, Amir, et al.
Pubblicazione: (2026)
Conversational Prompt Engineering
di: Ein-Dor, Liat, et al.
Pubblicazione: (2024)
di: Ein-Dor, Liat, et al.
Pubblicazione: (2024)
A Systematic Review of NLP for Dementia -- Tasks, Datasets and Opportunities
di: Peled-Cohen, Lotem, et al.
Pubblicazione: (2024)
di: Peled-Cohen, Lotem, et al.
Pubblicazione: (2024)
Multilingual Fine-Grained News Headline Hallucination Detection
di: Shen, Jiaming, et al.
Pubblicazione: (2024)
di: Shen, Jiaming, et al.
Pubblicazione: (2024)
Will it Merge? On The Causes of Model Mergeability
di: Rahamim, Adir, et al.
Pubblicazione: (2026)
di: Rahamim, Adir, et al.
Pubblicazione: (2026)
Genie: Achieving Human Parity in Content-Grounded Datasets Generation
di: Yehudai, Asaf, et al.
Pubblicazione: (2024)
di: Yehudai, Asaf, et al.
Pubblicazione: (2024)
LIBERTy: A Causal Framework for Benchmarking Concept-Based Explanations of LLMs with Structural Counterfactuals
di: Toker, Gilat, et al.
Pubblicazione: (2026)
di: Toker, Gilat, et al.
Pubblicazione: (2026)
Cross-Layer Attention Probing for Fine-Grained Hallucination Detection
di: Suresh, Malavika, et al.
Pubblicazione: (2025)
di: Suresh, Malavika, et al.
Pubblicazione: (2025)
TabSTAR: A Tabular Foundation Model for Tabular Data with Text Fields
di: Arazi, Alan, et al.
Pubblicazione: (2025)
di: Arazi, Alan, et al.
Pubblicazione: (2025)
Can LLMs Replace Economic Choice Prediction Labs? The Case of Language-based Persuasion Games
di: Shapira, Eilam, et al.
Pubblicazione: (2024)
di: Shapira, Eilam, et al.
Pubblicazione: (2024)
Donors and Recipients: On Asymmetric Transfer Across Tasks and Languages with Parameter-Efficient Fine-Tuning
di: Dymkiewicz, Kajetan, et al.
Pubblicazione: (2025)
di: Dymkiewicz, Kajetan, et al.
Pubblicazione: (2025)
Motivation in Large Language Models
di: Nahum, Omer, et al.
Pubblicazione: (2026)
di: Nahum, Omer, et al.
Pubblicazione: (2026)
Detecting Stylistic Fingerprints of Large Language Models
di: Bitton, Yehonatan, et al.
Pubblicazione: (2025)
di: Bitton, Yehonatan, et al.
Pubblicazione: (2025)
The Poisoned Apple Effect: Strategic Manipulation of Mediated Markets via Technology Expansion of AI Agents
di: Shapira, Eilam, et al.
Pubblicazione: (2026)
di: Shapira, Eilam, et al.
Pubblicazione: (2026)
Efficient Benchmarking of Language Models
di: Perlitz, Yotam, et al.
Pubblicazione: (2023)
di: Perlitz, Yotam, et al.
Pubblicazione: (2023)
The Colorful Future of LLMs: Evaluating and Improving LLMs as Emotional Supporters for Queer Youth
di: Lissak, Shir, et al.
Pubblicazione: (2024)
di: Lissak, Shir, et al.
Pubblicazione: (2024)
Stay Tuned: An Empirical Study of the Impact of Hyperparameters on LLM Tuning in Real-World Applications
di: Halfon, Alon, et al.
Pubblicazione: (2024)
di: Halfon, Alon, et al.
Pubblicazione: (2024)
Seeing is Believing: Rich-Context Hallucination Detection for MLLMs via Backward Visual Grounding
di: Guo, Pinxue, et al.
Pubblicazione: (2025)
di: Guo, Pinxue, et al.
Pubblicazione: (2025)
Predicting Decisions of AI Agents from Limited Interaction through Text-Tabular Modeling
di: Shapira, Eilam, et al.
Pubblicazione: (2026)
di: Shapira, Eilam, et al.
Pubblicazione: (2026)
Systematic Biases in LLM Simulations of Debates
di: Taubenfeld, Amir, et al.
Pubblicazione: (2024)
di: Taubenfeld, Amir, et al.
Pubblicazione: (2024)
Med-HEAL: Analyzing and Mitigating Hallucinations in Medical LLMs with Hallucination-Aware In-Context Learning
di: Liao, Yiming, et al.
Pubblicazione: (2026)
di: Liao, Yiming, et al.
Pubblicazione: (2026)
Documenti analoghi
-
Multi-Domain Explainability of Preferences
di: Calderon, Nitay, et al.
Pubblicazione: (2025) -
Does Fine-Tuning LLMs on New Knowledge Encourage Hallucinations?
di: Gekhman, Zorik, et al.
Pubblicazione: (2024) -
Can LLMs Learn Macroeconomic Narratives from Social Media?
di: Gueta, Almog, et al.
Pubblicazione: (2024) -
Thinking to Recall: How Reasoning Unlocks Parametric Knowledge in LLMs
di: Gekhman, Zorik, et al.
Pubblicazione: (2026) -
LLMs Know More Than They Show: On the Intrinsic Representation of LLM Hallucinations
di: Orgad, Hadas, et al.
Pubblicazione: (2024)