Ask a Local: Detecting Hallucinations With Specialized Model Divergence
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Creo, Aldan, Cerezo-Costas, Héctor, Alonso-Doval, Pedro, Hormazábal-Lagos, Maximiliano |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
MRT at IberLEF-2025 PRESTA Task: Maximizing Recovery from Tables with Multiple Steps
von: Lagos, Maximiliano Hormazábal, et al.
Veröffentlicht: (2025)
von: Lagos, Maximiliano Hormazábal, et al.
Veröffentlicht: (2025)
ExpliCIT-QA: Explainable Code-Based Image Table Question Answering
von: Lagos, Maximiliano Hormazábal, et al.
Veröffentlicht: (2025)
von: Lagos, Maximiliano Hormazábal, et al.
Veröffentlicht: (2025)
Spatially Grounded Explanations in Vision Language Models for Document Visual Question Answering
von: Lagos, Maximiliano Hormazábal, et al.
Veröffentlicht: (2025)
von: Lagos, Maximiliano Hormazábal, et al.
Veröffentlicht: (2025)
MRT at SemEval-2025 Task 8: Maximizing Recovery from Tables with Multiple Steps
von: Lagos, Maximiliano Hormazábal, et al.
Veröffentlicht: (2025)
von: Lagos, Maximiliano Hormazábal, et al.
Veröffentlicht: (2025)
Complete Evasion, Zero Modification: PDF Attacks on AI Text Detection
von: Creo, Aldan
Veröffentlicht: (2025)
von: Creo, Aldan
Veröffentlicht: (2025)
SilverSpeak: Evading AI-Generated Text Detectors using Homoglyphs
von: Creo, Aldan, et al.
Veröffentlicht: (2024)
von: Creo, Aldan, et al.
Veröffentlicht: (2024)
Mass-Scale Analysis of In-the-Wild Conversations Reveals Complexity Bounds on LLM Jailbreaking
von: Creo, Aldan, et al.
Veröffentlicht: (2025)
von: Creo, Aldan, et al.
Veröffentlicht: (2025)
Hallucination Detection in LLMs with Topological Divergence on Attention Graphs
von: Bazarova, Alexandra, et al.
Veröffentlicht: (2025)
von: Bazarova, Alexandra, et al.
Veröffentlicht: (2025)
DAHRS: Divergence-Aware Hallucination-Remediated SRL Projection
von: Youm, Sangpil, et al.
Veröffentlicht: (2024)
von: Youm, Sangpil, et al.
Veröffentlicht: (2024)
Prompt-Response Semantic Divergence Metrics for Faithfulness Hallucination and Misalignment Detection in Large Language Models
von: Halperin, Igor
Veröffentlicht: (2025)
von: Halperin, Igor
Veröffentlicht: (2025)
Hallucination Detection and Hallucination Mitigation: An Investigation
von: Luo, Junliang, et al.
Veröffentlicht: (2024)
von: Luo, Junliang, et al.
Veröffentlicht: (2024)
Hallucination Detection with Small Language Models
von: Cheung, Ming
Veröffentlicht: (2025)
von: Cheung, Ming
Veröffentlicht: (2025)
Ask Good Questions for Large Language Models
von: Wu, Qi, et al.
Veröffentlicht: (2025)
von: Wu, Qi, et al.
Veröffentlicht: (2025)
IntelliAsk: Learning to Ask High-Quality Research Questions via RLVR
von: Sharma, Karun, et al.
Veröffentlicht: (2026)
von: Sharma, Karun, et al.
Veröffentlicht: (2026)
Enhancing Uncertainty Modeling with Semantic Graph for Hallucination Detection
von: Chen, Kedi, et al.
Veröffentlicht: (2025)
von: Chen, Kedi, et al.
Veröffentlicht: (2025)
Can LLMs Ask Good Questions?
von: Zhang, Yueheng, et al.
Veröffentlicht: (2025)
von: Zhang, Yueheng, et al.
Veröffentlicht: (2025)
Turk-LettuceDetect: A Hallucination Detection Models for Turkish RAG Applications
von: Taş, Selva, et al.
Veröffentlicht: (2025)
von: Taş, Selva, et al.
Veröffentlicht: (2025)
Neural Probe-Based Hallucination Detection for Large Language Models
von: Liang, Shize, et al.
Veröffentlicht: (2025)
von: Liang, Shize, et al.
Veröffentlicht: (2025)
Counterfactual Probing for Hallucination Detection and Mitigation in Large Language Models
von: Feng, Yijun
Veröffentlicht: (2025)
von: Feng, Yijun
Veröffentlicht: (2025)
The Energy of Falsehood: Detecting Hallucinations via Diffusion Model Likelihoods
von: Gautam, Arpit Singh, et al.
Veröffentlicht: (2026)
von: Gautam, Arpit Singh, et al.
Veröffentlicht: (2026)
Reference-free Hallucination Detection for Large Vision-Language Models
von: Li, Qing, et al.
Veröffentlicht: (2024)
von: Li, Qing, et al.
Veröffentlicht: (2024)
STaR-GATE: Teaching Language Models to Ask Clarifying Questions
von: Andukuri, Chinmaya, et al.
Veröffentlicht: (2024)
von: Andukuri, Chinmaya, et al.
Veröffentlicht: (2024)
Hallucination Detection with the Internal Layers of LLMs
von: Preiß, Martin
Veröffentlicht: (2025)
von: Preiß, Martin
Veröffentlicht: (2025)
Towards Long Context Hallucination Detection
von: Liu, Siyi, et al.
Veröffentlicht: (2025)
von: Liu, Siyi, et al.
Veröffentlicht: (2025)
HIDE and Seek: Detecting Hallucinations in Language Models via Decoupled Representations
von: Chatterjee, Anwoy, et al.
Veröffentlicht: (2025)
von: Chatterjee, Anwoy, et al.
Veröffentlicht: (2025)
Hallucination Detection: Robustly Discerning Reliable Answers in Large Language Models
von: Chen, Yuyan, et al.
Veröffentlicht: (2024)
von: Chen, Yuyan, et al.
Veröffentlicht: (2024)
Alleviating Hallucinations of Large Language Models through Induced Hallucinations
von: Zhang, Yue, et al.
Veröffentlicht: (2023)
von: Zhang, Yue, et al.
Veröffentlicht: (2023)
LettuceDetect: A Hallucination Detection Framework for RAG Applications
von: Kovács, Ádám, et al.
Veröffentlicht: (2025)
von: Kovács, Ádám, et al.
Veröffentlicht: (2025)
Enhancing Hallucination Detection via Future Context
von: Lee, Joosung, et al.
Veröffentlicht: (2025)
von: Lee, Joosung, et al.
Veröffentlicht: (2025)
Unsupervised Hallucination Detection by Inspecting Reasoning Processes
von: Srey, Ponhvoan, et al.
Veröffentlicht: (2025)
von: Srey, Ponhvoan, et al.
Veröffentlicht: (2025)
FaithLens: Detecting and Explaining Faithfulness Hallucination
von: Si, Shuzheng, et al.
Veröffentlicht: (2025)
von: Si, Shuzheng, et al.
Veröffentlicht: (2025)
Comparing Hallucination Detection Metrics for Multilingual Generation
von: Kang, Haoqiang, et al.
Veröffentlicht: (2024)
von: Kang, Haoqiang, et al.
Veröffentlicht: (2024)
On Early Detection of Hallucinations in Factual Question Answering
von: Snyder, Ben, et al.
Veröffentlicht: (2023)
von: Snyder, Ben, et al.
Veröffentlicht: (2023)
Sanity Checks for Long-Form Hallucination Detection
von: Zollicoffer, Geigh, et al.
Veröffentlicht: (2026)
von: Zollicoffer, Geigh, et al.
Veröffentlicht: (2026)
GLSim: Detecting Object Hallucinations in LVLMs via Global-Local Similarity
von: Park, Seongheon, et al.
Veröffentlicht: (2025)
von: Park, Seongheon, et al.
Veröffentlicht: (2025)
From Out-of-Distribution Detection to Hallucination Detection: A Geometric View
von: Liu, Litian, et al.
Veröffentlicht: (2026)
von: Liu, Litian, et al.
Veröffentlicht: (2026)
Measuring the Impact of Lexical Training Data Coverage on Hallucination Detection in Large Language Models
von: Zhang, Shuo, et al.
Veröffentlicht: (2025)
von: Zhang, Shuo, et al.
Veröffentlicht: (2025)
Attention-guided Self-reflection for Zero-shot Hallucination Detection in Large Language Models
von: Liu, Qiang, et al.
Veröffentlicht: (2025)
von: Liu, Qiang, et al.
Veröffentlicht: (2025)
HalluSAE: Detecting Hallucinations in Large Language Models via Sparse Auto-Encoders
von: Chen, Boshui, et al.
Veröffentlicht: (2026)
von: Chen, Boshui, et al.
Veröffentlicht: (2026)
Unsupervised Real-Time Hallucination Detection based on the Internal States of Large Language Models
von: Su, Weihang, et al.
Veröffentlicht: (2024)
von: Su, Weihang, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
MRT at IberLEF-2025 PRESTA Task: Maximizing Recovery from Tables with Multiple Steps
von: Lagos, Maximiliano Hormazábal, et al.
Veröffentlicht: (2025) -
ExpliCIT-QA: Explainable Code-Based Image Table Question Answering
von: Lagos, Maximiliano Hormazábal, et al.
Veröffentlicht: (2025) -
Spatially Grounded Explanations in Vision Language Models for Document Visual Question Answering
von: Lagos, Maximiliano Hormazábal, et al.
Veröffentlicht: (2025) -
MRT at SemEval-2025 Task 8: Maximizing Recovery from Tables with Multiple Steps
von: Lagos, Maximiliano Hormazábal, et al.
Veröffentlicht: (2025) -
Complete Evasion, Zero Modification: PDF Attacks on AI Text Detection
von: Creo, Aldan
Veröffentlicht: (2025)