Learned Hallucination Detection in Black-Box LLMs using Token-level Entropy Production Rate
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Moslonka, Charles, Randrianarivo, Hicham, Garnier, Arthur, Malherbe, Emmanuel |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
FactSelfCheck: Fact-Level Black-Box Hallucination Detection for LLMs
von: Sawczyn, Albert, et al.
Veröffentlicht: (2025)
von: Sawczyn, Albert, et al.
Veröffentlicht: (2025)
Black-Box Hallucination Detection via Consistency Under the Uncertain Expression
von: Joo, Seongho, et al.
Veröffentlicht: (2025)
von: Joo, Seongho, et al.
Veröffentlicht: (2025)
Verify when Uncertain: Beyond Self-Consistency in Black Box Hallucination Detection
von: Xue, Yihao, et al.
Veröffentlicht: (2025)
von: Xue, Yihao, et al.
Veröffentlicht: (2025)
Semantic Entropy Probes: Robust and Cheap Hallucination Detection in LLMs
von: Kossen, Jannik, et al.
Veröffentlicht: (2024)
von: Kossen, Jannik, et al.
Veröffentlicht: (2024)
TECP: Token-Entropy Conformal Prediction for LLMs
von: Xu, Beining, et al.
Veröffentlicht: (2025)
von: Xu, Beining, et al.
Veröffentlicht: (2025)
Matryoshka Pilot: Learning to Drive Black-Box LLMs with LLMs
von: Li, Changhao, et al.
Veröffentlicht: (2024)
von: Li, Changhao, et al.
Veröffentlicht: (2024)
Text Corpora as Concept Fields: Black-Box Hallucination and Novelty Measurement
von: Kersting, Nicholas S., et al.
Veröffentlicht: (2026)
von: Kersting, Nicholas S., et al.
Veröffentlicht: (2026)
A Survey of Calibration Process for Black-Box LLMs
von: Xie, Liangru, et al.
Veröffentlicht: (2024)
von: Xie, Liangru, et al.
Veröffentlicht: (2024)
Hallucination Detection with the Internal Layers of LLMs
von: Preiß, Martin
Veröffentlicht: (2025)
von: Preiß, Martin
Veröffentlicht: (2025)
TPA: Next Token Probability Attribution for Detecting Hallucinations in RAG
von: Lu, Pengqian, et al.
Veröffentlicht: (2025)
von: Lu, Pengqian, et al.
Veröffentlicht: (2025)
The First Token Knows: Single-Decode Confidence for Hallucination Detection
von: Gabriel, Mina
Veröffentlicht: (2026)
von: Gabriel, Mina
Veröffentlicht: (2026)
AggTruth: Contextual Hallucination Detection using Aggregated Attention Scores in LLMs
von: Matys, Piotr, et al.
Veröffentlicht: (2025)
von: Matys, Piotr, et al.
Veröffentlicht: (2025)
ALTER: Asymmetric LoRA for Token-Entropy-Guided Unlearning of LLMs
von: Chen, Xunlei, et al.
Veröffentlicht: (2026)
von: Chen, Xunlei, et al.
Veröffentlicht: (2026)
The Two Sides of the Coin: Hallucination Generation and Detection with LLMs as Evaluators for LLMs
von: Bui, Anh Thu Maria, et al.
Veröffentlicht: (2024)
von: Bui, Anh Thu Maria, et al.
Veröffentlicht: (2024)
SINdex: Semantic INconsistency Index for Hallucination Detection in LLMs
von: Abdaljalil, Samir, et al.
Veröffentlicht: (2025)
von: Abdaljalil, Samir, et al.
Veröffentlicht: (2025)
Hallucination Detection in LLMs with Topological Divergence on Attention Graphs
von: Bazarova, Alexandra, et al.
Veröffentlicht: (2025)
von: Bazarova, Alexandra, et al.
Veröffentlicht: (2025)
Can LLMs Detect Intrinsic Hallucinations in Paraphrasing and Machine Translation?
von: Gogoulou, Evangelia, et al.
Veröffentlicht: (2025)
von: Gogoulou, Evangelia, et al.
Veröffentlicht: (2025)
Graphing the Truth: Structured Visualizations for Automated Hallucination Detection in LLMs
von: Agrawal, Tanmay
Veröffentlicht: (2025)
von: Agrawal, Tanmay
Veröffentlicht: (2025)
Self-Instructed Derived Prompt Generation Meets In-Context Learning: Unlocking New Potential of Black-Box LLMs
von: Li, Zhuo, et al.
Veröffentlicht: (2024)
von: Li, Zhuo, et al.
Veröffentlicht: (2024)
Cost-Effective Hallucination Detection for LLMs
von: Valentin, Simon, et al.
Veröffentlicht: (2024)
von: Valentin, Simon, et al.
Veröffentlicht: (2024)
BERT-as-a-Judge: A Robust Alternative to Lexical Methods for Efficient Reference-Based LLM Evaluation
von: Gisserot-Boukhlef, Hippolyte, et al.
Veröffentlicht: (2026)
von: Gisserot-Boukhlef, Hippolyte, et al.
Veröffentlicht: (2026)
Scalable Token-Level Hallucination Detection in Large Language Models
von: Min, Rui, et al.
Veröffentlicht: (2026)
von: Min, Rui, et al.
Veröffentlicht: (2026)
Lie to Me: Knowledge Graphs for Robust Hallucination Self-Detection in LLMs
von: Kale, Sahil, et al.
Veröffentlicht: (2025)
von: Kale, Sahil, et al.
Veröffentlicht: (2025)
DSC2025 -- ViHallu Challenge: Detecting Hallucination in Vietnamese LLMs
von: Nguyen, Anh Thi-Hoang, et al.
Veröffentlicht: (2026)
von: Nguyen, Anh Thi-Hoang, et al.
Veröffentlicht: (2026)
Graph of Attacks: Improved Black-Box and Interpretable Jailbreaks for LLMs
von: Akbar-Tajari, Mohammad, et al.
Veröffentlicht: (2025)
von: Akbar-Tajari, Mohammad, et al.
Veröffentlicht: (2025)
In-Context Explainers: Harnessing LLMs for Explaining Black Box Models
von: Kroeger, Nicholas, et al.
Veröffentlicht: (2023)
von: Kroeger, Nicholas, et al.
Veröffentlicht: (2023)
Hallucination Detection and Hallucination Mitigation: An Investigation
von: Luo, Junliang, et al.
Veröffentlicht: (2024)
von: Luo, Junliang, et al.
Veröffentlicht: (2024)
Token-Guard: Towards Token-Level Hallucination Control via Self-Checking Decoding
von: Zhu, Yifan, et al.
Veröffentlicht: (2026)
von: Zhu, Yifan, et al.
Veröffentlicht: (2026)
You've Changed: Detecting Modification of Black-Box Large Language Models
von: Dima, Alden, et al.
Veröffentlicht: (2025)
von: Dima, Alden, et al.
Veröffentlicht: (2025)
Is LLMs Hallucination Usable? LLM-based Negative Reasoning for Fake News Detection
von: Zhang, Chaowei, et al.
Veröffentlicht: (2025)
von: Zhang, Chaowei, et al.
Veröffentlicht: (2025)
ICR Probe: Tracking Hidden State Dynamics for Reliable Hallucination Detection in LLMs
von: Zhang, Zhenliang, et al.
Veröffentlicht: (2025)
von: Zhang, Zhenliang, et al.
Veröffentlicht: (2025)
Black-Box Visual Prompt Engineering for Mitigating Object Hallucination in Large Vision Language Models
von: Woo, Sangmin, et al.
Veröffentlicht: (2025)
von: Woo, Sangmin, et al.
Veröffentlicht: (2025)
ShED-HD: A Shannon Entropy Distribution Framework for Lightweight Hallucination Detection on Edge Devices
von: Vathul, Aneesh, et al.
Veröffentlicht: (2025)
von: Vathul, Aneesh, et al.
Veröffentlicht: (2025)
Effective and Efficient Jailbreaks of Black-Box LLMs with Cross-Behavior Attacks
von: Gohil, Vasudev
Veröffentlicht: (2025)
von: Gohil, Vasudev
Veröffentlicht: (2025)
Training Deliberative Monitors for Black-Box Scheming Detection
von: Sinha, Aditya, et al.
Veröffentlicht: (2026)
von: Sinha, Aditya, et al.
Veröffentlicht: (2026)
A Geometric Taxonomy of Hallucinations in LLMs
von: Marín, Javier
Veröffentlicht: (2026)
von: Marín, Javier
Veröffentlicht: (2026)
The Illusion of Progress: Re-evaluating Hallucination Detection in LLMs
von: Janiak, Denis, et al.
Veröffentlicht: (2025)
von: Janiak, Denis, et al.
Veröffentlicht: (2025)
Cannot or Should Not? Automatic Analysis of Refusal Composition in IFT/RLHF Datasets and Refusal Behavior of Black-Box LLMs
von: von Recum, Alexander, et al.
Veröffentlicht: (2024)
von: von Recum, Alexander, et al.
Veröffentlicht: (2024)
How to Train Your Advisor: Steering Black-Box LLMs with Advisor Models
von: Asawa, Parth, et al.
Veröffentlicht: (2025)
von: Asawa, Parth, et al.
Veröffentlicht: (2025)
Bias Similarity Measurement: A Black-Box Audit of Fairness Across LLMs
von: Jeong, Hyejun, et al.
Veröffentlicht: (2024)
von: Jeong, Hyejun, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
FactSelfCheck: Fact-Level Black-Box Hallucination Detection for LLMs
von: Sawczyn, Albert, et al.
Veröffentlicht: (2025) -
Black-Box Hallucination Detection via Consistency Under the Uncertain Expression
von: Joo, Seongho, et al.
Veröffentlicht: (2025) -
Verify when Uncertain: Beyond Self-Consistency in Black Box Hallucination Detection
von: Xue, Yihao, et al.
Veröffentlicht: (2025) -
Semantic Entropy Probes: Robust and Cheap Hallucination Detection in LLMs
von: Kossen, Jannik, et al.
Veröffentlicht: (2024) -
TECP: Token-Entropy Conformal Prediction for LLMs
von: Xu, Beining, et al.
Veröffentlicht: (2025)