Detecting LLM Fact-conflicting Hallucinations Enhanced by Temporal-logic-based Reasoning
Fuente:
arXiv
Guardado en:
| Autores principales: | Li, Ningke, Song, Yahui, Wang, Kailong, Li, Yuekang, Shi, Ling, Liu, Yi, Wang, Haoyu |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Drowzee: Metamorphic Testing for Fact-Conflicting Hallucination Detection in Large Language Models
por: Li, Ningke, et al.
Publicado: (2024)
por: Li, Ningke, et al.
Publicado: (2024)
Beyond Correctness: Exposing LLM-generated Logical Flaws in Reasoning via Multi-step Automated Theorem Proving
por: Zheng, Xinyi, et al.
Publicado: (2025)
por: Zheng, Xinyi, et al.
Publicado: (2025)
Glitch Tokens in Large Language Models: Categorization Taxonomy and Effective Detection
por: Li, Yuxi, et al.
Publicado: (2024)
por: Li, Yuxi, et al.
Publicado: (2024)
Continuous Embedding Attacks via Clipped Inputs in Jailbreaking Large Language Models
por: Xu, Zihao, et al.
Publicado: (2024)
por: Xu, Zihao, et al.
Publicado: (2024)
Digger: Detecting Copyright Content Mis-usage in Large Language Model Training
por: Li, Haodong, et al.
Publicado: (2024)
por: Li, Haodong, et al.
Publicado: (2024)
OnionEval: An Unified Evaluation of Fact-conflicting Hallucination for Small-Large Language Models
por: Sun, Chongren, et al.
Publicado: (2025)
por: Sun, Chongren, et al.
Publicado: (2025)
Circumventing Safety Alignment in Large Language Models Through Embedding Space Toxicity Attenuation
por: Zhang, Zhibo, et al.
Publicado: (2025)
por: Zhang, Zhibo, et al.
Publicado: (2025)
Prompt Injection attack against LLM-integrated Applications
por: Liu, Yi, et al.
Publicado: (2023)
por: Liu, Yi, et al.
Publicado: (2023)
GlitchProber: Advancing Effective Detection and Mitigation of Glitch Tokens in Large Language Models
por: Zhang, Zhibo, et al.
Publicado: (2024)
por: Zhang, Zhibo, et al.
Publicado: (2024)
FactCHD: Benchmarking Fact-Conflicting Hallucination Detection
por: Chen, Xiang, et al.
Publicado: (2023)
por: Chen, Xiang, et al.
Publicado: (2023)
ChronoFact: Timeline-based Temporal Fact Verification
por: Barik, Anab Maulana, et al.
Publicado: (2024)
por: Barik, Anab Maulana, et al.
Publicado: (2024)
Detection and Mitigation of Hallucination in Large Reasoning Models: A Mechanistic Perspective
por: Sun, Zhongxiang, et al.
Publicado: (2025)
por: Sun, Zhongxiang, et al.
Publicado: (2025)
STEAMROLLER: A Multi-Agent System for Inclusive Automatic Speech Recognition for People who Stutter
por: Xu, Ziqi, et al.
Publicado: (2026)
por: Xu, Ziqi, et al.
Publicado: (2026)
Is LLMs Hallucination Usable? LLM-based Negative Reasoning for Fake News Detection
por: Zhang, Chaowei, et al.
Publicado: (2025)
por: Zhang, Chaowei, et al.
Publicado: (2025)
Large Language Models are overconfident and amplify human bias
por: Sun, Fengfei, et al.
Publicado: (2025)
por: Sun, Fengfei, et al.
Publicado: (2025)
FactCheckmate: Preemptively Detecting and Mitigating Hallucinations in LMs
por: Alnuhait, Deema, et al.
Publicado: (2024)
por: Alnuhait, Deema, et al.
Publicado: (2024)
LLM Hallucination Detection: HSAD
por: Li, JinXin, et al.
Publicado: (2025)
por: Li, JinXin, et al.
Publicado: (2025)
Towards Unification of Hallucination Detection and Fact Verification for Large Language Models
por: Su, Weihang, et al.
Publicado: (2025)
por: Su, Weihang, et al.
Publicado: (2025)
LLM Hallucination Detection: A Fast Fourier Transform Method Based on Hidden Layer Temporal Signals
por: Li, Jinxin, et al.
Publicado: (2025)
por: Li, Jinxin, et al.
Publicado: (2025)
Joint Evaluation of Answer and Reasoning Consistency for Hallucination Detection in Large Reasoning Models
por: Wang, Changyue, et al.
Publicado: (2025)
por: Wang, Changyue, et al.
Publicado: (2025)
LLM-Guided Knowledge Distillation for Temporal Knowledge Graph Reasoning
por: Xing, Wang, et al.
Publicado: (2026)
por: Xing, Wang, et al.
Publicado: (2026)
Reducing Hallucinations in Entity Abstract Summarization with Facts-Template Decomposition
por: Zhu, Fangwei, et al.
Publicado: (2024)
por: Zhu, Fangwei, et al.
Publicado: (2024)
Steer LLM Latents for Hallucination Detection
por: Park, Seongheon, et al.
Publicado: (2025)
por: Park, Seongheon, et al.
Publicado: (2025)
How to Detect and Defeat Molecular Mirage: A Metric-Driven Benchmark for Hallucination in LLM-based Molecular Comprehension
por: Li, Hao, et al.
Publicado: (2025)
por: Li, Hao, et al.
Publicado: (2025)
FactSelfCheck: Fact-Level Black-Box Hallucination Detection for LLMs
por: Sawczyn, Albert, et al.
Publicado: (2025)
por: Sawczyn, Albert, et al.
Publicado: (2025)
VeriFact: Enhancing Long-Form Factuality Evaluation with Refined Fact Extraction and Reference Facts
por: Liu, Xin, et al.
Publicado: (2025)
por: Liu, Xin, et al.
Publicado: (2025)
Beyond Translation: LLM-Based Data Generation for Multilingual Fact-Checking
por: Chung, Yi-Ling, et al.
Publicado: (2025)
por: Chung, Yi-Ling, et al.
Publicado: (2025)
AgentHallu: Benchmarking Automated Hallucination Attribution of LLM-based Agents
por: Liu, Xuannan, et al.
Publicado: (2026)
por: Liu, Xuannan, et al.
Publicado: (2026)
Mitigating Hallucinations in Large Vision-Language Models with Internal Fact-based Contrastive Decoding
por: Wang, Chao, et al.
Publicado: (2025)
por: Wang, Chao, et al.
Publicado: (2025)
Jailbreaking ChatGPT via Prompt Engineering: An Empirical Study
por: Liu, Yi, et al.
Publicado: (2023)
por: Liu, Yi, et al.
Publicado: (2023)
Enhancing Temporal Sensitivity and Reasoning for Time-Sensitive Question Answering
por: Yang, Wanqi, et al.
Publicado: (2024)
por: Yang, Wanqi, et al.
Publicado: (2024)
Socrates or Smartypants: Testing Logic Reasoning Capabilities of Large Language Models with Logic Programming-based Test Oracles
por: Xu, Zihao, et al.
Publicado: (2025)
por: Xu, Zihao, et al.
Publicado: (2025)
HARP: Hallucination Detection via Reasoning Subspace Projection
por: Hu, Junjie, et al.
Publicado: (2025)
por: Hu, Junjie, et al.
Publicado: (2025)
Let's Reason Formally: Natural-Formal Hybrid Reasoning Enhances LLM's Math Capability
por: Wang, Ruida, et al.
Publicado: (2025)
por: Wang, Ruida, et al.
Publicado: (2025)
Consistency Is the Key: Detecting Hallucinations in LLM Generated Text By Checking Inconsistencies About Key Facts
por: Gupta, Raavi, et al.
Publicado: (2025)
por: Gupta, Raavi, et al.
Publicado: (2025)
MedFact: A Large-scale Chinese Dataset for Evidence-based Medical Fact-checking of LLM Responses
por: Chen, Tong, et al.
Publicado: (2025)
por: Chen, Tong, et al.
Publicado: (2025)
From Hallucinations to Facts: Enhancing Language Models with Curated Knowledge Graphs
por: Joshi, Ratnesh Kumar, et al.
Publicado: (2024)
por: Joshi, Ratnesh Kumar, et al.
Publicado: (2024)
Uncovering Logit Suppression Vulnerabilities in LLM Safety Alignment
por: Li, Yuxi, et al.
Publicado: (2024)
por: Li, Yuxi, et al.
Publicado: (2024)
Probability-Consistent Preference Optimization for Enhanced LLM Reasoning
por: Yang, Yunqiao, et al.
Publicado: (2025)
por: Yang, Yunqiao, et al.
Publicado: (2025)
IndexRAG: Bridging Facts for Cross-Document Reasoning at Index Time
por: Bao, Zhenghua, et al.
Publicado: (2026)
por: Bao, Zhenghua, et al.
Publicado: (2026)
Ejemplares similares
-
Drowzee: Metamorphic Testing for Fact-Conflicting Hallucination Detection in Large Language Models
por: Li, Ningke, et al.
Publicado: (2024) -
Beyond Correctness: Exposing LLM-generated Logical Flaws in Reasoning via Multi-step Automated Theorem Proving
por: Zheng, Xinyi, et al.
Publicado: (2025) -
Glitch Tokens in Large Language Models: Categorization Taxonomy and Effective Detection
por: Li, Yuxi, et al.
Publicado: (2024) -
Continuous Embedding Attacks via Clipped Inputs in Jailbreaking Large Language Models
por: Xu, Zihao, et al.
Publicado: (2024) -
Digger: Detecting Copyright Content Mis-usage in Large Language Model Training
por: Li, Haodong, et al.
Publicado: (2024)