Gespeichert in:
| Hauptverfasser: | Li, Ningke, Song, Yahui, Wang, Kailong, Li, Yuekang, Shi, Ling, Liu, Yi, Wang, Haoyu |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2502.13416 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Drowzee: Metamorphic Testing for Fact-Conflicting Hallucination Detection in Large Language Models
von: Li, Ningke, et al.
Veröffentlicht: (2024)
von: Li, Ningke, et al.
Veröffentlicht: (2024)
Beyond Correctness: Exposing LLM-generated Logical Flaws in Reasoning via Multi-step Automated Theorem Proving
von: Zheng, Xinyi, et al.
Veröffentlicht: (2025)
von: Zheng, Xinyi, et al.
Veröffentlicht: (2025)
Glitch Tokens in Large Language Models: Categorization Taxonomy and Effective Detection
von: Li, Yuxi, et al.
Veröffentlicht: (2024)
von: Li, Yuxi, et al.
Veröffentlicht: (2024)
Continuous Embedding Attacks via Clipped Inputs in Jailbreaking Large Language Models
von: Xu, Zihao, et al.
Veröffentlicht: (2024)
von: Xu, Zihao, et al.
Veröffentlicht: (2024)
Digger: Detecting Copyright Content Mis-usage in Large Language Model Training
von: Li, Haodong, et al.
Veröffentlicht: (2024)
von: Li, Haodong, et al.
Veröffentlicht: (2024)
Prompt Injection attack against LLM-integrated Applications
von: Liu, Yi, et al.
Veröffentlicht: (2023)
von: Liu, Yi, et al.
Veröffentlicht: (2023)
Circumventing Safety Alignment in Large Language Models Through Embedding Space Toxicity Attenuation
von: Zhang, Zhibo, et al.
Veröffentlicht: (2025)
von: Zhang, Zhibo, et al.
Veröffentlicht: (2025)
OnionEval: An Unified Evaluation of Fact-conflicting Hallucination for Small-Large Language Models
von: Sun, Chongren, et al.
Veröffentlicht: (2025)
von: Sun, Chongren, et al.
Veröffentlicht: (2025)
GlitchProber: Advancing Effective Detection and Mitigation of Glitch Tokens in Large Language Models
von: Zhang, Zhibo, et al.
Veröffentlicht: (2024)
von: Zhang, Zhibo, et al.
Veröffentlicht: (2024)
STEAMROLLER: A Multi-Agent System for Inclusive Automatic Speech Recognition for People who Stutter
von: Xu, Ziqi, et al.
Veröffentlicht: (2026)
von: Xu, Ziqi, et al.
Veröffentlicht: (2026)
FactCHD: Benchmarking Fact-Conflicting Hallucination Detection
von: Chen, Xiang, et al.
Veröffentlicht: (2023)
von: Chen, Xiang, et al.
Veröffentlicht: (2023)
Large Language Models are overconfident and amplify human bias
von: Sun, Fengfei, et al.
Veröffentlicht: (2025)
von: Sun, Fengfei, et al.
Veröffentlicht: (2025)
Detection and Mitigation of Hallucination in Large Reasoning Models: A Mechanistic Perspective
von: Sun, Zhongxiang, et al.
Veröffentlicht: (2025)
von: Sun, Zhongxiang, et al.
Veröffentlicht: (2025)
ChronoFact: Timeline-based Temporal Fact Verification
von: Barik, Anab Maulana, et al.
Veröffentlicht: (2024)
von: Barik, Anab Maulana, et al.
Veröffentlicht: (2024)
Is LLMs Hallucination Usable? LLM-based Negative Reasoning for Fake News Detection
von: Zhang, Chaowei, et al.
Veröffentlicht: (2025)
von: Zhang, Chaowei, et al.
Veröffentlicht: (2025)
Jailbreaking ChatGPT via Prompt Engineering: An Empirical Study
von: Liu, Yi, et al.
Veröffentlicht: (2023)
von: Liu, Yi, et al.
Veröffentlicht: (2023)
Towards Unification of Hallucination Detection and Fact Verification for Large Language Models
von: Su, Weihang, et al.
Veröffentlicht: (2025)
von: Su, Weihang, et al.
Veröffentlicht: (2025)
FactCheckmate: Preemptively Detecting and Mitigating Hallucinations in LMs
von: Alnuhait, Deema, et al.
Veröffentlicht: (2024)
von: Alnuhait, Deema, et al.
Veröffentlicht: (2024)
Uncovering Logit Suppression Vulnerabilities in LLM Safety Alignment
von: Li, Yuxi, et al.
Veröffentlicht: (2024)
von: Li, Yuxi, et al.
Veröffentlicht: (2024)
LLM Hallucination Detection: HSAD
von: Li, JinXin, et al.
Veröffentlicht: (2025)
von: Li, JinXin, et al.
Veröffentlicht: (2025)
MiniScope: Automated UI Exploration and Privacy Inconsistency Detection of MiniApps via Two-phase Iterative Hybrid Analysis
von: Wang, Shenao, et al.
Veröffentlicht: (2024)
von: Wang, Shenao, et al.
Veröffentlicht: (2024)
LLM Hallucination Detection: A Fast Fourier Transform Method Based on Hidden Layer Temporal Signals
von: Li, Jinxin, et al.
Veröffentlicht: (2025)
von: Li, Jinxin, et al.
Veröffentlicht: (2025)
Steer LLM Latents for Hallucination Detection
von: Park, Seongheon, et al.
Veröffentlicht: (2025)
von: Park, Seongheon, et al.
Veröffentlicht: (2025)
Joint Evaluation of Answer and Reasoning Consistency for Hallucination Detection in Large Reasoning Models
von: Wang, Changyue, et al.
Veröffentlicht: (2025)
von: Wang, Changyue, et al.
Veröffentlicht: (2025)
Reducing Hallucinations in Entity Abstract Summarization with Facts-Template Decomposition
von: Zhu, Fangwei, et al.
Veröffentlicht: (2024)
von: Zhu, Fangwei, et al.
Veröffentlicht: (2024)
MeTMaP: Metamorphic Testing for Detecting False Vector Matching Problems in LLM Augmented Generation
von: Wang, Guanyu, et al.
Veröffentlicht: (2024)
von: Wang, Guanyu, et al.
Veröffentlicht: (2024)
FactSelfCheck: Fact-Level Black-Box Hallucination Detection for LLMs
von: Sawczyn, Albert, et al.
Veröffentlicht: (2025)
von: Sawczyn, Albert, et al.
Veröffentlicht: (2025)
LLM-Guided Knowledge Distillation for Temporal Knowledge Graph Reasoning
von: Xing, Wang, et al.
Veröffentlicht: (2026)
von: Xing, Wang, et al.
Veröffentlicht: (2026)
Beyond Translation: LLM-Based Data Generation for Multilingual Fact-Checking
von: Chung, Yi-Ling, et al.
Veröffentlicht: (2025)
von: Chung, Yi-Ling, et al.
Veröffentlicht: (2025)
Socrates or Smartypants: Testing Logic Reasoning Capabilities of Large Language Models with Logic Programming-based Test Oracles
von: Xu, Zihao, et al.
Veröffentlicht: (2025)
von: Xu, Zihao, et al.
Veröffentlicht: (2025)
How to Detect and Defeat Molecular Mirage: A Metric-Driven Benchmark for Hallucination in LLM-based Molecular Comprehension
von: Li, Hao, et al.
Veröffentlicht: (2025)
von: Li, Hao, et al.
Veröffentlicht: (2025)
Mitigating Hallucinations in Large Vision-Language Models with Internal Fact-based Contrastive Decoding
von: Wang, Chao, et al.
Veröffentlicht: (2025)
von: Wang, Chao, et al.
Veröffentlicht: (2025)
VeriFact: Enhancing Long-Form Factuality Evaluation with Refined Fact Extraction and Reference Facts
von: Liu, Xin, et al.
Veröffentlicht: (2025)
von: Liu, Xin, et al.
Veröffentlicht: (2025)
Enhancing Temporal Sensitivity and Reasoning for Time-Sensitive Question Answering
von: Yang, Wanqi, et al.
Veröffentlicht: (2024)
von: Yang, Wanqi, et al.
Veröffentlicht: (2024)
Consistency Is the Key: Detecting Hallucinations in LLM Generated Text By Checking Inconsistencies About Key Facts
von: Gupta, Raavi, et al.
Veröffentlicht: (2025)
von: Gupta, Raavi, et al.
Veröffentlicht: (2025)
AgentHallu: Benchmarking Automated Hallucination Attribution of LLM-based Agents
von: Liu, Xuannan, et al.
Veröffentlicht: (2026)
von: Liu, Xuannan, et al.
Veröffentlicht: (2026)
HARP: Hallucination Detection via Reasoning Subspace Projection
von: Hu, Junjie, et al.
Veröffentlicht: (2025)
von: Hu, Junjie, et al.
Veröffentlicht: (2025)
Let's Reason Formally: Natural-Formal Hybrid Reasoning Enhances LLM's Math Capability
von: Wang, Ruida, et al.
Veröffentlicht: (2025)
von: Wang, Ruida, et al.
Veröffentlicht: (2025)
IndexRAG: Bridging Facts for Cross-Document Reasoning at Index Time
von: Bao, Zhenghua, et al.
Veröffentlicht: (2026)
von: Bao, Zhenghua, et al.
Veröffentlicht: (2026)
From Hallucinations to Facts: Enhancing Language Models with Curated Knowledge Graphs
von: Joshi, Ratnesh Kumar, et al.
Veröffentlicht: (2024)
von: Joshi, Ratnesh Kumar, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Drowzee: Metamorphic Testing for Fact-Conflicting Hallucination Detection in Large Language Models
von: Li, Ningke, et al.
Veröffentlicht: (2024) -
Beyond Correctness: Exposing LLM-generated Logical Flaws in Reasoning via Multi-step Automated Theorem Proving
von: Zheng, Xinyi, et al.
Veröffentlicht: (2025) -
Glitch Tokens in Large Language Models: Categorization Taxonomy and Effective Detection
von: Li, Yuxi, et al.
Veröffentlicht: (2024) -
Continuous Embedding Attacks via Clipped Inputs in Jailbreaking Large Language Models
von: Xu, Zihao, et al.
Veröffentlicht: (2024) -
Digger: Detecting Copyright Content Mis-usage in Large Language Model Training
von: Li, Haodong, et al.
Veröffentlicht: (2024)