Shaking to Reveal: Perturbation-Based Detection of LLM Hallucinations
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Luo, Jinyuan, Fang, Zhen, Li, Yixuan, Park, Seongheon, Chen, Ling |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Steer LLM Latents for Hallucination Detection
von: Park, Seongheon, et al.
Veröffentlicht: (2025)
von: Park, Seongheon, et al.
Veröffentlicht: (2025)
GLSim: Detecting Object Hallucinations in LVLMs via Global-Local Similarity
von: Park, Seongheon, et al.
Veröffentlicht: (2025)
von: Park, Seongheon, et al.
Veröffentlicht: (2025)
Beyond In-Domain Detection: SpikeScore for Cross-Domain Hallucination Detection
von: Deng, Yongxin, et al.
Veröffentlicht: (2026)
von: Deng, Yongxin, et al.
Veröffentlicht: (2026)
ConjNorm: Tractable Density Estimation for Out-of-Distribution Detection
von: Peng, Bo, et al.
Veröffentlicht: (2024)
von: Peng, Bo, et al.
Veröffentlicht: (2024)
Whispers that Shake Foundations: Analyzing and Mitigating False Premise Hallucinations in Large Language Models
von: Yuan, Hongbang, et al.
Veröffentlicht: (2024)
von: Yuan, Hongbang, et al.
Veröffentlicht: (2024)
Understanding Language Prior of LVLMs by Contrasting Chain-of-Embedding
von: Long, Lin, et al.
Veröffentlicht: (2025)
von: Long, Lin, et al.
Veröffentlicht: (2025)
Beyond Functional Correctness: Exploring Hallucinations in LLM-Generated Code
von: Liu, Fang, et al.
Veröffentlicht: (2024)
von: Liu, Fang, et al.
Veröffentlicht: (2024)
Enhancing Hallucination Detection through Perturbation-Based Synthetic Data Generation in System Responses
von: Zhang, Dongxu, et al.
Veröffentlicht: (2024)
von: Zhang, Dongxu, et al.
Veröffentlicht: (2024)
Hallucination Detection and Hallucination Mitigation: An Investigation
von: Luo, Junliang, et al.
Veröffentlicht: (2024)
von: Luo, Junliang, et al.
Veröffentlicht: (2024)
Probing LLM Hallucination from Within: Perturbation-Driven Approach via Internal Knowledge
von: Lee, Seongmin, et al.
Veröffentlicht: (2024)
von: Lee, Seongmin, et al.
Veröffentlicht: (2024)
Assessing and Mitigating Miscalibration in LLM-Based Social Science Measurement
von: Wang, Jinyuan, et al.
Veröffentlicht: (2026)
von: Wang, Jinyuan, et al.
Veröffentlicht: (2026)
LLM-CAS: Dynamic Neuron Perturbation for Real-Time Hallucination Correction
von: Zhang, Jensen, et al.
Veröffentlicht: (2025)
von: Zhang, Jensen, et al.
Veröffentlicht: (2025)
HalluEntity: Benchmarking and Understanding Entity-Level Hallucination Detection
von: Yeh, Min-Hsuan, et al.
Veröffentlicht: (2025)
von: Yeh, Min-Hsuan, et al.
Veröffentlicht: (2025)
The MedPerturb Dataset: What Non-Content Perturbations Reveal About Human and Clinical LLM Decision Making
von: Gourabathina, Abinitha, et al.
Veröffentlicht: (2025)
von: Gourabathina, Abinitha, et al.
Veröffentlicht: (2025)
Enhancing Mathematical Reasoning in Large Language Models with Self-Consistency-Based Hallucination Detection
von: Liu, MingShan, et al.
Veröffentlicht: (2025)
von: Liu, MingShan, et al.
Veröffentlicht: (2025)
Uncertainty Quantification in LLM Agents: Foundations, Emerging Challenges, and Opportunities
von: Oh, Changdae, et al.
Veröffentlicht: (2026)
von: Oh, Changdae, et al.
Veröffentlicht: (2026)
On the Structural Memory of LLM Agents
von: Zeng, Ruihong, et al.
Veröffentlicht: (2024)
von: Zeng, Ruihong, et al.
Veröffentlicht: (2024)
Toxic HallucinAItions: Perturbing Prompts and Tracing LLM Circuits
von: Shimgekar, Soorya Ram, et al.
Veröffentlicht: (2026)
von: Shimgekar, Soorya Ram, et al.
Veröffentlicht: (2026)
Zero-resource Hallucination Detection for Text Generation via Graph-based Contextual Knowledge Triples Modeling
von: Fang, Xinyue, et al.
Veröffentlicht: (2024)
von: Fang, Xinyue, et al.
Veröffentlicht: (2024)
Shaking the Fake: Detecting Deepfake Videos in Real Time via Active Probes
von: Xie, Zhixin, et al.
Veröffentlicht: (2024)
von: Xie, Zhixin, et al.
Veröffentlicht: (2024)
Microsaccade-Inspired Probing: Positional Encoding Perturbations Reveal LLM Misbehaviours
von: Melo, Rui, et al.
Veröffentlicht: (2025)
von: Melo, Rui, et al.
Veröffentlicht: (2025)
Enhancing Hallucination Detection via Future Context
von: Lee, Joosung, et al.
Veröffentlicht: (2025)
von: Lee, Joosung, et al.
Veröffentlicht: (2025)
Hide-and-Seek in Trajectories: Discovering Failure Signals for VLA Runtime Monitoring
von: Park, Seongheon, et al.
Veröffentlicht: (2026)
von: Park, Seongheon, et al.
Veröffentlicht: (2026)
LLM Lies: Hallucinations are not Bugs, but Features as Adversarial Examples
von: Yao, Jia-Yu, et al.
Veröffentlicht: (2023)
von: Yao, Jia-Yu, et al.
Veröffentlicht: (2023)
EF-LLM: Energy Forecasting LLM with AI-assisted Automation, Enhanced Sparse Prediction, Hallucination Detection
von: Qiu, Zihang, et al.
Veröffentlicht: (2024)
von: Qiu, Zihang, et al.
Veröffentlicht: (2024)
Revealing Multi-View Hallucination in Large Vision-Language Models
von: Park, Wooje, et al.
Veröffentlicht: (2026)
von: Park, Wooje, et al.
Veröffentlicht: (2026)
OptArgus: A Multi-Agent System to Detect Hallucinations in LLM-based Optimization Modeling
von: Li, Zhong, et al.
Veröffentlicht: (2026)
von: Li, Zhong, et al.
Veröffentlicht: (2026)
RepreGuard: Detecting LLM-Generated Text by Revealing Hidden Representation Patterns
von: Chen, Xin, et al.
Veröffentlicht: (2025)
von: Chen, Xin, et al.
Veröffentlicht: (2025)
FaithLens: Detecting and Explaining Faithfulness Hallucination
von: Si, Shuzheng, et al.
Veröffentlicht: (2025)
von: Si, Shuzheng, et al.
Veröffentlicht: (2025)
Hallucination Detection and Mitigation in Large Language Models
von: Pesaranghader, Ahmad, et al.
Veröffentlicht: (2026)
von: Pesaranghader, Ahmad, et al.
Veröffentlicht: (2026)
Is LLMs Hallucination Usable? LLM-based Negative Reasoning for Fake News Detection
von: Zhang, Chaowei, et al.
Veröffentlicht: (2025)
von: Zhang, Chaowei, et al.
Veröffentlicht: (2025)
Respecting Modality Gap in Post-hoc Out-of-distribution Detection with Pre-trained Vision-Language Models
von: Hu, Yuanwei, et al.
Veröffentlicht: (2026)
von: Hu, Yuanwei, et al.
Veröffentlicht: (2026)
Rethinking Evaluation for LLM Hallucination Detection: A Desiderata, A New RAG-based Benchmark, New Insights
von: Chen, Wenbo, et al.
Veröffentlicht: (2026)
von: Chen, Wenbo, et al.
Veröffentlicht: (2026)
Two Pathways to Truthfulness: On the Intrinsic Encoding of LLM Hallucinations
von: Luo, Wen, et al.
Veröffentlicht: (2026)
von: Luo, Wen, et al.
Veröffentlicht: (2026)
PerturboLLaVA: Reducing Multimodal Hallucinations with Perturbative Visual Training
von: Chen, Cong, et al.
Veröffentlicht: (2025)
von: Chen, Cong, et al.
Veröffentlicht: (2025)
AutoPBO: LLM-powered Optimization for Local Search PBO Solvers
von: Li, Jinyuan, et al.
Veröffentlicht: (2025)
von: Li, Jinyuan, et al.
Veröffentlicht: (2025)
NoiseBoost: Alleviating Hallucination with Noise Perturbation for Multimodal Large Language Models
von: Wu, Kai, et al.
Veröffentlicht: (2024)
von: Wu, Kai, et al.
Veröffentlicht: (2024)
VAUQ: Vision-Aware Uncertainty Quantification for LVLM Self-Evaluation
von: Park, Seongheon, et al.
Veröffentlicht: (2026)
von: Park, Seongheon, et al.
Veröffentlicht: (2026)
Failure Ontology: A Lifelong Learning Framework for Blind Spot Detection and Resilience Design
von: Sun, Yuan, et al.
Veröffentlicht: (2026)
von: Sun, Yuan, et al.
Veröffentlicht: (2026)
Nonsense Helps: Prompt Space Perturbation Broadens Reasoning Exploration
von: Huang, Langlin, et al.
Veröffentlicht: (2026)
von: Huang, Langlin, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Steer LLM Latents for Hallucination Detection
von: Park, Seongheon, et al.
Veröffentlicht: (2025) -
GLSim: Detecting Object Hallucinations in LVLMs via Global-Local Similarity
von: Park, Seongheon, et al.
Veröffentlicht: (2025) -
Beyond In-Domain Detection: SpikeScore for Cross-Domain Hallucination Detection
von: Deng, Yongxin, et al.
Veröffentlicht: (2026) -
ConjNorm: Tractable Density Estimation for Out-of-Distribution Detection
von: Peng, Bo, et al.
Veröffentlicht: (2024) -
Whispers that Shake Foundations: Analyzing and Mitigating False Premise Hallucinations in Large Language Models
von: Yuan, Hongbang, et al.
Veröffentlicht: (2024)