The Facade of Truth: Uncovering and Mitigating LLM Susceptibility to Deceptive Evidence
Fuente:
arXiv
Saved in:
| Main Authors: | Wan, Herun, Wu, Jiaying, Luo, Minnan, Li, Fanxiao, Zeng, Zhi, Kan, Min-Yen |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Truth over Tricks: Measuring and Mitigating Shortcut Learning in Misinformation Detection
by: Wan, Herun, et al.
Published: (2025)
by: Wan, Herun, et al.
Published: (2025)
Seeing Through Deception: Uncovering Misleading Creator Intent in Multimodal News with Vision-Language Models
by: Wu, Jiaying, et al.
Published: (2025)
by: Wu, Jiaying, et al.
Published: (2025)
DiFaR: Enhancing Multimodal Misinformation Detection with Diverse, Factual, and Relevant Rationales
by: Wan, Herun, et al.
Published: (2025)
by: Wan, Herun, et al.
Published: (2025)
On the Risk of Evidence Pollution for Malicious Social Text Detection in the Era of LLMs
by: Wan, Herun, et al.
Published: (2024)
by: Wan, Herun, et al.
Published: (2024)
Bot Meets Shortcut: How Can LLMs Aid in Handling Unknown Invariance OOD Scenarios?
by: Zheng, Shiyan, et al.
Published: (2025)
by: Zheng, Shiyan, et al.
Published: (2025)
Better with Experience: Self-Evolving LLM Agents for Evidence-Grounded Health Community Notes
by: Fu, Zihang, et al.
Published: (2026)
by: Fu, Zihang, et al.
Published: (2026)
DELL: Generating Reactions and Explanations for LLM-Based Misinformation Detection
by: Wan, Herun, et al.
Published: (2024)
by: Wan, Herun, et al.
Published: (2024)
Beyond the Crowd: LLM-Augmented Community Notes for Governing Health Misinformation
by: Wu, Jiaying, et al.
Published: (2025)
by: Wu, Jiaying, et al.
Published: (2025)
What's Left Unsaid? Detecting and Correcting Misleading Omissions in Multimodal News Previews
by: Li, Fanxiao, et al.
Published: (2026)
by: Li, Fanxiao, et al.
Published: (2026)
HACo-Det: A Study Towards Fine-Grained Machine-Generated Text Detection under Human-AI Coauthoring
by: Su, Zhixiong, et al.
Published: (2025)
by: Su, Zhixiong, et al.
Published: (2025)
What Does the Bot Say? Opportunities and Risks of Large Language Models in Social Media Bot Detection
by: Feng, Shangbin, et al.
Published: (2024)
by: Feng, Shangbin, et al.
Published: (2024)
GuessBench: Sensemaking Multimodal Creativity in the Wild
by: Zhu, Zifeng, et al.
Published: (2025)
by: Zhu, Zifeng, et al.
Published: (2025)
Forecasting the Buzz: Enriching Hashtag Popularity Prediction with LLM Reasoning
by: Xu, Yifei, et al.
Published: (2025)
by: Xu, Yifei, et al.
Published: (2025)
To Tell The Truth: Language of Deception and Language Models
by: Hazra, Sanchaita, et al.
Published: (2023)
by: Hazra, Sanchaita, et al.
Published: (2023)
CCSBench: Evaluating Compositional Controllability in LLMs for Scientific Document Summarization
by: Ding, Yixi, et al.
Published: (2024)
by: Ding, Yixi, et al.
Published: (2024)
How Do Social Bots Participate in Misinformation Spread? A Comprehensive Dataset and Analysis
by: Wan, Herun, et al.
Published: (2024)
by: Wan, Herun, et al.
Published: (2024)
From Manipulation to Mistrust: Explaining Diverse Micro-Video Misinformation for Robust Debunking in the Wild
by: Zeng, Zhi, et al.
Published: (2026)
by: Zeng, Zhi, et al.
Published: (2026)
CMIE: Combining MLLM Insights with External Evidence for Explainable Out-of-Context Misinformation Detection
by: Li, Fanxiao, et al.
Published: (2025)
by: Li, Fanxiao, et al.
Published: (2025)
FlowSteer: Prompt-Only Workflow Steering Exposes Planning-Time Vulnerabilities in Multi-Agent LLM Systems
by: Li, Fanxiao, et al.
Published: (2026)
by: Li, Fanxiao, et al.
Published: (2026)
Incorporating External Knowledge and Goal Guidance for LLM-based Conversational Recommender Systems
by: Li, Chuang, et al.
Published: (2024)
by: Li, Chuang, et al.
Published: (2024)
Each Fake News is Fake in its Own Way: An Attribution Multi-Granularity Benchmark for Multimodal Fake News Detection
by: Guo, Hao, et al.
Published: (2024)
by: Guo, Hao, et al.
Published: (2024)
LH-Deception: Simulating and Understanding LLM Deceptive Behaviors in Long-Horizon Interactions
by: Xu, Yang, et al.
Published: (2025)
by: Xu, Yang, et al.
Published: (2025)
Drifting Away from Truth: GenAI-Driven News Diversity Challenges LVLM-Based Misinformation Detection
by: Li, Fanxiao, et al.
Published: (2025)
by: Li, Fanxiao, et al.
Published: (2025)
When Truth Is Overridden: Uncovering the Internal Origins of Sycophancy in Large Language Models
by: Wang, Keyu, et al.
Published: (2025)
by: Wang, Keyu, et al.
Published: (2025)
Mitigating Forgetting in LLM Fine-Tuning via Low-Perplexity Token Learning
by: Wu, Chao-Chung, et al.
Published: (2025)
by: Wu, Chao-Chung, et al.
Published: (2025)
DeceptGuard :A Constitutional Oversight Framework For Detecting Deception in LLM Agents
by: Mukhopadhyay, Snehasis
Published: (2026)
by: Mukhopadhyay, Snehasis
Published: (2026)
Discursive Circuits: How Do Language Models Understand Discourse Relations?
by: Miao, Yisong, et al.
Published: (2025)
by: Miao, Yisong, et al.
Published: (2025)
UNO-DST: Leveraging Unlabelled Data in Zero-Shot Dialogue State Tracking
by: Li, Chuang, et al.
Published: (2023)
by: Li, Chuang, et al.
Published: (2023)
Continuously Steering LLMs Sensitivity to Contextual Knowledge with Proxy Models
by: Wang, Yilin, et al.
Published: (2025)
by: Wang, Yilin, et al.
Published: (2025)
OpenDeception: Learning Deception and Trust in Human-AI Interaction via Multi-Agent Simulation
by: Wu, Yichen, et al.
Published: (2025)
by: Wu, Yichen, et al.
Published: (2025)
GitSearch: Enhancing Community Notes Generation with Gap-Informed Targeted Search
by: Singh, Sahajpreet, et al.
Published: (2026)
by: Singh, Sahajpreet, et al.
Published: (2026)
DECOR: Auditing LLM Deception via Information Manipulation Theory
by: Cai, Linyue, et al.
Published: (2026)
by: Cai, Linyue, et al.
Published: (2026)
The Devil is in the Neurons: Interpreting and Mitigating Social Biases in Pre-trained Language Models
by: Liu, Yan, et al.
Published: (2024)
by: Liu, Yan, et al.
Published: (2024)
ISQA: Informative Factuality Feedback for Scientific Summarization
by: Li, Zekai, et al.
Published: (2024)
by: Li, Zekai, et al.
Published: (2024)
OI-Bench: An Option Injection Benchmark for Evaluating LLM Susceptibility to Directive Interference
by: Liou, Yow-Fu, et al.
Published: (2026)
by: Liou, Yow-Fu, et al.
Published: (2026)
TruthTorchLM: A Comprehensive Library for Predicting Truthfulness in LLM Outputs
by: Yaldiz, Duygu Nur, et al.
Published: (2025)
by: Yaldiz, Duygu Nur, et al.
Published: (2025)
Rethinking Verification for LLM Code Generation: From Generation to Testing
by: Ma, Zihan, et al.
Published: (2025)
by: Ma, Zihan, et al.
Published: (2025)
Two Pathways to Truthfulness: On the Intrinsic Encoding of LLM Hallucinations
by: Luo, Wen, et al.
Published: (2026)
by: Luo, Wen, et al.
Published: (2026)
Are Knowledge and Reference in Multilingual Language Models Cross-Lingually Consistent?
by: Ai, Xi, et al.
Published: (2025)
by: Ai, Xi, et al.
Published: (2025)
TruthEval: A Dataset to Evaluate LLM Truthfulness and Reliability
by: Khatun, Aisha, et al.
Published: (2024)
by: Khatun, Aisha, et al.
Published: (2024)
Similar Items
-
Truth over Tricks: Measuring and Mitigating Shortcut Learning in Misinformation Detection
by: Wan, Herun, et al.
Published: (2025) -
Seeing Through Deception: Uncovering Misleading Creator Intent in Multimodal News with Vision-Language Models
by: Wu, Jiaying, et al.
Published: (2025) -
DiFaR: Enhancing Multimodal Misinformation Detection with Diverse, Factual, and Relevant Rationales
by: Wan, Herun, et al.
Published: (2025) -
On the Risk of Evidence Pollution for Malicious Social Text Detection in the Era of LLMs
by: Wan, Herun, et al.
Published: (2024) -
Bot Meets Shortcut: How Can LLMs Aid in Handling Unknown Invariance OOD Scenarios?
by: Zheng, Shiyan, et al.
Published: (2025)