Hidden in Plain Sight: Evaluation of the Deception Detection Capabilities of LLMs in Multimodal Settings
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Miah, Md Messal Monem, Anika, Adrita, Shi, Xi, Huang, Ruihong |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Evaluating LLMs' Reasoning Over Ordered Procedural Steps
von: Anika, Adrita, et al.
Veröffentlicht: (2025)
von: Anika, Adrita, et al.
Veröffentlicht: (2025)
Multimodal Contextual Dialogue Breakdown Detection for Conversational AI Models
von: Miah, Md Messal Monem, et al.
Veröffentlicht: (2024)
von: Miah, Md Messal Monem, et al.
Veröffentlicht: (2024)
EMONA: Event-level Moral Opinions in News Articles
von: Lei, Yuanyuan, et al.
Veröffentlicht: (2024)
von: Lei, Yuanyuan, et al.
Veröffentlicht: (2024)
From Scarcity to Capability: Empowering Fake News Detection in Low-Resource Languages with LLMs
von: Shibu, Hrithik Majumdar, et al.
Veröffentlicht: (2025)
von: Shibu, Hrithik Majumdar, et al.
Veröffentlicht: (2025)
CliME: Evaluating Multimodal Climate Discourse on Social Media and the Climate Alignment Quotient (CAQ)
von: Borah, Abhilekh, et al.
Veröffentlicht: (2025)
von: Borah, Abhilekh, et al.
Veröffentlicht: (2025)
Hidden in Plain Sight: Where Developers Confess Self-Admitted Technical Debt
von: Sridharan, Murali, et al.
Veröffentlicht: (2025)
von: Sridharan, Murali, et al.
Veröffentlicht: (2025)
Mitigating Gender Bias via Fostering Exploratory Thinking in LLMs
von: Wei, Kangda, et al.
Veröffentlicht: (2025)
von: Wei, Kangda, et al.
Veröffentlicht: (2025)
Evaluating Gender Bias of LLMs in Making Morality Judgements
von: Bajaj, Divij, et al.
Veröffentlicht: (2024)
von: Bajaj, Divij, et al.
Veröffentlicht: (2024)
Boosting Logical Fallacy Reasoning in LLMs via Logical Structure Tree
von: Lei, Yuanyuan, et al.
Veröffentlicht: (2024)
von: Lei, Yuanyuan, et al.
Veröffentlicht: (2024)
Dynamic Emotion and Personality Profiling for Multimodal Deception Detection
von: Zheng, Li, et al.
Veröffentlicht: (2026)
von: Zheng, Li, et al.
Veröffentlicht: (2026)
Multi-document Summarization through Multi-document Event Relation Graph Reasoning in LLMs: a case study in Framing Bias Mitigation
von: Lei, Yuanyuan, et al.
Veröffentlicht: (2025)
von: Lei, Yuanyuan, et al.
Veröffentlicht: (2025)
Insight Over Sight: Exploring the Vision-Knowledge Conflicts in Multimodal LLMs
von: Liu, Xiaoyuan, et al.
Veröffentlicht: (2024)
von: Liu, Xiaoyuan, et al.
Veröffentlicht: (2024)
Can Deception Detection Go Deeper? Dataset, Evaluation, and Benchmark for Deception Reasoning
von: Chen, Kang, et al.
Veröffentlicht: (2024)
von: Chen, Kang, et al.
Veröffentlicht: (2024)
Counterfactual Evaluation Reveals Hidden Capability Profiles in Clinical LLMs and Agents
von: Turk, Matt
Veröffentlicht: (2026)
von: Turk, Matt
Veröffentlicht: (2026)
ShredBench: Evaluating the Semantic Reasoning Capabilities of Multimodal LLMs in Document Reconstruction
von: Guo, Zichun, et al.
Veröffentlicht: (2026)
von: Guo, Zichun, et al.
Veröffentlicht: (2026)
Are LLMs Good Annotators for Discourse-level Event Relation Extraction?
von: Wei, Kangda, et al.
Veröffentlicht: (2024)
von: Wei, Kangda, et al.
Veröffentlicht: (2024)
Hiding in Plain Sight: Finding MAHA on Reddit
von: Ahmed, Sabit, et al.
Veröffentlicht: (2026)
von: Ahmed, Sabit, et al.
Veröffentlicht: (2026)
From Sight to Insight: Improving Visual Reasoning Capabilities of Multimodal Models via Reinforcement Learning
von: Sharif, Omar, et al.
Veröffentlicht: (2026)
von: Sharif, Omar, et al.
Veröffentlicht: (2026)
TRN-R1-Zero: Text-rich Network Reasoning via LLMs with Reinforcement Learning Only
von: Liu, Yilun, et al.
Veröffentlicht: (2026)
von: Liu, Yilun, et al.
Veröffentlicht: (2026)
Hidden in Plain Text: Measuring LLM Deception Quality Against Human Baselines Using Social Deduction Games
von: Kao, Christopher, et al.
Veröffentlicht: (2026)
von: Kao, Christopher, et al.
Veröffentlicht: (2026)
Evaluating Linguistic Capabilities of Multimodal LLMs in the Lens of Few-Shot Learning
von: Dogan, Mustafa, et al.
Veröffentlicht: (2024)
von: Dogan, Mustafa, et al.
Veröffentlicht: (2024)
Voting-based Multimodal Automatic Deception Detection
von: Touma, Lana, et al.
Veröffentlicht: (2023)
von: Touma, Lana, et al.
Veröffentlicht: (2023)
Hidden in Plain Text: Emergence & Mitigation of Steganographic Collusion in LLMs
von: Mathew, Yohan, et al.
Veröffentlicht: (2024)
von: Mathew, Yohan, et al.
Veröffentlicht: (2024)
Sentence-level Media Bias Analysis with Event Relation Graph
von: Lei, Yuanyuan, et al.
Veröffentlicht: (2024)
von: Lei, Yuanyuan, et al.
Veröffentlicht: (2024)
How Easy is It to Fool Your Multimodal LLMs? An Empirical Analysis on Deceptive Prompts
von: Qian, Yusu, et al.
Veröffentlicht: (2024)
von: Qian, Yusu, et al.
Veröffentlicht: (2024)
LieCraft: A Multi-Agent Framework for Evaluating Deceptive Capabilities in Language Models
von: Olson, Matthew Lyle, et al.
Veröffentlicht: (2026)
von: Olson, Matthew Lyle, et al.
Veröffentlicht: (2026)
A Structured Framework for Evaluating and Enhancing Interpretive Capabilities of Multimodal LLMs in Culturally Situated Tasks
von: Yu, Haorui, et al.
Veröffentlicht: (2025)
von: Yu, Haorui, et al.
Veröffentlicht: (2025)
Honeyfile Camouflage: Hiding Fake Files in Plain Sight
von: Timmer, Roelien C., et al.
Veröffentlicht: (2024)
von: Timmer, Roelien C., et al.
Veröffentlicht: (2024)
Evaluating LLMs' Multilingual Capabilities for Bengali: Benchmark Creation and Performance Analysis
von: Bhowmik, Shimanto, et al.
Veröffentlicht: (2025)
von: Bhowmik, Shimanto, et al.
Veröffentlicht: (2025)
'Since Lawyers are Males..': Examining Implicit Gender Bias in Hindi Language Generation by LLMs
von: Joshi, Ishika, et al.
Veröffentlicht: (2024)
von: Joshi, Ishika, et al.
Veröffentlicht: (2024)
"Hiding in Plain Sight": Designing Synthetic Dialog Generation for Uncovering Socially Situated Norms
von: Wu, Chengfei, et al.
Veröffentlicht: (2024)
von: Wu, Chengfei, et al.
Veröffentlicht: (2024)
Are the Hidden States Hiding Something? Testing the Limits of Factuality-Encoding Capabilities in LLMs
von: Servedio, Giovanni, et al.
Veröffentlicht: (2025)
von: Servedio, Giovanni, et al.
Veröffentlicht: (2025)
DeceptGuard :A Constitutional Oversight Framework For Detecting Deception in LLM Agents
von: Mukhopadhyay, Snehasis
Veröffentlicht: (2026)
von: Mukhopadhyay, Snehasis
Veröffentlicht: (2026)
What if Deception Cannot be Detected? A Cross-Linguistic Study on the Limits of Deception Detection from Text
von: Velutharambath, Aswathy, et al.
Veröffentlicht: (2025)
von: Velutharambath, Aswathy, et al.
Veröffentlicht: (2025)
Code-Vision: Evaluating Multimodal LLMs Logic Understanding and Code Generation Capabilities
von: Wang, Hanbin, et al.
Veröffentlicht: (2025)
von: Wang, Hanbin, et al.
Veröffentlicht: (2025)
Hidden in Plain Sight: Reasoning in Underspecified and Misspecified Scenarios for Multimodal LLMs
von: Yan, Qianqi, et al.
Veröffentlicht: (2025)
von: Yan, Qianqi, et al.
Veröffentlicht: (2025)
Towards Automatic Evaluation for LLMs' Clinical Capabilities: Metric, Data, and Algorithm
von: Liu, Lei, et al.
Veröffentlicht: (2024)
von: Liu, Lei, et al.
Veröffentlicht: (2024)
Unmasking the Shadows of AI: Investigating Deceptive Capabilities in Large Language Models
von: Guo, Linge
Veröffentlicht: (2024)
von: Guo, Linge
Veröffentlicht: (2024)
Text Meets Topology: Rethinking Out-of-distribution Detection in Text-Rich Networks
von: Wang, Danny, et al.
Veröffentlicht: (2025)
von: Wang, Danny, et al.
Veröffentlicht: (2025)
Hidden in Plain Sight: Detecting Illicit Massage Businesses from Mobility Data
von: Shomali, Roya, et al.
Veröffentlicht: (2026)
von: Shomali, Roya, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Evaluating LLMs' Reasoning Over Ordered Procedural Steps
von: Anika, Adrita, et al.
Veröffentlicht: (2025) -
Multimodal Contextual Dialogue Breakdown Detection for Conversational AI Models
von: Miah, Md Messal Monem, et al.
Veröffentlicht: (2024) -
EMONA: Event-level Moral Opinions in News Articles
von: Lei, Yuanyuan, et al.
Veröffentlicht: (2024) -
From Scarcity to Capability: Empowering Fake News Detection in Low-Resource Languages with LLMs
von: Shibu, Hrithik Majumdar, et al.
Veröffentlicht: (2025) -
CliME: Evaluating Multimodal Climate Discourse on Social Media and the Climate Alignment Quotient (CAQ)
von: Borah, Abhilekh, et al.
Veröffentlicht: (2025)