CAMOUFLAGE: Exploiting Misinformation Detection Systems Through LLM-driven Adversarial Claim Transformation
Fuente:
arXiv
Saved in:
| Main Authors: | Bethany, Mazal, Vishwamitra, Nishant, Chiang, Cho-Yu Jason, Najafirad, Peyman |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Enhancing Event Reasoning in Large Language Models through Instruction Fine-Tuning with Semantic Causal Graphs
by: Bethany, Mazal, et al.
Published: (2024)
by: Bethany, Mazal, et al.
Published: (2024)
Deciphering Textual Authenticity: A Generalized Strategy through the Lens of Large Language Semantics for Detecting Human vs. Machine-Generated Text
by: Bethany, Mazal, et al.
Published: (2024)
by: Bethany, Mazal, et al.
Published: (2024)
Image Safeguarding: Reasoning with Conditional Vision Language Model and Obfuscating Unsafe Content Counterfactually
by: Bethany, Mazal, et al.
Published: (2024)
by: Bethany, Mazal, et al.
Published: (2024)
Lateral Phishing With Large Language Models: A Large Organization Comparative Study
by: Bethany, Mazal, et al.
Published: (2024)
by: Bethany, Mazal, et al.
Published: (2024)
Reflective Agreement: Combining Self-Mixture of Agents with a Sequence Tagger for Robust Event Extraction
by: Haji, Fatemeh, et al.
Published: (2025)
by: Haji, Fatemeh, et al.
Published: (2025)
Agentic Adversarial Rewriting Exposes Architectural Vulnerabilities in Black-Box NLP Pipelines
by: Bethany, Mazal, et al.
Published: (2026)
by: Bethany, Mazal, et al.
Published: (2026)
Can Reinforcement Learning Unlock the Hidden Dangers in Aligned Large Language Models?
by: Karkevandi, Mohammad Bahrami, et al.
Published: (2024)
by: Karkevandi, Mohammad Bahrami, et al.
Published: (2024)
Jailbreaking Large Language Models with Symbolic Mathematics
by: Bethany, Emet, et al.
Published: (2024)
by: Bethany, Emet, et al.
Published: (2024)
Improving LLM Reasoning with Multi-Agent Tree-of-Thought Validator Agent
by: Haji, Fatemeh, et al.
Published: (2024)
by: Haji, Fatemeh, et al.
Published: (2024)
RASTeR: Robust, Agentic, and Structured Temporal Reasoning
by: Schumacher, Dan, et al.
Published: (2024)
by: Schumacher, Dan, et al.
Published: (2024)
Robustness of Misinformation Classification Systems to Adversarial Examples Through BeamAttack
by: Fazla, Arnisa, et al.
Published: (2025)
by: Fazla, Arnisa, et al.
Published: (2025)
Unintentional Security Flaws in Code: Automated Defense via Root Cause Analysis
by: Islam, Nafis Tanveer, et al.
Published: (2024)
by: Islam, Nafis Tanveer, et al.
Published: (2024)
Toward Verifiable Misinformation Detection: A Multi-Tool LLM Agent Framework
by: Cui, Zikun, et al.
Published: (2025)
by: Cui, Zikun, et al.
Published: (2025)
Same Claim, Different Judgment: Benchmarking Scenario-Induced Bias in Multilingual Financial Misinformation Detection
by: Liu, Zhiwei, et al.
Published: (2026)
by: Liu, Zhiwei, et al.
Published: (2026)
Attacking Misinformation Detection Using Adversarial Examples Generated by Language Models
by: Przybyła, Piotr, et al.
Published: (2024)
by: Przybyła, Piotr, et al.
Published: (2024)
Combating Biomedical Misinformation through Multi-modal Claim Detection and Evidence-based Verification
by: Barone, Mariano, et al.
Published: (2025)
by: Barone, Mariano, et al.
Published: (2025)
An Investigation of Large Language Models for Real-World Hate Speech Detection
by: Guo, Keyan, et al.
Published: (2024)
by: Guo, Keyan, et al.
Published: (2024)
Robust Claim Verification Through Fact Detection
by: Jafari, Nazanin, et al.
Published: (2024)
by: Jafari, Nazanin, et al.
Published: (2024)
Unmasking Digital Falsehoods: A Comparative Analysis of LLM-Based Misinformation Detection Strategies
by: Huang, Tianyi, et al.
Published: (2025)
by: Huang, Tianyi, et al.
Published: (2025)
DELL: Generating Reactions and Explanations for LLM-Based Misinformation Detection
by: Wan, Herun, et al.
Published: (2024)
by: Wan, Herun, et al.
Published: (2024)
MisinfoTeleGraph: Network-driven Misinformation Detection for German Telegram Messages
by: Kalkbrenner, Lu, et al.
Published: (2025)
by: Kalkbrenner, Lu, et al.
Published: (2025)
Emotion Detection for Misinformation: A Review
by: Liu, Zhiwei, et al.
Published: (2023)
by: Liu, Zhiwei, et al.
Published: (2023)
Can LLM-Generated Misinformation Be Detected?
by: Chen, Canyu, et al.
Published: (2023)
by: Chen, Canyu, et al.
Published: (2023)
Fighting Fire with Fire: Adversarial Prompting to Generate a Misinformation Detection Dataset
by: Satapara, Shrey, et al.
Published: (2024)
by: Satapara, Shrey, et al.
Published: (2024)
Unraveling Misinformation Propagation in LLM Reasoning
by: Feng, Yiyang, et al.
Published: (2025)
by: Feng, Yiyang, et al.
Published: (2025)
When Claims Evolve: Evaluating and Enhancing the Robustness of Embedding Models Against Misinformation Edits
by: Magomere, Jabez, et al.
Published: (2025)
by: Magomere, Jabez, et al.
Published: (2025)
OUTFOX: LLM-Generated Essay Detection Through In-Context Learning with Adversarially Generated Examples
by: Koike, Ryuto, et al.
Published: (2023)
by: Koike, Ryuto, et al.
Published: (2023)
SeQwen at the Financial Misinformation Detection Challenge Task: Sequential Learning for Claim Verification and Explanation Generation in Financial Domains
by: Purbey, Jebish, et al.
Published: (2024)
by: Purbey, Jebish, et al.
Published: (2024)
LLM-driven Constrained Copy Generation through Iterative Refinement
by: Vasudevan, Varun, et al.
Published: (2025)
by: Vasudevan, Varun, et al.
Published: (2025)
Merging Facts, Crafting Fallacies: Evaluating the Contradictory Nature of Aggregated Factual Claims in Long-Form Generations
by: Chiang, Cheng-Han, et al.
Published: (2024)
by: Chiang, Cheng-Han, et al.
Published: (2024)
From Insight to Exploit: Leveraging LLM Collaboration for Adaptive Adversarial Text Generation
by: Sultana, Najrin, et al.
Published: (2025)
by: Sultana, Najrin, et al.
Published: (2025)
ContextClaim: A Context-Driven Paradigm for Verifiable Claim Detection
by: Li, Yufeng, et al.
Published: (2026)
by: Li, Yufeng, et al.
Published: (2026)
Ontology Enhanced Claim Detection
by: Hüsünbeyi, Zehra Melce, et al.
Published: (2024)
by: Hüsünbeyi, Zehra Melce, et al.
Published: (2024)
IAI Group at CheckThat! 2024: Transformer Models and Data Augmentation for Checkworthy Claim Detection
by: Aarnes, Peter Røysland, et al.
Published: (2024)
by: Aarnes, Peter Røysland, et al.
Published: (2024)
HiEAG: Evidence-Augmented Generation for Out-of-Context Misinformation Detection
by: Wu, Junjie, et al.
Published: (2025)
by: Wu, Junjie, et al.
Published: (2025)
Exploring the Impact of Instruction-Tuning on LLM's Susceptibility to Misinformation
by: Han, Kyubeen, et al.
Published: (2025)
by: Han, Kyubeen, et al.
Published: (2025)
Moderating New Waves of Online Hate with Chain-of-Thought Reasoning in Large Language Models
by: Vishwamitra, Nishant, et al.
Published: (2023)
by: Vishwamitra, Nishant, et al.
Published: (2023)
Supporting Automated Fact-checking across Topics: Similarity-driven Gradual Topic Learning for Claim Detection
by: Abumansour, Amani S., et al.
Published: (2024)
by: Abumansour, Amani S., et al.
Published: (2024)
Uncertainty Resolution in Misinformation Detection
by: Orlovskiy, Yury, et al.
Published: (2024)
by: Orlovskiy, Yury, et al.
Published: (2024)
ReviewScore: Misinformed Peer Review Detection with Large Language Models
by: Ryu, Hyun, et al.
Published: (2025)
by: Ryu, Hyun, et al.
Published: (2025)
Similar Items
-
Enhancing Event Reasoning in Large Language Models through Instruction Fine-Tuning with Semantic Causal Graphs
by: Bethany, Mazal, et al.
Published: (2024) -
Deciphering Textual Authenticity: A Generalized Strategy through the Lens of Large Language Semantics for Detecting Human vs. Machine-Generated Text
by: Bethany, Mazal, et al.
Published: (2024) -
Image Safeguarding: Reasoning with Conditional Vision Language Model and Obfuscating Unsafe Content Counterfactually
by: Bethany, Mazal, et al.
Published: (2024) -
Lateral Phishing With Large Language Models: A Large Organization Comparative Study
by: Bethany, Mazal, et al.
Published: (2024) -
Reflective Agreement: Combining Self-Mixture of Agents with a Sequence Tagger for Robust Event Extraction
by: Haji, Fatemeh, et al.
Published: (2025)