Imitation Game for Adversarial Disillusion with Chain-of-Thought Reasoning in Generative AI
Fuente:
arXiv
Saved in:
| Main Authors: | Chang, Ching-Chun, Chen, Fan-Yun, Gu, Shih-Hong, Gao, Kai, Wang, Hanrui, Echizen, Isao |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Agentic Copyright Watermarking against Adversarial Evidence Forgery with Purification-Agnostic Curriculum Proxy Learning
by: Bao, Erjin, et al.
Published: (2024)
by: Bao, Erjin, et al.
Published: (2024)
Tell-Tale Watermarks for Explanatory Reasoning in Synthetic Media Forensics
by: Chang, Ching-Chun, et al.
Published: (2025)
by: Chang, Ching-Chun, et al.
Published: (2025)
GreedyPixel: Fine-Grained Black-Box Adversarial Attack Via Greedy Algorithm
by: Wang, Hanrui, et al.
Published: (2025)
by: Wang, Hanrui, et al.
Published: (2025)
Steganography Beyond Space-Time with Chain of Multimodal AI
by: Chang, Ching-Chun, et al.
Published: (2025)
by: Chang, Ching-Chun, et al.
Published: (2025)
Steganography in Game Actions
by: Chang, Ching-Chun, et al.
Published: (2024)
by: Chang, Ching-Chun, et al.
Published: (2024)
Minimal Cascade Gradient Smoothing for Fast Transferable Preemptive Adversarial Defense
by: Wang, Hanrui, et al.
Published: (2024)
by: Wang, Hanrui, et al.
Published: (2024)
Chronology of Multi-Agent Interactions for Provenance of Evolving Information
by: Chang, Ching-Chun, et al.
Published: (2025)
by: Chang, Ching-Chun, et al.
Published: (2025)
DiffMI: Breaking Face Recognition Privacy via Diffusion-Driven Training-Free Model Inversion
by: Wang, Hanrui, et al.
Published: (2025)
by: Wang, Hanrui, et al.
Published: (2025)
On the Origin of Synthetic Information by Means of Steganographic Inheritance
by: Chang, Ching-Chun, et al.
Published: (2026)
by: Chang, Ching-Chun, et al.
Published: (2026)
Cyber-Physical Steganography in Robotic Motion Control
by: Chang, Ching-Chun, et al.
Published: (2025)
by: Chang, Ching-Chun, et al.
Published: (2025)
Cyber Vaccine for Deepfake Immunity
by: Chang, Ching-Chun, et al.
Published: (2023)
by: Chang, Ching-Chun, et al.
Published: (2023)
Hypnopaedia-Aware Machine Unlearning via Psychometrics of Artificial Mental Imagery
by: Chang, Ching-Chun, et al.
Published: (2024)
by: Chang, Ching-Chun, et al.
Published: (2024)
Stop Reasoning! When Multimodal LLM with Chain-of-Thought Reasoning Meets Adversarial Image
by: Wang, Zefeng, et al.
Published: (2024)
by: Wang, Zefeng, et al.
Published: (2024)
Low-Cost Hard-Label Adversarial Attack with Theoretical Foundations
by: Liu, Jun, et al.
Published: (2026)
by: Liu, Jun, et al.
Published: (2026)
Mitigating Backdoor Attacks using Activation-Guided Model Editing
by: Hsieh, Felix, et al.
Published: (2024)
by: Hsieh, Felix, et al.
Published: (2024)
A Multi-task Adversarial Attack Against Face Authentication
by: Wang, Hanrui, et al.
Published: (2024)
by: Wang, Hanrui, et al.
Published: (2024)
The Blockchain Imitation Game
by: Qin, Kaihua, et al.
Published: (2023)
by: Qin, Kaihua, et al.
Published: (2023)
ExplainableGuard: Interpretable Adversarial Defense for Large Language Models Using Chain-of-Thought Reasoning
by: Guan, Shaowei, et al.
Published: (2025)
by: Guan, Shaowei, et al.
Published: (2025)
Efficient and Privacy-Preserving Federated Learning based on Full Homomorphic Encryption
by: Guo, Yuqi, et al.
Published: (2024)
by: Guo, Yuqi, et al.
Published: (2024)
Thought-Transfer: Indirect Targeted Poisoning Attacks on Chain-of-Thought Reasoning Models
by: Chaudhari, Harsh, et al.
Published: (2026)
by: Chaudhari, Harsh, et al.
Published: (2026)
Unreal Thinking: Chain-of-Thought Hijacking via Two-stage Backdoor
by: Chang, Wenhan, et al.
Published: (2026)
by: Chang, Wenhan, et al.
Published: (2026)
SwitchPatch: Physical Adversarial Attack Strategy with Switchable Adversarial Objectives
by: Jiang, Hanrui, et al.
Published: (2025)
by: Jiang, Hanrui, et al.
Published: (2025)
From Thinking to Output: Chain-of-Thought and Text Generation Characteristics in Reasoning Language Models
by: Liu, Junhao, et al.
Published: (2025)
by: Liu, Junhao, et al.
Published: (2025)
The Imitation Game: Using Large Language Models as Chatbots to Combat Chat-Based Cybercrimes
by: Yao, Yifan, et al.
Published: (2025)
by: Yao, Yifan, et al.
Published: (2025)
DiffProtect: Generate Adversarial Examples with Diffusion Models for Facial Privacy Protection
by: Liu, Jiang, et al.
Published: (2023)
by: Liu, Jiang, et al.
Published: (2023)
Natias: Neuron Attribution based Transferable Image Adversarial Steganography
by: Fan, Zexin, et al.
Published: (2024)
by: Fan, Zexin, et al.
Published: (2024)
Can Reasoning Models Obfuscate Reasoning? Stress-Testing Chain-of-Thought Monitorability
by: Zolkowski, Artur, et al.
Published: (2025)
by: Zolkowski, Artur, et al.
Published: (2025)
R-CoT: A Reasoning-Layer Watermark via Redundant Chain-of-Thought in Large Language Models
by: Zhang, Ziming, et al.
Published: (2026)
by: Zhang, Ziming, et al.
Published: (2026)
Echoes within the Reasoning: Stealthy and Effective Watermarking via Chain of Thought
by: Lu, Jiacheng, et al.
Published: (2026)
by: Lu, Jiacheng, et al.
Published: (2026)
The Adversarial AI-Art: Understanding, Generation, Detection, and Benchmarking
by: Li, Yuying, et al.
Published: (2024)
by: Li, Yuying, et al.
Published: (2024)
Iterative Window Mean Filter: Thwarting Diffusion-based Adversarial Purification
by: Wang, Hanrui, et al.
Published: (2024)
by: Wang, Hanrui, et al.
Published: (2024)
SAP-DIFF: Semantic Adversarial Patch Generation for Black-Box Face Recognition Models via Diffusion Models
by: Wang, Mingsi, et al.
Published: (2025)
by: Wang, Mingsi, et al.
Published: (2025)
Chain-of-Code Collapse: Reasoning Failures in LLMs via Adversarial Prompting in Code Generation
by: Roh, Jaechul, et al.
Published: (2025)
by: Roh, Jaechul, et al.
Published: (2025)
Preemptive Answer "Attacks" on Chain-of-Thought Reasoning
by: Xu, Rongwu, et al.
Published: (2024)
by: Xu, Rongwu, et al.
Published: (2024)
Defensive Unlearning with Adversarial Training for Robust Concept Erasure in Diffusion Models
by: Zhang, Yimeng, et al.
Published: (2024)
by: Zhang, Yimeng, et al.
Published: (2024)
Vulnerabilities in AI-generated Image Detection: The Challenge of Adversarial Attacks
by: Diao, Yunfeng, et al.
Published: (2024)
by: Diao, Yunfeng, et al.
Published: (2024)
Defensive Adversarial CAPTCHA: A Semantics-Driven Framework for Natural Adversarial Example Generation
by: Du, Xia, et al.
Published: (2025)
by: Du, Xia, et al.
Published: (2025)
Facial Recognition Leveraging Generative Adversarial Networks
by: Li, Zhongwen, et al.
Published: (2025)
by: Li, Zhongwen, et al.
Published: (2025)
BadThink: Triggered Overthinking Attacks on Chain-of-Thought Reasoning in Large Language Models
by: Liu, Shuaitong, et al.
Published: (2025)
by: Liu, Shuaitong, et al.
Published: (2025)
Reasoning Under Pressure: How do Training Incentives Influence Chain-of-Thought Monitorability?
by: MacDermott, Matt, et al.
Published: (2025)
by: MacDermott, Matt, et al.
Published: (2025)
Similar Items
-
Agentic Copyright Watermarking against Adversarial Evidence Forgery with Purification-Agnostic Curriculum Proxy Learning
by: Bao, Erjin, et al.
Published: (2024) -
Tell-Tale Watermarks for Explanatory Reasoning in Synthetic Media Forensics
by: Chang, Ching-Chun, et al.
Published: (2025) -
GreedyPixel: Fine-Grained Black-Box Adversarial Attack Via Greedy Algorithm
by: Wang, Hanrui, et al.
Published: (2025) -
Steganography Beyond Space-Time with Chain of Multimodal AI
by: Chang, Ching-Chun, et al.
Published: (2025) -
Steganography in Game Actions
by: Chang, Ching-Chun, et al.
Published: (2024)