The Wolf Within: Covert Injection of Malice into MLLM Societies via an MLLM Operative
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Tan, Zhen, Zhao, Chengshuai, Moraffah, Raha, Li, Yifan, Kong, Yu, Chen, Tianlong, Liu, Huan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
"Glue pizza and eat rocks" -- Exploiting Vulnerabilities in Retrieval-Augmented Generative Models
von: Tan, Zhen, et al.
Veröffentlicht: (2024)
von: Tan, Zhen, et al.
Veröffentlicht: (2024)
A Generative Approach to Surrogate-based Black-box Attacks
von: Moraffah, Raha, et al.
Veröffentlicht: (2024)
von: Moraffah, Raha, et al.
Veröffentlicht: (2024)
Exploiting Class Probabilities for Black-box Sentence-level Attacks
von: Moraffah, Raha, et al.
Veröffentlicht: (2024)
von: Moraffah, Raha, et al.
Veröffentlicht: (2024)
Adversarial Text Purification: A Large Language Model Approach for Defense
von: Moraffah, Raha, et al.
Veröffentlicht: (2024)
von: Moraffah, Raha, et al.
Veröffentlicht: (2024)
Community Covert Communication - Dynamic Mass Covert Communication Through Social Media
von: Filiol, Eric
Veröffentlicht: (2025)
von: Filiol, Eric
Veröffentlicht: (2025)
Enhanced MLLM Black-Box Jailbreaking Attacks and Defenses
von: Zhong, Xingwei, et al.
Veröffentlicht: (2025)
von: Zhong, Xingwei, et al.
Veröffentlicht: (2025)
Medical Malice: A Dataset for Context-Aware Safety in Healthcare LLMs
von: D'addario, Andrew Maranhão Ventura
Veröffentlicht: (2025)
von: D'addario, Andrew Maranhão Ventura
Veröffentlicht: (2025)
Technical Report for ICML 2024 TiFA Workshop MLLM Attack Challenge: Suffix Injection and Projected Gradient Descent Can Easily Fool An MLLM
von: Guo, Yangyang, et al.
Veröffentlicht: (2024)
von: Guo, Yangyang, et al.
Veröffentlicht: (2024)
MLLM-Protector: Ensuring MLLM's Safety without Hurting Performance
von: Pi, Renjie, et al.
Veröffentlicht: (2024)
von: Pi, Renjie, et al.
Veröffentlicht: (2024)
DualTAP: A Dual-Task Adversarial Protector for Mobile MLLM Agents
von: Zhang, Fuyao, et al.
Veröffentlicht: (2025)
von: Zhang, Fuyao, et al.
Veröffentlicht: (2025)
Q-MLLM: Vector Quantization for Robust Multimodal Large Language Model Security
von: Zhao, Wei, et al.
Veröffentlicht: (2025)
von: Zhao, Wei, et al.
Veröffentlicht: (2025)
When Your Reviewer is an LLM: Biases, Divergence, and Prompt Injection Risks in Peer Review
von: Zhu, Changjia, et al.
Veröffentlicht: (2025)
von: Zhu, Changjia, et al.
Veröffentlicht: (2025)
Prompt Injection Vulnerability of Consensus Generating Applications in Digital Democracy
von: Gudiño-Rosero, Jairo, et al.
Veröffentlicht: (2025)
von: Gudiño-Rosero, Jairo, et al.
Veröffentlicht: (2025)
Covert Surveillance in Smart Devices: A SCOUR Framework Analysis of Youth Privacy Implications
von: Shouli, Austin, et al.
Veröffentlicht: (2025)
von: Shouli, Austin, et al.
Veröffentlicht: (2025)
To See is Not to Learn: Protecting Multimodal Data from Unauthorized Fine-Tuning of Large Vision-Language Model
von: Zhao, Chengshuai, et al.
Veröffentlicht: (2026)
von: Zhao, Chengshuai, et al.
Veröffentlicht: (2026)
Backdoor Cleaning without External Guidance in MLLM Fine-tuning
von: Rong, Xuankun, et al.
Veröffentlicht: (2025)
von: Rong, Xuankun, et al.
Veröffentlicht: (2025)
Analysing Multidisciplinary Approaches to Fight Large-Scale Digital Influence Operations
von: Arroyo, David, et al.
Veröffentlicht: (2025)
von: Arroyo, David, et al.
Veröffentlicht: (2025)
Blockchain-Anchored Audit Trail Model for Transparent Inter-Operator Settlement
von: Kunthu, Balakumar Ravindranath, et al.
Veröffentlicht: (2025)
von: Kunthu, Balakumar Ravindranath, et al.
Veröffentlicht: (2025)
AI Agents May Always Fall for Prompt Injections
von: Abdelnabi, Sahar, et al.
Veröffentlicht: (2026)
von: Abdelnabi, Sahar, et al.
Veröffentlicht: (2026)
ComMark: Covert and Robust Black-Box Model Watermarking with Compressed Samples
von: Yang, Yunfei, et al.
Veröffentlicht: (2025)
von: Yang, Yunfei, et al.
Veröffentlicht: (2025)
AmbShield: Enhancing Physical Layer Security with Ambient Backscatter Devices against Eavesdroppers
von: Zhang, Yifan, et al.
Veröffentlicht: (2026)
von: Zhang, Yifan, et al.
Veröffentlicht: (2026)
Mind the Third Eye! Benchmarking Privacy Awareness in MLLM-powered Smartphone Agents
von: Lin, Zhixin, et al.
Veröffentlicht: (2025)
von: Lin, Zhixin, et al.
Veröffentlicht: (2025)
A Survey of Operating System Kernel Fuzzing
von: Xu, Jiacheng, et al.
Veröffentlicht: (2025)
von: Xu, Jiacheng, et al.
Veröffentlicht: (2025)
Sleeper Cell: Injecting Latent Malice Temporal Backdoors into Tool-Using LLMs
von: Pallakonda, Bhanu, et al.
Veröffentlicht: (2026)
von: Pallakonda, Bhanu, et al.
Veröffentlicht: (2026)
A High-performance Real-time Container File Monitoring Approach Based on Virtual Machine Introspection
von: Tan, Kai, et al.
Veröffentlicht: (2025)
von: Tan, Kai, et al.
Veröffentlicht: (2025)
Careful About What App Promotion Ads Recommend! Detecting and Explaining Malware Promotion via App Promotion Graph
von: Ma, Shang, et al.
Veröffentlicht: (2024)
von: Ma, Shang, et al.
Veröffentlicht: (2024)
Infrastructure for Valuable, Tradable, and Verifiable Agent Memory
von: Li, Mengyuan, et al.
Veröffentlicht: (2026)
von: Li, Mengyuan, et al.
Veröffentlicht: (2026)
TEMPEST-LoRa: Cross-Technology Covert Communication
von: Sun, Xieyang, et al.
Veröffentlicht: (2025)
von: Sun, Xieyang, et al.
Veröffentlicht: (2025)
SDN-Based Dynamic Cybersecurity Framework of IEC-61850 Communications in Smart Grid
von: Girdhar, Mansi, et al.
Veröffentlicht: (2023)
von: Girdhar, Mansi, et al.
Veröffentlicht: (2023)
Reconnecting Citizens to Politics via Blockchain - Starting the Debate
von: Serdült, Uwe
Veröffentlicht: (2025)
von: Serdült, Uwe
Veröffentlicht: (2025)
SemFuzz: A Semantics-Aware Fuzzing Framework for Network Protocol Implementations
von: Sun, Yanbang, et al.
Veröffentlicht: (2026)
von: Sun, Yanbang, et al.
Veröffentlicht: (2026)
AI Safety vs. AI Security: Demystifying the Distinction and Boundaries
von: Lin, Zhiqiang, et al.
Veröffentlicht: (2025)
von: Lin, Zhiqiang, et al.
Veröffentlicht: (2025)
vApps: Verifiable Applications at Internet Scale
von: Zhang, Isaac, et al.
Veröffentlicht: (2025)
von: Zhang, Isaac, et al.
Veröffentlicht: (2025)
Measuring Real-World Prompt Injection Attacks in LLM-based Resume Screening
von: Zhang, Mohan, et al.
Veröffentlicht: (2026)
von: Zhang, Mohan, et al.
Veröffentlicht: (2026)
SQL Injection Jailbreak: A Structural Disaster of Large Language Models
von: Zhao, Jiawei, et al.
Veröffentlicht: (2024)
von: Zhao, Jiawei, et al.
Veröffentlicht: (2024)
Global Web, Local Privacy? An International Review of Web Tracking
von: Yu, Harry, et al.
Veröffentlicht: (2026)
von: Yu, Harry, et al.
Veröffentlicht: (2026)
Provenance Verification of AI-Generated Images via a Perceptual Hash Registry Anchored on Blockchain
von: Mohit, Apoorv, et al.
Veröffentlicht: (2026)
von: Mohit, Apoorv, et al.
Veröffentlicht: (2026)
Privacy Computing Meets Metaverse: Necessity, Taxonomy and Challenges
von: Chen, Chuan, et al.
Veröffentlicht: (2023)
von: Chen, Chuan, et al.
Veröffentlicht: (2023)
LightDefense: A Lightweight Uncertainty-Driven Defense against Jailbreaks via Shifted Token Distribution
von: Yang, Zhuoran, et al.
Veröffentlicht: (2025)
von: Yang, Zhuoran, et al.
Veröffentlicht: (2025)
Cyclic Adaptive Private Synthesis for Sharing Real-World Data in Education
von: Ito, Hibiki, et al.
Veröffentlicht: (2026)
von: Ito, Hibiki, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
"Glue pizza and eat rocks" -- Exploiting Vulnerabilities in Retrieval-Augmented Generative Models
von: Tan, Zhen, et al.
Veröffentlicht: (2024) -
A Generative Approach to Surrogate-based Black-box Attacks
von: Moraffah, Raha, et al.
Veröffentlicht: (2024) -
Exploiting Class Probabilities for Black-box Sentence-level Attacks
von: Moraffah, Raha, et al.
Veröffentlicht: (2024) -
Adversarial Text Purification: A Large Language Model Approach for Defense
von: Moraffah, Raha, et al.
Veröffentlicht: (2024) -
Community Covert Communication - Dynamic Mass Covert Communication Through Social Media
von: Filiol, Eric
Veröffentlicht: (2025)