Hidden in Plain Text: Emergence & Mitigation of Steganographic Collusion in LLMs
Fuente:
arXiv
Saved in:
| Main Authors: | Mathew, Yohan, Matthews, Ollie, McCarthy, Robert, Velja, Joan, de Witt, Christian Schroeder, Cope, Dylan, Schoots, Nandi |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Early Signs of Steganographic Capabilities in Frontier LLMs
by: Zolkowski, Artur, et al.
Published: (2025)
by: Zolkowski, Artur, et al.
Published: (2025)
Mitigating Collusion in Proofs of Liabilities
by: Mohamed, Malcom, et al.
Published: (2026)
by: Mohamed, Malcom, et al.
Published: (2026)
Synthetic Embedding of Hidden Information in Industrial Control System Network Protocols for Evaluation of Steganographic Malware
by: Neubert, Tom, et al.
Published: (2024)
by: Neubert, Tom, et al.
Published: (2024)
Secret Collusion among AI Agents: Multi-Agent Deception via Steganography
by: Motwani, Sumeet Ramesh, et al.
Published: (2024)
by: Motwani, Sumeet Ramesh, et al.
Published: (2024)
Hiding in Plain Sight: A Steganographic Approach to Stealthy LLM Jailbreaks
by: Geng, Jianing, et al.
Published: (2025)
by: Geng, Jianing, et al.
Published: (2025)
Safeguarding LLMs Against Misuse and AI-Driven Malware Using Steganographic Canaries
by: Raz, Md, et al.
Published: (2026)
by: Raz, Md, et al.
Published: (2026)
The Hidden Threat in Plain Text: Attacking RAG Data Loaders
by: Castagnaro, Alberto, et al.
Published: (2025)
by: Castagnaro, Alberto, et al.
Published: (2025)
MonitoringBench: Semi-Automated Red-Teaming for Agent Monitoring
by: Jotautaitė, Monika, et al.
Published: (2026)
by: Jotautaitė, Monika, et al.
Published: (2026)
Unraveling Log4Shell: Analyzing the Impact and Response to the Log4j Vulnerabil
by: Doll, John, et al.
Published: (2025)
by: Doll, John, et al.
Published: (2025)
NEST: Nascent Encoded Steganographic Thoughts
by: Karpov, Artem
Published: (2026)
by: Karpov, Artem
Published: (2026)
GIFDL: Generated Image Fluctuation Distortion Learning for Enhancing Steganographic Security
by: Wang, Xiangkun, et al.
Published: (2025)
by: Wang, Xiangkun, et al.
Published: (2025)
Hidden in Plain Sound: Environmental Backdoor Poisoning Attacks on Whisper, and Mitigations
by: Bartolini, Jonatan, et al.
Published: (2024)
by: Bartolini, Jonatan, et al.
Published: (2024)
Invisible Injections: Exploiting Vision-Language Models Through Steganographic Prompt Embedding
by: Pathade, Chetan
Published: (2025)
by: Pathade, Chetan
Published: (2025)
A Character-based Diffusion Embedding Algorithm for Enhancing the Generation Quality of Generative Linguistic Steganographic Texts
by: Chen, Yingquan, et al.
Published: (2025)
by: Chen, Yingquan, et al.
Published: (2025)
Adaptive Fuzzy Logic-Based Steganographic Encryption Framework: A Comprehensive Experimental Evaluation
by: Joshi, Aadi, et al.
Published: (2026)
by: Joshi, Aadi, et al.
Published: (2026)
A Public and Reproducible Assessment of the Topics API on Real Data
by: Beugin, Yohan, et al.
Published: (2024)
by: Beugin, Yohan, et al.
Published: (2024)
Technical Report: The Need for a (Research) Sandstorm through the Privacy Sandbox
by: Beugin, Yohan, et al.
Published: (2025)
by: Beugin, Yohan, et al.
Published: (2025)
Steganographic Embeddings as an Effective Data Augmentation
by: DiSalvo, Nicholas
Published: (2025)
by: DiSalvo, Nicholas
Published: (2025)
Computing Low-Entropy Couplings for Large-Support Distributions
by: Sokota, Samuel, et al.
Published: (2024)
by: Sokota, Samuel, et al.
Published: (2024)
The Steganographic Potentials of Language Models
by: Karpov, Artem, et al.
Published: (2025)
by: Karpov, Artem, et al.
Published: (2025)
Purified and Unified Steganographic Network
by: Li, Guobiao, et al.
Published: (2024)
by: Li, Guobiao, et al.
Published: (2024)
Hidden-in-Plain-Text: A Benchmark for Social-Web Indirect Prompt Injection in RAG
by: Guo, Haoze, et al.
Published: (2026)
by: Guo, Haoze, et al.
Published: (2026)
PKE and ABE with Collusion-Resistant Secure Key Leasing
by: Kitagawa, Fuyuki, et al.
Published: (2025)
by: Kitagawa, Fuyuki, et al.
Published: (2025)
Multichannel Steganography: A Provably Secure Hybrid Steganographic Model for Secure Communication
by: Omego, Obinna, et al.
Published: (2025)
by: Omego, Obinna, et al.
Published: (2025)
ADLM -- stega: A Universal Adaptive Token Selection Algorithm for Improving Steganographic Text Quality via Information Entropy
by: Qin, Zezheng, et al.
Published: (2024)
by: Qin, Zezheng, et al.
Published: (2024)
Extending the OWASP Multi-Agentic System Threat Modeling Guide: Insights from Multi-Agent Security Research
by: Krawiecka, Klaudia, et al.
Published: (2025)
by: Krawiecka, Klaudia, et al.
Published: (2025)
Unclonable Cryptography with Unbounded Collusions and Impossibility of Hyperefficient Shadow Tomography
by: Çakan, Alper, et al.
Published: (2023)
by: Çakan, Alper, et al.
Published: (2023)
Collusion-Driven Impersonation Attack on Channel-Resistant RF Fingerprinting
by: Xu, Zhou, et al.
Published: (2025)
by: Xu, Zhou, et al.
Published: (2025)
A Dual-Layer Image Encryption Framework Using Chaotic AES with Dynamic S-Boxes and Steganographic QR Codes
by: Bayesh, Md Rishadul, et al.
Published: (2025)
by: Bayesh, Md Rishadul, et al.
Published: (2025)
VET Your Agent: Towards Host-Independent Autonomy via Verifiable Execution Traces
by: Grigor, Artem, et al.
Published: (2025)
by: Grigor, Artem, et al.
Published: (2025)
BlackCATT: Black-box Collusion Aware Traitor Tracing in Federated Learning
by: Rodríguez-Lois, Elena, et al.
Published: (2026)
by: Rodríguez-Lois, Elena, et al.
Published: (2026)
A Collusion-Resistance Privacy-Preserving Smart Metering Protocol for Operational Utility
by: Zaredar, Farid, et al.
Published: (2025)
by: Zaredar, Farid, et al.
Published: (2025)
Vulnerability of Text-Matching in ML/AI Conference Reviewer Assignments to Collusions
by: Hsieh, Jhih-Yi, et al.
Published: (2024)
by: Hsieh, Jhih-Yi, et al.
Published: (2024)
Analysing India's Cyber Warfare Readiness and Developing a Defence Strategy
by: Fernandes, Yohan, et al.
Published: (2024)
by: Fernandes, Yohan, et al.
Published: (2024)
ADCA: Attention-Driven Multi-Party Collusion Attack in Federated Self-Supervised Learning
by: Wang, Jiayao, et al.
Published: (2026)
by: Wang, Jiayao, et al.
Published: (2026)
Targeting Alignment: Extracting Safety Classifiers of Aligned LLMs
by: Ferrand, Jean-Charles Noirot, et al.
Published: (2025)
by: Ferrand, Jean-Charles Noirot, et al.
Published: (2025)
Breaking Free: Efficient Multi-Party Private Set Union Without Non-Collusion Assumptions
by: Dong, Minglang, et al.
Published: (2024)
by: Dong, Minglang, et al.
Published: (2024)
On the Origin of Synthetic Information by Means of Steganographic Inheritance
by: Chang, Ching-Chun, et al.
Published: (2026)
by: Chang, Ching-Chun, et al.
Published: (2026)
TrojanStego: Your Language Model Can Secretly Be A Steganographic Privacy Leaking Agent
by: Meier, Dominik, et al.
Published: (2025)
by: Meier, Dominik, et al.
Published: (2025)
SAB:A Stealing and Robust Backdoor Attack based on Steganographic Algorithm against Federated Learning
by: Xu, Weida, et al.
Published: (2024)
by: Xu, Weida, et al.
Published: (2024)
Similar Items
-
Early Signs of Steganographic Capabilities in Frontier LLMs
by: Zolkowski, Artur, et al.
Published: (2025) -
Mitigating Collusion in Proofs of Liabilities
by: Mohamed, Malcom, et al.
Published: (2026) -
Synthetic Embedding of Hidden Information in Industrial Control System Network Protocols for Evaluation of Steganographic Malware
by: Neubert, Tom, et al.
Published: (2024) -
Secret Collusion among AI Agents: Multi-Agent Deception via Steganography
by: Motwani, Sumeet Ramesh, et al.
Published: (2024) -
Hiding in Plain Sight: A Steganographic Approach to Stealthy LLM Jailbreaks
by: Geng, Jianing, et al.
Published: (2025)