Hallucinating AI Hijacking Attack: Large Language Models and Malicious Code Recommenders
Fuente:
arXiv
Saved in:
| Main Authors: | Noever, David, McKee, Forrest |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Infecting Generative AI With Viruses
by: Noever, David, et al.
Published: (2025)
by: Noever, David, et al.
Published: (2025)
Transparency Attacks: How Imperceptible Image Layers Can Fool AI Perception
by: McKee, Forrest, et al.
Published: (2024)
by: McKee, Forrest, et al.
Published: (2024)
Safeguarding Voice Privacy: Harnessing Near-Ultrasonic Interference To Protect Against Unauthorized Audio Recording
by: McKee, Forrest, et al.
Published: (2024)
by: McKee, Forrest, et al.
Published: (2024)
Favicon Trojans: Executable Steganography Via Ico Alpha Channel Exploitation
by: Noever, David, et al.
Published: (2025)
by: Noever, David, et al.
Published: (2025)
Language Models And A Second Opinion Use Case: The Pocket Professional
by: Noever, David
Published: (2024)
by: Noever, David
Published: (2024)
HijackRAG: Hijacking Attacks against Retrieval-Augmented Large Language Models
by: Zhang, Yucheng, et al.
Published: (2024)
by: Zhang, Yucheng, et al.
Published: (2024)
Servant, Stalker, Predator: How An Honest, Helpful, And Harmless (3H) Agent Unlocks Adversarial Skills
by: Noever, David
Published: (2025)
by: Noever, David
Published: (2025)
Make Split, not Hijack: Preventing Feature-Space Hijacking Attacks in Split Learning
by: Khan, Tanveer, et al.
Published: (2024)
by: Khan, Tanveer, et al.
Published: (2024)
Forbidden Science: Dual-Use AI Challenge Benchmark and Scientific Refusal Tests
by: Noever, David, et al.
Published: (2025)
by: Noever, David, et al.
Published: (2025)
Vocabulary Attack to Hijack Large Language Model Applications
by: Levi, Patrick, et al.
Published: (2024)
by: Levi, Patrick, et al.
Published: (2024)
Can AI Freelancers Compete? Benchmarking Earnings, Reliability, and Task Success at Scale
by: Noever, David, et al.
Published: (2025)
by: Noever, David, et al.
Published: (2025)
Leveraging Large Language Models to Detect npm Malicious Packages
by: Zahan, Nusrat, et al.
Published: (2024)
by: Zahan, Nusrat, et al.
Published: (2024)
ImportSnare: Directed "Code Manual" Hijacking in Retrieval-Augmented Code Generation
by: Ye, Kai, et al.
Published: (2025)
by: Ye, Kai, et al.
Published: (2025)
Prompt-in-Content Attacks: Exploiting Uploaded Inputs to Hijack LLM Behavior
by: Lian, Zhuotao, et al.
Published: (2025)
by: Lian, Zhuotao, et al.
Published: (2025)
Hijacking Agent Memory: Stealthy Trojan Attacks Through Conversational Interaction
by: Wang, Hongtao, et al.
Published: (2026)
by: Wang, Hongtao, et al.
Published: (2026)
The Impossible Test: A 2024 Unsolvable Dataset and A Chance for an AGI Quiz
by: Noever, David, et al.
Published: (2024)
by: Noever, David, et al.
Published: (2024)
AirTag, You're It: Reverse Logistics and Last Mile Dynamics
by: Noever, David, et al.
Published: (2025)
by: Noever, David, et al.
Published: (2025)
Malla: Demystifying Real-world Large Language Model Integrated Malicious Services
by: Lin, Zilong, et al.
Published: (2024)
by: Lin, Zilong, et al.
Published: (2024)
Blind PRNG Hijacking: An Undetectable Integrity-Preserving Attack Against LLM Watermarking
by: You, Ziyang, et al.
Published: (2026)
by: You, Ziyang, et al.
Published: (2026)
Moravec's Paradox: Towards an Auditory Turing Test
by: Noever, David, et al.
Published: (2025)
by: Noever, David, et al.
Published: (2025)
Hijacking Large Audio-Language Models via Context-Agnostic and Imperceptible Auditory Prompt Injection
by: Chen, Meng, et al.
Published: (2026)
by: Chen, Meng, et al.
Published: (2026)
Invisible Prompts, Visible Threats: Malicious Font Injection in External Resources for Large Language Models
by: Xiong, Junjie, et al.
Published: (2025)
by: Xiong, Junjie, et al.
Published: (2025)
Predictive Coding and Information Bottleneck for Hallucination Detection in Large Language Models
by: Bhatt, Manish
Published: (2026)
by: Bhatt, Manish
Published: (2026)
Hallucination-Resistant Security Planning with a Large Language Model
by: Hammar, Kim, et al.
Published: (2026)
by: Hammar, Kim, et al.
Published: (2026)
Hide Your Malicious Goal Into Benign Narratives: Jailbreak Large Language Models through Carrier Articles
by: Wang, Zhilong, et al.
Published: (2024)
by: Wang, Zhilong, et al.
Published: (2024)
Hidden You Malicious Goal Into Benign Narratives: Jailbreak Large Language Models through Logic Chain Injection
by: Wang, Zhilong, et al.
Published: (2024)
by: Wang, Zhilong, et al.
Published: (2024)
CrossGuard: Safeguarding MLLMs against Joint-Modal Implicit Malicious Attacks
by: Zhang, Xu, et al.
Published: (2025)
by: Zhang, Xu, et al.
Published: (2025)
A Survey of Attacks on Large Language Models
by: Xu, Wenrui, et al.
Published: (2025)
by: Xu, Wenrui, et al.
Published: (2025)
Breaking MCP with Function Hijacking Attacks: Novel Threats for Function Calling and Agentic Models
by: Belkhiter, Yannis, et al.
Published: (2026)
by: Belkhiter, Yannis, et al.
Published: (2026)
CHAI: Command Hijacking against embodied AI
by: Burbano, Luis, et al.
Published: (2025)
by: Burbano, Luis, et al.
Published: (2025)
Membership Inference Attacks on Tokenizers of Large Language Models
by: Tong, Meng, et al.
Published: (2025)
by: Tong, Meng, et al.
Published: (2025)
A Cross-Language Investigation into Jailbreak Attacks in Large Language Models
by: Li, Jie, et al.
Published: (2024)
by: Li, Jie, et al.
Published: (2024)
Incident Response Planning Using a Lightweight Large Language Model with Reduced Hallucination
by: Hammar, Kim, et al.
Published: (2025)
by: Hammar, Kim, et al.
Published: (2025)
Breaking the Code: Security Assessment of AI Code Agents Through Systematic Jailbreaking Attacks
by: Saha, Shoumik, et al.
Published: (2025)
by: Saha, Shoumik, et al.
Published: (2025)
Persona Attack: Incremental Memory Injection Jailbreak Attack against Large Language Models
by: Park, Junyoung, et al.
Published: (2026)
by: Park, Junyoung, et al.
Published: (2026)
Automatically Generating Rules of Malicious Software Packages via Large Language Model
by: Zhang, XiangRui, et al.
Published: (2025)
by: Zhang, XiangRui, et al.
Published: (2025)
Recent Advances in Attack and Defense Approaches of Large Language Models
by: Cui, Jing, et al.
Published: (2024)
by: Cui, Jing, et al.
Published: (2024)
Chain-of-Scrutiny: Detecting Backdoor Attacks for Large Language Models
by: Li, Xi, et al.
Published: (2024)
by: Li, Xi, et al.
Published: (2024)
Attacking LLMs and AI Agents: Advertisement Embedding Attacks Against Large Language Models
by: Guo, Qiming, et al.
Published: (2025)
by: Guo, Qiming, et al.
Published: (2025)
CORVUS: Red-Teaming Hallucination Detectors via Internal Signal Camouflage in Large Language Models
by: Min, Nay Myat, et al.
Published: (2026)
by: Min, Nay Myat, et al.
Published: (2026)
Similar Items
-
Infecting Generative AI With Viruses
by: Noever, David, et al.
Published: (2025) -
Transparency Attacks: How Imperceptible Image Layers Can Fool AI Perception
by: McKee, Forrest, et al.
Published: (2024) -
Safeguarding Voice Privacy: Harnessing Near-Ultrasonic Interference To Protect Against Unauthorized Audio Recording
by: McKee, Forrest, et al.
Published: (2024) -
Favicon Trojans: Executable Steganography Via Ico Alpha Channel Exploitation
by: Noever, David, et al.
Published: (2025) -
Language Models And A Second Opinion Use Case: The Pocket Professional
by: Noever, David
Published: (2024)