Are My Optimized Prompts Compromised? Exploring Vulnerabilities of LLM-based Optimizers
Fuente:
arXiv
Salvato in:
| Autori principali: | Zhao, Andrew, Ghosh, Reshmi, Carvalho, Vitor, Lawton, Emily, Hines, Keegan, Huang, Gao, Stokes, Jack W. |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
VLMGuard: Defending VLMs against Malicious Prompts via Unlabeled Data
di: Du, Xuefeng, et al.
Pubblicazione: (2024)
di: Du, Xuefeng, et al.
Pubblicazione: (2024)
Never Compromise to Vulnerabilities: A Comprehensive Survey on AI Governance
di: Jiang, Yuchu, et al.
Pubblicazione: (2025)
di: Jiang, Yuchu, et al.
Pubblicazione: (2025)
The Vulnerability of LLM Rankers to Prompt Injection Attacks
di: Yin, Yu, et al.
Pubblicazione: (2026)
di: Yin, Yu, et al.
Pubblicazione: (2026)
Optimization-based Prompt Injection Attack to LLM-as-a-Judge
di: Shi, Jiawen, et al.
Pubblicazione: (2024)
di: Shi, Jiawen, et al.
Pubblicazione: (2024)
Defending Against Indirect Prompt Injection Attacks With Spotlighting
di: Hines, Keegan, et al.
Pubblicazione: (2024)
di: Hines, Keegan, et al.
Pubblicazione: (2024)
The Trigger in the Haystack: Extracting and Reconstructing LLM Backdoor Triggers
di: Bullwinkel, Blake, et al.
Pubblicazione: (2026)
di: Bullwinkel, Blake, et al.
Pubblicazione: (2026)
Token by Token, Compromised: Backdoor Vulnerabilities in Unified Autoregressive Models
di: Braun, Tobias, et al.
Pubblicazione: (2026)
di: Braun, Tobias, et al.
Pubblicazione: (2026)
SecureForge: Finding and Preventing Vulnerabilities in LLM-Generated Code via Prompt Optimization
di: Liu, Houjun, et al.
Pubblicazione: (2026)
di: Liu, Houjun, et al.
Pubblicazione: (2026)
Joint Optimization of Prompt Security and System Performance in Edge-Cloud LLM Systems
di: Huang, Haiyang, et al.
Pubblicazione: (2025)
di: Huang, Haiyang, et al.
Pubblicazione: (2025)
Prompt Optimization and Evaluation for LLM Automated Red Teaming
di: Freenor, Michael, et al.
Pubblicazione: (2025)
di: Freenor, Michael, et al.
Pubblicazione: (2025)
Fun-tuning: Characterizing the Vulnerability of Proprietary LLMs to Optimization-based Prompt Injection Attacks via the Fine-Tuning Interface
di: Labunets, Andrey, et al.
Pubblicazione: (2025)
di: Labunets, Andrey, et al.
Pubblicazione: (2025)
AgentShield: Deception-based Compromise Detection for Tool-using LLM Agents
di: Rassul, Yassin H., et al.
Pubblicazione: (2026)
di: Rassul, Yassin H., et al.
Pubblicazione: (2026)
Propagating Unsafe Actions in LLM Controlled Multi-Robot Collaboration via Single Robot Compromise
di: Huang, Zhen, et al.
Pubblicazione: (2026)
di: Huang, Zhen, et al.
Pubblicazione: (2026)
Prompt Engineering vs. Fine-Tuning for LLM-Based Vulnerability Detection in Solana and Algorand Smart Contracts
di: Boi, Biagio, et al.
Pubblicazione: (2025)
di: Boi, Biagio, et al.
Pubblicazione: (2025)
Who Moved My Transaction? Uncovering Post-Transaction Auditability Vulnerabilities in Modern Super Apps
di: Liu, Junlin, et al.
Pubblicazione: (2025)
di: Liu, Junlin, et al.
Pubblicazione: (2025)
VulBinLLM: LLM-powered Vulnerability Detection for Stripped Binaries
di: Hussain, Nasir, et al.
Pubblicazione: (2025)
di: Hussain, Nasir, et al.
Pubblicazione: (2025)
FlowSteer: Prompt-Only Workflow Steering Exposes Planning-Time Vulnerabilities in Multi-Agent LLM Systems
di: Li, Fanxiao, et al.
Pubblicazione: (2026)
di: Li, Fanxiao, et al.
Pubblicazione: (2026)
Evil Vizier: Vulnerabilities of LLM-Integrated XR Systems
di: Zhang, Yicheng, et al.
Pubblicazione: (2025)
di: Zhang, Yicheng, et al.
Pubblicazione: (2025)
VULSOLVER: Vulnerability Detection via LLM-Driven Constraint Solving
di: Li, Xiang, et al.
Pubblicazione: (2025)
di: Li, Xiang, et al.
Pubblicazione: (2025)
SAGE: Signal-Amplified Guided Embeddings for LLM-based Vulnerability Detection
di: Shan, Zhengyang, et al.
Pubblicazione: (2026)
di: Shan, Zhengyang, et al.
Pubblicazione: (2026)
CLASP: Cost-Optimized LLM-based Agentic System for Phishing Detection
di: Trad, Fouad, et al.
Pubblicazione: (2025)
di: Trad, Fouad, et al.
Pubblicazione: (2025)
Benchmarking LLMs and LLM-based Agents in Practical Vulnerability Detection for Code Repositories
di: Yildiz, Alperen, et al.
Pubblicazione: (2025)
di: Yildiz, Alperen, et al.
Pubblicazione: (2025)
JailGuard: A Universal Detection Framework for LLM Prompt-based Attacks
di: Zhang, Xiaoyu, et al.
Pubblicazione: (2023)
di: Zhang, Xiaoyu, et al.
Pubblicazione: (2023)
Tailored Prompts, Targeted Protection: Vulnerability-Specific LLM Analysis for Smart Contracts
di: Zhang, Xing, et al.
Pubblicazione: (2026)
di: Zhang, Xing, et al.
Pubblicazione: (2026)
ASPI: Seeking Ambiguity Clarification Amplifies Prompt Injection Vulnerability in LLM Agents
di: Sehwag, Udari Madhushani, et al.
Pubblicazione: (2026)
di: Sehwag, Udari Madhushani, et al.
Pubblicazione: (2026)
Breaking Agents: Compromising Autonomous LLM Agents Through Malfunction Amplification
di: Zhang, Boyang, et al.
Pubblicazione: (2024)
di: Zhang, Boyang, et al.
Pubblicazione: (2024)
Yama: Precise Opcode-based Data Flow Analysis for Detecting PHP Applications Vulnerabilities
di: Jiazhen, Zhao, et al.
Pubblicazione: (2024)
di: Jiazhen, Zhao, et al.
Pubblicazione: (2024)
Exploring Vulnerabilities and Concerns in Solana Smart Contracts
di: Wu, Xiangfan, et al.
Pubblicazione: (2025)
di: Wu, Xiangfan, et al.
Pubblicazione: (2025)
PSM: Prompt Sensitivity Minimization via LLM-Guided Black-Box Optimization
di: Jawad, Huseein, et al.
Pubblicazione: (2025)
di: Jawad, Huseein, et al.
Pubblicazione: (2025)
Framework and Classification of Indicator of Compromise for physics-based attacks
di: Tan, Vincent
Pubblicazione: (2024)
di: Tan, Vincent
Pubblicazione: (2024)
Public-key encryption from a trapdoor one-way embedding of $SL_2(\mathbb{N}$)
di: Hines, Robert
Pubblicazione: (2024)
di: Hines, Robert
Pubblicazione: (2024)
Boosting Cybersecurity Vulnerability Scanning based on LLM-supported Static Application Security Testing
di: Keltek, Mete, et al.
Pubblicazione: (2024)
di: Keltek, Mete, et al.
Pubblicazione: (2024)
VULPO: Context-Aware Vulnerability Detection via On-Policy LLM Optimization
di: Li, Youpeng, et al.
Pubblicazione: (2025)
di: Li, Youpeng, et al.
Pubblicazione: (2025)
PathSeeker: Exploring LLM Security Vulnerabilities with a Reinforcement Learning-Based Jailbreak Approach
di: Lin, Zhihao, et al.
Pubblicazione: (2024)
di: Lin, Zhihao, et al.
Pubblicazione: (2024)
Demystifying RCE Vulnerabilities in LLM-Integrated Apps
di: Liu, Tong, et al.
Pubblicazione: (2023)
di: Liu, Tong, et al.
Pubblicazione: (2023)
All You Need Is A Fuzzing Brain: An LLM-Powered System for Automated Vulnerability Detection and Patching
di: Sheng, Ze, et al.
Pubblicazione: (2025)
di: Sheng, Ze, et al.
Pubblicazione: (2025)
The Power of Bamboo: On the Post-Compromise Security for Searchable Symmetric Encryption
di: Chen, Tianyang, et al.
Pubblicazione: (2024)
di: Chen, Tianyang, et al.
Pubblicazione: (2024)
Quantifying Memory Cells Vulnerability for DRAM Security
di: Hu, Zilong, et al.
Pubblicazione: (2026)
di: Hu, Zilong, et al.
Pubblicazione: (2026)
LogicScan: An LLM-driven Framework for Detecting Business Logic Vulnerabilities in Smart Contracts
di: Gao, Jiaqi, et al.
Pubblicazione: (2026)
di: Gao, Jiaqi, et al.
Pubblicazione: (2026)
RLCracker: Evaluating the Worst-Case Vulnerability of LLM Watermarks with Adaptive RL Attacks
di: Huang, Hanbo, et al.
Pubblicazione: (2025)
di: Huang, Hanbo, et al.
Pubblicazione: (2025)
Documenti analoghi
-
VLMGuard: Defending VLMs against Malicious Prompts via Unlabeled Data
di: Du, Xuefeng, et al.
Pubblicazione: (2024) -
Never Compromise to Vulnerabilities: A Comprehensive Survey on AI Governance
di: Jiang, Yuchu, et al.
Pubblicazione: (2025) -
The Vulnerability of LLM Rankers to Prompt Injection Attacks
di: Yin, Yu, et al.
Pubblicazione: (2026) -
Optimization-based Prompt Injection Attack to LLM-as-a-Judge
di: Shi, Jiawen, et al.
Pubblicazione: (2024) -
Defending Against Indirect Prompt Injection Attacks With Spotlighting
di: Hines, Keegan, et al.
Pubblicazione: (2024)