When Alignment Isn't Enough: Response-Path Attacks on LLM Agents
Fuente:
arXiv
Salvato in:
| Autori principali: | Luo, Mingyu, Zhang, Zihan, Liu, Zesen, Xie, Yuchong, Zhang, Zhixiang, Yeung, Dung Hiu Hilton, Lai, Wai Ip, Chen, Ping, Wen, Ming, She, Dongdong |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
CompressionAttack: Exploiting Prompt Compression as a New Attack Surface in LLM-Powered Agents
di: Liu, Zesen, et al.
Pubblicazione: (2025)
di: Liu, Zesen, et al.
Pubblicazione: (2025)
When Trust Isn't Enough.
di: Behrman, Sara
Pubblicazione: (1998)
di: Behrman, Sara
Pubblicazione: (1998)
From Similarity to Vulnerability: Key Collision Attack on LLM Semantic Caching
di: Zhang, Zhixiang, et al.
Pubblicazione: (2026)
di: Zhang, Zhixiang, et al.
Pubblicazione: (2026)
Knowing the Answer Isn't Enough: Fixing Reasoning Path Failures in LVLMs
di: Wang, Chaoyang, et al.
Pubblicazione: (2025)
di: Wang, Chaoyang, et al.
Pubblicazione: (2025)
When One Moment Isn't Enough: Multi-Moment Retrieval with Cross-Moment Interactions
di: Cao, Zhuo, et al.
Pubblicazione: (2025)
di: Cao, Zhuo, et al.
Pubblicazione: (2025)
When Standard Newborn Screening Isn't Enough: Diagnostic Challenges in the Age of Globalization
di: Marina Ortúzar Menéndez, et al.
Pubblicazione: (2026)
di: Marina Ortúzar Menéndez, et al.
Pubblicazione: (2026)
Awake ECMO for Mid‐Tracheal Obstruction: When a Tracheostomy Isn't Enough
di: Jacob Beiriger, et al.
Pubblicazione: (2026)
di: Jacob Beiriger, et al.
Pubblicazione: (2026)
Recall Isn't Enough: Bounding Commitments in Personalized Language Systems
di: Tang, Rui, et al.
Pubblicazione: (2026)
di: Tang, Rui, et al.
Pubblicazione: (2026)
Explainable AI Isn't Enough! Rethinking Algorithmic Contestability
di: Freiesleben, Timo, et al.
Pubblicazione: (2026)
di: Freiesleben, Timo, et al.
Pubblicazione: (2026)
QueryIPI: Query-agnostic Indirect Prompt Injection on Coding Agents
di: Xie, Yuchong, et al.
Pubblicazione: (2025)
di: Xie, Yuchong, et al.
Pubblicazione: (2025)
Red-Teaming Coding Agents from a Tool-Invocation Perspective: An Empirical Security Assessment
di: Xie, Yuchong, et al.
Pubblicazione: (2025)
di: Xie, Yuchong, et al.
Pubblicazione: (2025)
Comment on “Awake ECMO for Mid‐Tracheal Obstruction: When a Tracheostomy Isn't Enough”
di: Ilaria Onorati, et al.
Pubblicazione: (2026)
di: Ilaria Onorati, et al.
Pubblicazione: (2026)
Reasoning Isn't Enough: Examining Truth-Bias and Sycophancy in LLMs
di: Barkett, Emilio, et al.
Pubblicazione: (2025)
di: Barkett, Emilio, et al.
Pubblicazione: (2025)
Inverse Scaling: When Bigger Isn't Better
di: McKenzie, Ian R., et al.
Pubblicazione: (2023)
di: McKenzie, Ian R., et al.
Pubblicazione: (2023)
Strong Reasoning Isn't Enough: Evaluating Evidence Elicitation in Interactive Diagnosis
di: Long, Zhuohan, et al.
Pubblicazione: (2026)
di: Long, Zhuohan, et al.
Pubblicazione: (2026)
If It Isn't Broken...Break It!
di: Voges, Mickie A.
Pubblicazione: (2000)
di: Voges, Mickie A.
Pubblicazione: (2000)
ZTaint-Havoc: From Havoc Mode to Zero-Execution Fuzzing-Driven Taint Inference
di: Xie, Yuchong, et al.
Pubblicazione: (2025)
di: Xie, Yuchong, et al.
Pubblicazione: (2025)
SAIE Framework: Support Alone Isn't Enough -- Advancing LLM Training with Adversarial Remarks
di: Loem, Mengsay, et al.
Pubblicazione: (2023)
di: Loem, Mengsay, et al.
Pubblicazione: (2023)
Access Isn't Enough: Merely Connecting People and Computers Won't Close the Digital Divide.
di: Blau, Andrew
Pubblicazione: (2002)
di: Blau, Andrew
Pubblicazione: (2002)
The Universe Isn't Expanding, It's Relaxing
di: Blouin, Sam
Pubblicazione: (2025)
di: Blouin, Sam
Pubblicazione: (2025)
The Universe Isn't Expanding — It's Relaxing
di: Blouin, Sam
Pubblicazione: (2025)
di: Blouin, Sam
Pubblicazione: (2025)
AI Isn't Creating Anything
di: Lee Skallerup Bessette
Pubblicazione: (2025)
di: Lee Skallerup Bessette
Pubblicazione: (2025)
Videodiscs: A Revolution That Isn't.
Pubblicazione: (1982)
Pubblicazione: (1982)
AI Isn't Creating Anything
di: Lee Skallerup Bessette
Pubblicazione: (2025)
di: Lee Skallerup Bessette
Pubblicazione: (2025)
Model-Behavior Alignment under Flexible Evaluation: When the Best-Fitting Model Isn't the Right One
di: Avitan, Itamar, et al.
Pubblicazione: (2025)
di: Avitan, Itamar, et al.
Pubblicazione: (2025)
From Metacognition to Computable Wisdom: Why "Thinking About Thinking" Isn't Enough for Agentic AI
di: Figurelli, Rogério
Pubblicazione: (2026)
di: Figurelli, Rogério
Pubblicazione: (2026)
Coverage Isn't Enough: SBFL-Driven Insights into Manually Created vs. Automatically Generated Tests
di: Shimizu, Sasara, et al.
Pubblicazione: (2025)
di: Shimizu, Sasara, et al.
Pubblicazione: (2025)
When More Isn't Better: The Curvilinear Effects of ESG on Firm Performance
di: Joel Victor Dossa, et al.
Pubblicazione: (2026)
di: Joel Victor Dossa, et al.
Pubblicazione: (2026)
Seeing Isn't Knowing: Do VLMs Know When Not to Answer Spatial Questions (and Why)?
di: Zhang, Yue, et al.
Pubblicazione: (2026)
di: Zhang, Yue, et al.
Pubblicazione: (2026)
ML Interpretability: Simple Isn't Easy
di: Räz, Tim
Pubblicazione: (2022)
di: Räz, Tim
Pubblicazione: (2022)
"Just Say No" Isn't Sex Education.
di: Osborn, Anne
Pubblicazione: (1991)
di: Osborn, Anne
Pubblicazione: (1991)
When Meaning Isn't Literal: Exploring Idiomatic Meaning Across Languages and Modalities
di: Das, Sarmistha, et al.
Pubblicazione: (2026)
di: Das, Sarmistha, et al.
Pubblicazione: (2026)
When Privacy Isn't Synthetic: Hidden Data Leakage in Generative AI Models
di: Mustaqim, S. M., et al.
Pubblicazione: (2025)
di: Mustaqim, S. M., et al.
Pubblicazione: (2025)
When Fairness Isn't Statistical: The Limits of Machine Learning in Evaluating Legal Reasoning
di: Barale, Claire, et al.
Pubblicazione: (2025)
di: Barale, Claire, et al.
Pubblicazione: (2025)
When Slower Isn't Truer: Inverse Scaling Law of Truthfulness in Multimodal Reasoning
di: Fang, Sitong, et al.
Pubblicazione: (2025)
di: Fang, Sitong, et al.
Pubblicazione: (2025)
The Newberys: Getting Them Read (It Isn't Easy)
di: Aborne, Carlene
Pubblicazione: (1974)
di: Aborne, Carlene
Pubblicazione: (1974)
The Future Isn't What It Used to Be: Videotex Is on the Way.
di: McKenzie, Jamieson A.
Pubblicazione: (1984)
di: McKenzie, Jamieson A.
Pubblicazione: (1984)
Excuse Me, Isn't That Your Library on Fire?
di: Grayson, Randall
Pubblicazione: (1998)
di: Grayson, Randall
Pubblicazione: (1998)
Take the Step—Waiting Isn’t a Strategy
di: David B. LaFrance
Pubblicazione: (2026)
di: David B. LaFrance
Pubblicazione: (2026)
When Correct Isn't Usable: Improving Structured Output Reliability in Small Language Models
di: Galeone, Cosimo, et al.
Pubblicazione: (2026)
di: Galeone, Cosimo, et al.
Pubblicazione: (2026)
Documenti analoghi
-
CompressionAttack: Exploiting Prompt Compression as a New Attack Surface in LLM-Powered Agents
di: Liu, Zesen, et al.
Pubblicazione: (2025) -
When Trust Isn't Enough.
di: Behrman, Sara
Pubblicazione: (1998) -
From Similarity to Vulnerability: Key Collision Attack on LLM Semantic Caching
di: Zhang, Zhixiang, et al.
Pubblicazione: (2026) -
Knowing the Answer Isn't Enough: Fixing Reasoning Path Failures in LVLMs
di: Wang, Chaoyang, et al.
Pubblicazione: (2025) -
When One Moment Isn't Enough: Multi-Moment Retrieval with Cross-Moment Interactions
di: Cao, Zhuo, et al.
Pubblicazione: (2025)