If you're waiting for a sign... that might not be it! Mitigating Trust Boundary Confusion from Visual Injections on Vision-Language Agentic Systems
Fuente:
arXiv
Salvato in:
| Autori principali: | Chang, Jiamin, Xue, Minhui, Sun, Ruoxi, Pang, Shuchao, Kanhere, Salil S., Pearce, Hammond |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
What's Pulling the Strings? Evaluating Integrity and Attribution in AI Training and Inference through Concept Shift
di: Chang, Jiamin, et al.
Pubblicazione: (2025)
di: Chang, Jiamin, et al.
Pubblicazione: (2025)
SoK: The Security-Safety Continuum of Multimodal Foundation Models through Information Flow and Global Game-Theoretic Analysis of Asymmetric Threats
di: Sun, Ruoxi, et al.
Pubblicazione: (2024)
di: Sun, Ruoxi, et al.
Pubblicazione: (2024)
SoK: Trusting Self-Sovereign Identity
di: Krul, Evan, et al.
Pubblicazione: (2024)
di: Krul, Evan, et al.
Pubblicazione: (2024)
The Americas. Thank you, general, you're dismissed
Pubblicazione: (1997)
Pubblicazione: (1997)
Label Shift Estimation With Incremental Prior Update
di: Zhang, Yunrui, et al.
Pubblicazione: (2026)
di: Zhang, Yunrui, et al.
Pubblicazione: (2026)
Revisit Time Series Classification Benchmark: The Impact of Temporal Information for Classification
di: Zhang, Yunrui, et al.
Pubblicazione: (2025)
di: Zhang, Yunrui, et al.
Pubblicazione: (2025)
Instance-Wise Monotonic Calibration by Constrained Transformation
di: Zhang, Yunrui, et al.
Pubblicazione: (2025)
di: Zhang, Yunrui, et al.
Pubblicazione: (2025)
GETA: Generalized Encrypted Traffic Analysis
di: Gunasekara, Ransika, et al.
Pubblicazione: (2026)
di: Gunasekara, Ransika, et al.
Pubblicazione: (2026)
Synthetic Trajectory Generation Through Convolutional Neural Networks
di: Merhi, Jesse, et al.
Pubblicazione: (2024)
di: Merhi, Jesse, et al.
Pubblicazione: (2024)
In Vino Veritas and Vulnerabilities: Examining LLM Safety via Drunk Language Inducement
di: Shetty, Anudeex, et al.
Pubblicazione: (2026)
di: Shetty, Anudeex, et al.
Pubblicazione: (2026)
AgentRAE: Remote Action Execution through Notification-based Visual Backdoors against Screenshots-based Mobile GUI Agents
di: Luo, Yutao, et al.
Pubblicazione: (2026)
di: Luo, Yutao, et al.
Pubblicazione: (2026)
Gender identity: You think you're confused?
di: Alison Knopf
Pubblicazione: (2024)
di: Alison Knopf
Pubblicazione: (2024)
Reconstruction of Differentially Private Text Sanitization via Large Language Models
di: Pang, Shuchao, et al.
Pubblicazione: (2024)
di: Pang, Shuchao, et al.
Pubblicazione: (2024)
SoK: Practical Aspects of Releasing Differentially Private Graphs
di: D'Silva, Nicholas, et al.
Pubblicazione: (2026)
di: D'Silva, Nicholas, et al.
Pubblicazione: (2026)
Ensuring you're insured : the role of the ombudsman for short-term insurance
Pubblicazione: (2003)
Pubblicazione: (2003)
Adversarially Guided Stateful Defense Against Backdoor Attacks in Federated Deep Learning
di: Ali, Hassan, et al.
Pubblicazione: (2024)
di: Ali, Hassan, et al.
Pubblicazione: (2024)
Beyond Life: A Digital Will Solution for Posthumous Data Management
di: Chen, Xinzhang, et al.
Pubblicazione: (2025)
di: Chen, Xinzhang, et al.
Pubblicazione: (2025)
Multi-MedChain: Multi-Party Multi-Blockchain Medical Supply Chain Management System
di: Saini, Akanksha, et al.
Pubblicazione: (2024)
di: Saini, Akanksha, et al.
Pubblicazione: (2024)
Proceed to your nearest emergency room … and wait
di: Edward Tabor
Pubblicazione: (2025)
di: Edward Tabor
Pubblicazione: (2025)
A Duty to Forget, a Right to be Assured? Exposing Vulnerabilities in Machine Unlearning Services
di: Hu, Hongsheng, et al.
Pubblicazione: (2023)
di: Hu, Hongsheng, et al.
Pubblicazione: (2023)
MediConfusion: Can you trust your AI radiologist? Probing the reliability of multimodal medical foundation models
di: Sepehri, Mohammad Shahab, et al.
Pubblicazione: (2024)
di: Sepehri, Mohammad Shahab, et al.
Pubblicazione: (2024)
À la recherche du sens perdu: your favourite LLM might have more to say than you can understand
di: Erziev, K. O. T.
Pubblicazione: (2025)
di: Erziev, K. O. T.
Pubblicazione: (2025)
Everybody needs good neighbours, even if you’re a winter flounder
di: William Bernard Perry
Pubblicazione: (2024)
di: William Bernard Perry
Pubblicazione: (2024)
Honeyfile Camouflage: Hiding Fake Files in Plain Sight
di: Timmer, Roelien C., et al.
Pubblicazione: (2024)
di: Timmer, Roelien C., et al.
Pubblicazione: (2024)
Masked Vector Quantization
di: Nguyen, David D., et al.
Pubblicazione: (2023)
di: Nguyen, David D., et al.
Pubblicazione: (2023)
Comparison of Multilingual and Bilingual Models for Satirical News Detection of Arabic and English
di: Abdalla, Omar W., et al.
Pubblicazione: (2024)
di: Abdalla, Omar W., et al.
Pubblicazione: (2024)
Prompt Injection as Role Confusion
di: Ye, Charles, et al.
Pubblicazione: (2026)
di: Ye, Charles, et al.
Pubblicazione: (2026)
Be sure you're recruiting the right person to the board—not just “checking a box”
Pubblicazione: (2025)
Pubblicazione: (2025)
Response to letter to the editor re: Worth waiting for?
di: Karen Joseph, et al.
Pubblicazione: (2024)
di: Karen Joseph, et al.
Pubblicazione: (2024)
‘You're not supposed to be gay, you're black’: Analysing race and LGBTQ+ youth identity through an intersectional lens
di: Lucy Jones
Pubblicazione: (2024)
di: Lucy Jones
Pubblicazione: (2024)
The boosted HP filter is more general than you might think
di: Mei, Ziwei, et al.
Pubblicazione: (2022)
di: Mei, Ziwei, et al.
Pubblicazione: (2022)
SoK: Can Trajectory Generation Combine Privacy and Utility?
di: Buchholz, Erik, et al.
Pubblicazione: (2024)
di: Buchholz, Erik, et al.
Pubblicazione: (2024)
Demo: TOSense -- What Did You Just Agree to?
di: Chen, Xinzhang, et al.
Pubblicazione: (2025)
di: Chen, Xinzhang, et al.
Pubblicazione: (2025)
Towards Weaknesses and Attack Patterns Prediction for IoT Devices
di: A., Carlos A. Rivera, et al.
Pubblicazione: (2024)
di: A., Carlos A. Rivera, et al.
Pubblicazione: (2024)
How to decide what to do: Why you're already a realist about value
di: Claire Kirwin
Pubblicazione: (2024)
di: Claire Kirwin
Pubblicazione: (2024)
“You think you’re ‘one of the boys’ but you never really are”: the impact of discriminatory violence on the retention of women in the construction industry in Quebec
di: Laurence Hamel-Roy (Author), et al.
Pubblicazione: (2021)
di: Laurence Hamel-Roy (Author), et al.
Pubblicazione: (2021)
The 2020 US Decennial Census is more private than you (might) think
di: Su, Buxin, et al.
Pubblicazione: (2024)
di: Su, Buxin, et al.
Pubblicazione: (2024)
"Think about it like you're a firefighter": Understanding How Reddit Moderators Use the Modqueue
di: Bajpai, Tanvi, et al.
Pubblicazione: (2025)
di: Bajpai, Tanvi, et al.
Pubblicazione: (2025)
‛Until you're in the chair and executing your role, you don't know’: A qualitative study of the needs and perspectives of people with stroke‐related communication disabilities when returning to vocational activity
di: Lucette Lanyon, et al.
Pubblicazione: (2024)
di: Lucette Lanyon, et al.
Pubblicazione: (2024)
The Invisible Game on the Internet: A Case Study of Decoding Deceptive Patterns
di: Shi, Zewei, et al.
Pubblicazione: (2024)
di: Shi, Zewei, et al.
Pubblicazione: (2024)
Documenti analoghi
-
What's Pulling the Strings? Evaluating Integrity and Attribution in AI Training and Inference through Concept Shift
di: Chang, Jiamin, et al.
Pubblicazione: (2025) -
SoK: The Security-Safety Continuum of Multimodal Foundation Models through Information Flow and Global Game-Theoretic Analysis of Asymmetric Threats
di: Sun, Ruoxi, et al.
Pubblicazione: (2024) -
SoK: Trusting Self-Sovereign Identity
di: Krul, Evan, et al.
Pubblicazione: (2024) -
The Americas. Thank you, general, you're dismissed
Pubblicazione: (1997) -
Label Shift Estimation With Incremental Prior Update
di: Zhang, Yunrui, et al.
Pubblicazione: (2026)