Agent Security is a Systems Problem
Fuente:
arXiv
Saved in:
| Main Authors: | Christodorescu, Mihai, Fernandes, Earlence, Hooda, Ashish, Jha, Somesh, Rehberger, Johann, Chaudhuri, Kamalika, Fu, Xiaohan, Shams, Khawaja, Amir, Guy, Choi, Jihye, Choudhary, Sarthak, Palumbo, Nils, Labunets, Andrey, Pandya, Nishit V. |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Systems Security Foundations for Agentic Computing
by: Christodorescu, Mihai, et al.
Published: (2025)
by: Christodorescu, Mihai, et al.
Published: (2025)
Fun-tuning: Characterizing the Vulnerability of Proprietary LLMs to Optimization-based Prompt Injection Attacks via the Fine-Tuning Interface
by: Labunets, Andrey, et al.
Published: (2025)
by: Labunets, Andrey, et al.
Published: (2025)
Through the Stealth Lens: Attention-Aware Defenses Against Poisoning in RAG
by: Choudhary, Sarthak, et al.
Published: (2025)
by: Choudhary, Sarthak, et al.
Published: (2025)
May I have your Attention? Breaking Fine-Tuning based Prompt Injection Defenses using Architecture-Aware Attacks
by: Pandya, Nishit V., et al.
Published: (2025)
by: Pandya, Nishit V., et al.
Published: (2025)
Formal Policy Enforcement for Real-World Agentic Systems
by: Palumbo, Nils, et al.
Published: (2026)
by: Palumbo, Nils, et al.
Published: (2026)
Undetectable Backdoors in Model Parameters: Hiding Sparse Secrets in High Dimensions
by: Choudhary, Sarthak, et al.
Published: (2026)
by: Choudhary, Sarthak, et al.
Published: (2026)
Do Large Code Models Understand Programming Concepts? Counterfactual Analysis for Code Predicates
by: Hooda, Ashish, et al.
Published: (2024)
by: Hooda, Ashish, et al.
Published: (2024)
Auto-SPT: Automating Semantic Preserving Transformations for Code
by: Hooda, Ashish, et al.
Published: (2025)
by: Hooda, Ashish, et al.
Published: (2025)
How Not to Detect Prompt Injections with an LLM
by: Choudhary, Sarthak, et al.
Published: (2025)
by: Choudhary, Sarthak, et al.
Published: (2025)
Dependency-Aware Privacy for Multi-turn Agents
by: Anshumaan, Divyam, et al.
Published: (2026)
by: Anshumaan, Divyam, et al.
Published: (2026)
Verifier-Guided Code Translation via Meta-Step Decoding
by: Zhou, Tianyang, et al.
Published: (2026)
by: Zhou, Tianyang, et al.
Published: (2026)
PRP: Propagating Universal Perturbations to Attack Large Language Model Guard-Rails
by: Mangaokar, Neal, et al.
Published: (2024)
by: Mangaokar, Neal, et al.
Published: (2024)
PolicyLR: A Logic Representation For Privacy Policies
by: Hooda, Ashish, et al.
Published: (2024)
by: Hooda, Ashish, et al.
Published: (2024)
Functional Homotopy: Smoothing Discrete Optimization via Continuous Parameters for LLM Jailbreak Attacks
by: Wang, Zi, et al.
Published: (2024)
by: Wang, Zi, et al.
Published: (2024)
Trust No AI: Prompt Injection Along The CIA Security Triad
by: Rehberger, Johann
Published: (2024)
by: Rehberger, Johann
Published: (2024)
SACTOR: LLM-Driven Correct and Idiomatic C to Rust Translation with Static Analysis and FFI-Based Verification
by: Zhou, Tianyang, et al.
Published: (2025)
by: Zhou, Tianyang, et al.
Published: (2025)
MALADE: Orchestration of LLM-powered Agents with Retrieval Augmented Generation for Pharmacovigilance
by: Choi, Jihye, et al.
Published: (2024)
by: Choi, Jihye, et al.
Published: (2024)
Securing the Future of GenAI: Policy and Technology
by: Christodorescu, Mihai, et al.
Published: (2024)
by: Christodorescu, Mihai, et al.
Published: (2024)
Rethinking Diversity in Deep Neural Network Testing
by: Wang, Zi, et al.
Published: (2023)
by: Wang, Zi, et al.
Published: (2023)
Adaptive Concept Bottleneck for Foundation Models Under Distribution Shifts
by: Choi, Jihye, et al.
Published: (2024)
by: Choi, Jihye, et al.
Published: (2024)
Two Heads are Actually Better than One: Towards Better Adversarial Robustness via Transduction and Rejection
by: Palumbo, Nils, et al.
Published: (2023)
by: Palumbo, Nils, et al.
Published: (2023)
Validating Mechanistic Interpretations: An Axiomatic Approach
by: Palumbo, Nils, et al.
Published: (2024)
by: Palumbo, Nils, et al.
Published: (2024)
Taint-Style Vulnerability Detection and Confirmation for Node.js Packages Using LLM Agent Reasoning
by: Ni, Ronghao, et al.
Published: (2026)
by: Ni, Ronghao, et al.
Published: (2026)
PolicyBank: Evolving Policy Understanding for LLM Agents
by: Choi, Jihye, et al.
Published: (2026)
by: Choi, Jihye, et al.
Published: (2026)
Defending Object Detectors against Patch Attacks with Out-of-Distribution Smoothing
by: Feng, Ryan, et al.
Published: (2022)
by: Feng, Ryan, et al.
Published: (2022)
ATLAS: Constraints-Aware Multi-Agent Collaboration for Real-World Travel Planning
by: Choi, Jihye, et al.
Published: (2025)
by: Choi, Jihye, et al.
Published: (2025)
Data Redaction from Conditional Generative Models
by: Kong, Zhifeng, et al.
Published: (2023)
by: Kong, Zhifeng, et al.
Published: (2023)
A Closer Look at the Learnability of Out-of-Distribution (OOD) Detection
by: Garov, Konstantin, et al.
Published: (2025)
by: Garov, Konstantin, et al.
Published: (2025)
Distribution Learning with Valid Outputs Beyond the Worst-Case
by: Rittler, Nick, et al.
Published: (2024)
by: Rittler, Nick, et al.
Published: (2024)
Quality control of PEN wavelength shifters for DarkSide-20k veto
by: Choudhary, Sarthak
Published: (2025)
by: Choudhary, Sarthak
Published: (2025)
What Really is a Member? Discrediting Membership Inference via Poisoning
by: Mangaokar, Neal, et al.
Published: (2025)
by: Mangaokar, Neal, et al.
Published: (2025)
SLVR: Securely Leveraging Client Validation for Robust Federated Learning
by: Choi, Jihye, et al.
Published: (2025)
by: Choi, Jihye, et al.
Published: (2025)
Unified Uncertainty Calibration
by: Chaudhuri, Kamalika, et al.
Published: (2023)
by: Chaudhuri, Kamalika, et al.
Published: (2023)
Identifying and Mitigating the Security Risks of Generative AI
by: Barrett, Clark, et al.
Published: (2023)
by: Barrett, Clark, et al.
Published: (2023)
Learning-Time Encoding Shapes Unlearning in LLMs
by: Wu, Ruihan, et al.
Published: (2025)
by: Wu, Ruihan, et al.
Published: (2025)
Better Membership Inference Privacy Measurement through Discrepancy
by: Wu, Ruihan, et al.
Published: (2024)
by: Wu, Ruihan, et al.
Published: (2024)
Beyond Discrepancy: A Closer Look at the Theory of Distribution Shift
by: Bhattacharjee, Robi, et al.
Published: (2024)
by: Bhattacharjee, Robi, et al.
Published: (2024)
Measuring Privacy Loss in Distributed Spatio-Temporal Data
by: Koga, Tatsuki, et al.
Published: (2024)
by: Koga, Tatsuki, et al.
Published: (2024)
Communication-Efficient Triangle Counting under Local Differential Privacy
by: Imola, Jacob, et al.
Published: (2021)
by: Imola, Jacob, et al.
Published: (2021)
Auditing $f$-Differential Privacy in One Run
by: Mahloujifar, Saeed, et al.
Published: (2024)
by: Mahloujifar, Saeed, et al.
Published: (2024)
Similar Items
-
Systems Security Foundations for Agentic Computing
by: Christodorescu, Mihai, et al.
Published: (2025) -
Fun-tuning: Characterizing the Vulnerability of Proprietary LLMs to Optimization-based Prompt Injection Attacks via the Fine-Tuning Interface
by: Labunets, Andrey, et al.
Published: (2025) -
Through the Stealth Lens: Attention-Aware Defenses Against Poisoning in RAG
by: Choudhary, Sarthak, et al.
Published: (2025) -
May I have your Attention? Breaking Fine-Tuning based Prompt Injection Defenses using Architecture-Aware Attacks
by: Pandya, Nishit V., et al.
Published: (2025) -
Formal Policy Enforcement for Real-World Agentic Systems
by: Palumbo, Nils, et al.
Published: (2026)