AutoRISE: Agent-Driven Strategy Evolution for Red-Teaming Large Language Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Gautam, Tanmay, Bahramali, Alireza, Atluri, Sandeep |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Effective Red-Teaming of Policy-Adherent Agents
von: Nakash, Itay, et al.
Veröffentlicht: (2025)
von: Nakash, Itay, et al.
Veröffentlicht: (2025)
Web Fraud Attacks Against LLM-Driven Multi-Agent Systems
von: Kong, Dezhang, et al.
Veröffentlicht: (2025)
von: Kong, Dezhang, et al.
Veröffentlicht: (2025)
Public and private blockchain for decentralized digital building twins and building automation system
von: Ly, Reachsak, et al.
Veröffentlicht: (2026)
von: Ly, Reachsak, et al.
Veröffentlicht: (2026)
CTFExplorer: Evaluating LLM Offensive Agents Through Multi-Target Web CTF Benchmarking
von: Rani, Nanda, et al.
Veröffentlicht: (2026)
von: Rani, Nanda, et al.
Veröffentlicht: (2026)
SV-LLM: An Agentic Approach for SoC Security Verification using Large Language Models
von: Saha, Dipayan, et al.
Veröffentlicht: (2025)
von: Saha, Dipayan, et al.
Veröffentlicht: (2025)
Who Owns This Agent? Tracing AI Agents Back to Their Owners
von: Chocron, Ruben, et al.
Veröffentlicht: (2026)
von: Chocron, Ruben, et al.
Veröffentlicht: (2026)
Agents for Agents: An Interrogator-Based Secure Framework for Autonomous Internet of Underwater Things
von: Akarma, Ali, et al.
Veröffentlicht: (2026)
von: Akarma, Ali, et al.
Veröffentlicht: (2026)
Open Challenges in Multi-Agent Security: Towards Secure Systems of Interacting AI Agents
von: de Witt, Christian Schroeder, et al.
Veröffentlicht: (2025)
von: de Witt, Christian Schroeder, et al.
Veröffentlicht: (2025)
Trusted AI Agents in the Cloud
von: Bodea, Teofil, et al.
Veröffentlicht: (2025)
von: Bodea, Teofil, et al.
Veröffentlicht: (2025)
SentinelAgent: Intent-Verified Delegation Chains for Securing Federal Multi-Agent AI Systems
von: Patil, KrishnaSaiReddy
Veröffentlicht: (2026)
von: Patil, KrishnaSaiReddy
Veröffentlicht: (2026)
The Art of Building Verifiers for Computer Use Agents
von: Rosset, Corby, et al.
Veröffentlicht: (2026)
von: Rosset, Corby, et al.
Veröffentlicht: (2026)
Agent Capability Negotiation and Binding Protocol (ACNBP)
von: Huang, Ken, et al.
Veröffentlicht: (2025)
von: Huang, Ken, et al.
Veröffentlicht: (2025)
Agent Name Service (ANS): A Proof-of-Concept Trust Layer for Secure AI Agent Discovery, Identity, and Governance in Kubernetes
von: Mittal, Akshay, et al.
Veröffentlicht: (2026)
von: Mittal, Akshay, et al.
Veröffentlicht: (2026)
Multi-Agent Actor-Critics in Autonomous Cyber Defense
von: Wang, Mingjun, et al.
Veröffentlicht: (2024)
von: Wang, Mingjun, et al.
Veröffentlicht: (2024)
Chronology of Multi-Agent Interactions for Provenance of Evolving Information
von: Chang, Ching-Chun, et al.
Veröffentlicht: (2025)
von: Chang, Ching-Chun, et al.
Veröffentlicht: (2025)
Prompt Flow Integrity to Prevent Privilege Escalation in LLM Agents
von: Kim, Juhee, et al.
Veröffentlicht: (2025)
von: Kim, Juhee, et al.
Veröffentlicht: (2025)
Towards Log Analysis with AI Agents: Cowrie Case Study
von: Karaarslan, Enis, et al.
Veröffentlicht: (2025)
von: Karaarslan, Enis, et al.
Veröffentlicht: (2025)
A Vision for Access Control in LLM-based Agent Systems
von: Li, Xinfeng, et al.
Veröffentlicht: (2025)
von: Li, Xinfeng, et al.
Veröffentlicht: (2025)
Heartbeat-Bound Hierarchical Credentials: Cryptographic Revocation for AI Agent Swarms
von: Deochake, Saurabh
Veröffentlicht: (2026)
von: Deochake, Saurabh
Veröffentlicht: (2026)
The Aegis Protocol: A Foundational Security Framework for Autonomous AI Agents
von: Adapala, Sai Teja Reddy, et al.
Veröffentlicht: (2025)
von: Adapala, Sai Teja Reddy, et al.
Veröffentlicht: (2025)
LegalSim: Multi-Agent Simulation of Legal Systems for Discovering Procedural Exploits
von: Badhe, Sanket
Veröffentlicht: (2025)
von: Badhe, Sanket
Veröffentlicht: (2025)
Prompt Infection: LLM-to-LLM Prompt Injection within Multi-Agent Systems
von: Lee, Donghyun, et al.
Veröffentlicht: (2024)
von: Lee, Donghyun, et al.
Veröffentlicht: (2024)
Digital Identity for Agentic Systems: Toward a Portable Authorization Standard for Autonomous Agents
von: Madhira, Partha
Veröffentlicht: (2026)
von: Madhira, Partha
Veröffentlicht: (2026)
OrchJail: Jailbreaking Tool-Calling Text-to-Image Agents by Orchestration-Guided Fuzzing
von: Chen, Jianming, et al.
Veröffentlicht: (2026)
von: Chen, Jianming, et al.
Veröffentlicht: (2026)
From Cloud-Native to Trust-Native: A Protocol for Verifiable Multi-Agent Systems
von: Li, Muyang
Veröffentlicht: (2025)
von: Li, Muyang
Veröffentlicht: (2025)
Attack the Messages, Not the Agents: A Multi-round Adaptive Stealthy Tampering Framework for LLM-MAS
von: Yan, Bingyu, et al.
Veröffentlicht: (2025)
von: Yan, Bingyu, et al.
Veröffentlicht: (2025)
GAMMAF: A Common Framework for Graph-Based Anomaly Monitoring Benchmarking in LLM Multi-Agent Systems
von: Mateo-Torrejón, Pablo, et al.
Veröffentlicht: (2026)
von: Mateo-Torrejón, Pablo, et al.
Veröffentlicht: (2026)
Ambient Persuasion in a Deployed AI Agent: Unauthorized Escalation Following Routine Non-Adversarial Content Exposure
von: Cuadros, Diego F., et al.
Veröffentlicht: (2026)
von: Cuadros, Diego F., et al.
Veröffentlicht: (2026)
Explainable and Fine-Grained Safeguarding of LLM Multi-Agent Systems via Bi-Level Graph Anomaly Detection
von: Pan, Junjun, et al.
Veröffentlicht: (2025)
von: Pan, Junjun, et al.
Veröffentlicht: (2025)
Secure Forgetting: A Framework for Privacy-Driven Unlearning in Large Language Model (LLM)-Based Agents
von: Ye, Dayong, et al.
Veröffentlicht: (2026)
von: Ye, Dayong, et al.
Veröffentlicht: (2026)
CRAKEN: Cybersecurity LLM Agent with Knowledge-Based Execution
von: Shao, Minghao, et al.
Veröffentlicht: (2025)
von: Shao, Minghao, et al.
Veröffentlicht: (2025)
Emerging Safety Attack and Defense in Federated Instruction Tuning of Large Language Models
von: Ye, Rui, et al.
Veröffentlicht: (2024)
von: Ye, Rui, et al.
Veröffentlicht: (2024)
X-Teaming: Multi-Turn Jailbreaks and Defenses with Adaptive Multi-Agents
von: Rahman, Salman, et al.
Veröffentlicht: (2025)
von: Rahman, Salman, et al.
Veröffentlicht: (2025)
LLM-Agent-UMF: LLM-based Agent Unified Modeling Framework for Seamless Design of Multi Active/Passive Core-Agent Architectures
von: Hassouna, Amine Ben, et al.
Veröffentlicht: (2024)
von: Hassouna, Amine Ben, et al.
Veröffentlicht: (2024)
DASH: Deception-Augmented Shared Mental Model for a Human-Machine Teaming System
von: Wan, Zelin, et al.
Veröffentlicht: (2025)
von: Wan, Zelin, et al.
Veröffentlicht: (2025)
Architectural Obsolescence of Unhardened Agentic-AI Runtimes
von: Metere, Alfredo
Veröffentlicht: (2026)
von: Metere, Alfredo
Veröffentlicht: (2026)
enclawed: A Configurable, Sector-Neutral Hardening Framework for Single-User AI Assistant Gateways
von: Metere, Alfredo
Veröffentlicht: (2026)
von: Metere, Alfredo
Veröffentlicht: (2026)
Governance-Constrained Agentic AI: Blockchain-Enforced Human Oversight for Safety-Critical Wildfire Monitoring
von: Akarma, Ali, et al.
Veröffentlicht: (2026)
von: Akarma, Ali, et al.
Veröffentlicht: (2026)
TessPay: Verify-then-Pay Infrastructure for Trusted Agentic Commerce
von: Goenka, Mehul, et al.
Veröffentlicht: (2026)
von: Goenka, Mehul, et al.
Veröffentlicht: (2026)
Protecting Context and Prompts: Deterministic Security for Non-Deterministic AI
von: Rajagopalan, Mohan, et al.
Veröffentlicht: (2026)
von: Rajagopalan, Mohan, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Effective Red-Teaming of Policy-Adherent Agents
von: Nakash, Itay, et al.
Veröffentlicht: (2025) -
Web Fraud Attacks Against LLM-Driven Multi-Agent Systems
von: Kong, Dezhang, et al.
Veröffentlicht: (2025) -
Public and private blockchain for decentralized digital building twins and building automation system
von: Ly, Reachsak, et al.
Veröffentlicht: (2026) -
CTFExplorer: Evaluating LLM Offensive Agents Through Multi-Target Web CTF Benchmarking
von: Rani, Nanda, et al.
Veröffentlicht: (2026) -
SV-LLM: An Agentic Approach for SoC Security Verification using Large Language Models
von: Saha, Dipayan, et al.
Veröffentlicht: (2025)