AgentLeak: A Full-Stack Benchmark for Privacy Leakage in Multi-Agent LLM Systems
Fuente:
arXiv
Guardado en:
| Autores principales: | Yagoubi, Faouzi El, Badu-Marfo, Godwin, Mallah, Ranwa Al |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Solving Zebra Puzzles Using Constraint-Guided Multi-Agent Systems
por: Berman, Shmuel, et al.
Publicado: (2024)
por: Berman, Shmuel, et al.
Publicado: (2024)
Controlling Long-Horizon Behavior in Language Model Agents with Explicit State Dynamics
por: Subaharan, Sukesh
Publicado: (2026)
por: Subaharan, Sukesh
Publicado: (2026)
Project Synapse: A Hierarchical Multi-Agent Framework with Hybrid Memory for Autonomous Resolution of Last-Mile Delivery Disruptions
por: Yadav, Arin Gopalan, et al.
Publicado: (2026)
por: Yadav, Arin Gopalan, et al.
Publicado: (2026)
NeuroState-Bench: A Human-Calibrated Benchmark for Commitment Integrity in LLM Agent Profiles
por: Jia, Xiao
Publicado: (2026)
por: Jia, Xiao
Publicado: (2026)
TRIZ Agents: A Multi-Agent LLM Approach for TRIZ-Based Innovation
por: Szczepanik, Kamil, et al.
Publicado: (2025)
por: Szczepanik, Kamil, et al.
Publicado: (2025)
From Helpfulness to Toxic Proactivity: Diagnosing Behavioral Misalignment in LLM Agents
por: Wang, Xinyue, et al.
Publicado: (2026)
por: Wang, Xinyue, et al.
Publicado: (2026)
Collaborative LLM Agents for C4 Software Architecture Design Automation
por: Szczepanik, Kamil, et al.
Publicado: (2025)
por: Szczepanik, Kamil, et al.
Publicado: (2025)
ART: Adaptive Response Tuning Framework -- A Multi-Agent Tournament-Based Approach to LLM Response Optimization
por: Khan, Omer Jauhar
Publicado: (2025)
por: Khan, Omer Jauhar
Publicado: (2025)
HumanMCP: A Human-Like Query Dataset for Evaluating MCP Tool Retrieval Performance
por: Laddha, Shubh, et al.
Publicado: (2025)
por: Laddha, Shubh, et al.
Publicado: (2025)
Agent Capsules: Quality-Gated Granularity Control for Multi-Agent LLM Pipelines
por: Ray, Aninda
Publicado: (2026)
por: Ray, Aninda
Publicado: (2026)
Multi-Agent Synergy-Driven Iterative Visual Narrative Synthesis
por: Xi, Wang, et al.
Publicado: (2025)
por: Xi, Wang, et al.
Publicado: (2025)
Applying Cognitive Design Patterns to General LLM Agents
por: Wray, Robert E., et al.
Publicado: (2025)
por: Wray, Robert E., et al.
Publicado: (2025)
AgentMisalignment: Measuring the Propensity for Misaligned Behaviour in LLM-Based Agents
por: Naik, Akshat, et al.
Publicado: (2025)
por: Naik, Akshat, et al.
Publicado: (2025)
ABot-Claw: A Foundation for Persistent, Cooperative, and Self-Evolving Robotic Agents
por: Huo, Dongjie, et al.
Publicado: (2026)
por: Huo, Dongjie, et al.
Publicado: (2026)
PARNESS: A Paper Harness for End-to-End Automated Scientific Research with Dynamic Workflows, Full-Text Indexing, and Cross-Run Knowledge Accumulation
por: Wang, Yuchen, et al.
Publicado: (2026)
por: Wang, Yuchen, et al.
Publicado: (2026)
Reservoir Computing with Evolved Critical Neural Cellular Automata
por: Pontes-Filho, Sidney, et al.
Publicado: (2025)
por: Pontes-Filho, Sidney, et al.
Publicado: (2025)
MEMTIER: Tiered Memory Architecture and Retrieval Bottleneck Analysis for Long-Running Autonomous AI Agents
por: Sidik, Bronislav, et al.
Publicado: (2026)
por: Sidik, Bronislav, et al.
Publicado: (2026)
Exploring Design of Multi-Agent LLM Dialogues for Research Ideation
por: Ueda, Keisuke, et al.
Publicado: (2025)
por: Ueda, Keisuke, et al.
Publicado: (2025)
ScrapMem: A Bio-inspired Framework for On-device Personalized Agent Memory via Optical Forgetting
por: Chang, Jiale, et al.
Publicado: (2026)
por: Chang, Jiale, et al.
Publicado: (2026)
Informed AI Regulation: Comparing the Ethical Frameworks of Leading LLM Chatbots Using an Ethics-Based Audit to Assess Moral Reasoning and Normative Values
por: Chun, Jon, et al.
Publicado: (2024)
por: Chun, Jon, et al.
Publicado: (2024)
Multi-Agent Object Detection Framework Based on Raspberry Pi YOLO Detector and Slack-Ollama Natural Language Interface
por: Kalušev, Vladimir, et al.
Publicado: (2026)
por: Kalušev, Vladimir, et al.
Publicado: (2026)
Do We Always Need Query-Level Workflows? Rethinking Agentic Workflow Generation for Multi-Agent Systems
por: Wang, Zixu, et al.
Publicado: (2026)
por: Wang, Zixu, et al.
Publicado: (2026)
PRIMA: Operational Patterns for Resilient Multi-Agent Research with Verifiable Identity and Convergent Feedback
por: Annapureddy, Sasank
Publicado: (2026)
por: Annapureddy, Sasank
Publicado: (2026)
Instruction-Level Weight Shaping: A Framework for Self-Improving AI Agents
por: Costa, Rimom
Publicado: (2025)
por: Costa, Rimom
Publicado: (2025)
Multi-Agent GraphRAG: A Text-to-Cypher Framework for Labeled Property Graphs
por: Gusarov, Anton, et al.
Publicado: (2025)
por: Gusarov, Anton, et al.
Publicado: (2025)
DeepPersona: A Generative Engine for Scaling Deep Synthetic Personas
por: Wang, Zhen, et al.
Publicado: (2025)
por: Wang, Zhen, et al.
Publicado: (2025)
Manipulating Transformer-Based Models: Controllability, Steerability, and Robust Interventions
por: Alpay, Faruk, et al.
Publicado: (2025)
por: Alpay, Faruk, et al.
Publicado: (2025)
Performance Evaluation of Sentiment Analysis on Text and Emoji Data Using End-to-End, Transfer Learning, Distributed and Explainable AI Models
por: Velampalli, Sirisha, et al.
Publicado: (2025)
por: Velampalli, Sirisha, et al.
Publicado: (2025)
On measuring grounding and generalizing grounding problems
por: Quigley, Daniel, et al.
Publicado: (2025)
por: Quigley, Daniel, et al.
Publicado: (2025)
Council Mode: A Heterogeneous Multi-Agent Consensus Framework for Reducing LLM Hallucination and Bias
por: Wu, Shuai, et al.
Publicado: (2026)
por: Wu, Shuai, et al.
Publicado: (2026)
Beyond Rating: A Comprehensive Evaluation and Benchmark for AI Reviews
por: Li, Bowen, et al.
Publicado: (2026)
por: Li, Bowen, et al.
Publicado: (2026)
Tool-RoCo: An Agent-as-Tool Self-organization Large Language Model Benchmark in Multi-robot Cooperation
por: Zhang, Ke, et al.
Publicado: (2025)
por: Zhang, Ke, et al.
Publicado: (2025)
Benchmarking Deception Probes via Black-to-White Performance Boosts
por: Parrack, Avi, et al.
Publicado: (2025)
por: Parrack, Avi, et al.
Publicado: (2025)
Multi-Paradigm Agent Interaction in Practice:A Systematic Analysis of Generator-Evaluator, ReAct Loop,and Adversarial Evaluation in the buddyMe Framework
por: Wang, Xiaohua, et al.
Publicado: (2026)
por: Wang, Xiaohua, et al.
Publicado: (2026)
Mimosa Framework: Toward Evolving Multi-Agent Systems for Scientific Research
por: Legrand, Martin, et al.
Publicado: (2026)
por: Legrand, Martin, et al.
Publicado: (2026)
Prima.cpp: Fast 30-70B LLM Inference on Heterogeneous and Low-Resource Home Clusters
por: Li, Zonghang, et al.
Publicado: (2025)
por: Li, Zonghang, et al.
Publicado: (2025)
Terrarium: Revisiting the Blackboard for Multi-Agent Safety, Privacy, and Security Studies
por: Nakamura, Mason, et al.
Publicado: (2025)
por: Nakamura, Mason, et al.
Publicado: (2025)
AI Agents: Evolution, Architecture, and Real-World Applications
por: Krishnan, Naveen
Publicado: (2025)
por: Krishnan, Naveen
Publicado: (2025)
Adaptive Minds: Empowering Agents with LoRA-as-Tools
por: Shekar, Pavan C, et al.
Publicado: (2025)
por: Shekar, Pavan C, et al.
Publicado: (2025)
Understanding Multi-Agent LLM Frameworks: A Unified Benchmark and Experimental Analysis
por: Orogat, Abdelghny, et al.
Publicado: (2026)
por: Orogat, Abdelghny, et al.
Publicado: (2026)
Ejemplares similares
-
Solving Zebra Puzzles Using Constraint-Guided Multi-Agent Systems
por: Berman, Shmuel, et al.
Publicado: (2024) -
Controlling Long-Horizon Behavior in Language Model Agents with Explicit State Dynamics
por: Subaharan, Sukesh
Publicado: (2026) -
Project Synapse: A Hierarchical Multi-Agent Framework with Hybrid Memory for Autonomous Resolution of Last-Mile Delivery Disruptions
por: Yadav, Arin Gopalan, et al.
Publicado: (2026) -
NeuroState-Bench: A Human-Calibrated Benchmark for Commitment Integrity in LLM Agent Profiles
por: Jia, Xiao
Publicado: (2026) -
TRIZ Agents: A Multi-Agent LLM Approach for TRIZ-Based Innovation
por: Szczepanik, Kamil, et al.
Publicado: (2025)