AgentBreeder: Mitigating the AI Safety Risks of Multi-Agent Scaffolds via Self-Improvement
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Rosser, J, Foerster, Jakob |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Portable Agent Memory: A Protocol for Cryptographically-Verified Memory Transfer Across Heterogeneous AI Agents
von: Ravindran, Santhosh Kumar
Veröffentlicht: (2026)
von: Ravindran, Santhosh Kumar
Veröffentlicht: (2026)
When the Agent Is the Adversary: Architectural Requirements for Agentic AI Containment After the April 2026 Frontier Model Escape
von: Mitchell, Richard Joseph
Veröffentlicht: (2026)
von: Mitchell, Richard Joseph
Veröffentlicht: (2026)
Reservoir Computing with Evolved Critical Neural Cellular Automata
von: Pontes-Filho, Sidney, et al.
Veröffentlicht: (2025)
von: Pontes-Filho, Sidney, et al.
Veröffentlicht: (2025)
Gradient Atoms: Unsupervised Discovery, Attribution and Steering of Model Behaviors via Sparse Decomposition of Training Gradients
von: Rosser, J
Veröffentlicht: (2026)
von: Rosser, J
Veröffentlicht: (2026)
Evolutionary Algorithms Approach For Search Based On Semantic Document Similarity
von: Muniyappa, Chandrashekar, et al.
Veröffentlicht: (2025)
von: Muniyappa, Chandrashekar, et al.
Veröffentlicht: (2025)
NeuroState-Bench: A Human-Calibrated Benchmark for Commitment Integrity in LLM Agent Profiles
von: Jia, Xiao
Veröffentlicht: (2026)
von: Jia, Xiao
Veröffentlicht: (2026)
An Organization-Scoped LLM Agent Runtime Architecture for Regulated Cybersecurity Operations
von: Fatouros, George, et al.
Veröffentlicht: (2026)
von: Fatouros, George, et al.
Veröffentlicht: (2026)
Cooperative Task Execution in Multi-Agent Systems
von: Karishma, et al.
Veröffentlicht: (2024)
von: Karishma, et al.
Veröffentlicht: (2024)
Infusion: Shaping Model Behavior by Editing Training Data via Influence Functions
von: Rosser, J, et al.
Veröffentlicht: (2026)
von: Rosser, J, et al.
Veröffentlicht: (2026)
Vertical Federated Graph Neural Network for Recommender System
von: Mai, Peihua, et al.
Veröffentlicht: (2023)
von: Mai, Peihua, et al.
Veröffentlicht: (2023)
ABot-Claw: A Foundation for Persistent, Cooperative, and Self-Evolving Robotic Agents
von: Huo, Dongjie, et al.
Veröffentlicht: (2026)
von: Huo, Dongjie, et al.
Veröffentlicht: (2026)
Quantigence: A Multi-Agent AI Framework for Quantum Security Research
von: Alquwayfili, Abdulmalik
Veröffentlicht: (2025)
von: Alquwayfili, Abdulmalik
Veröffentlicht: (2025)
Emergent Collective Memory in Decentralized Multi-Agent AI Systems
von: Khushiyant
Veröffentlicht: (2025)
von: Khushiyant
Veröffentlicht: (2025)
BlockA2A: Towards Secure and Verifiable Agent-to-Agent Interoperability
von: Zou, Zhenhua, et al.
Veröffentlicht: (2025)
von: Zou, Zhenhua, et al.
Veröffentlicht: (2025)
A Multi-Agent Framework for Medical AI: Leveraging Fine-Tuned GPT, LLaMA, and DeepSeek R1 for Evidence-Based and Bias-Aware Clinical Query Processing
von: Nourmohammadi, Naeimeh, et al.
Veröffentlicht: (2026)
von: Nourmohammadi, Naeimeh, et al.
Veröffentlicht: (2026)
A V2X-based Privacy Preserving Federated Measuring and Learning System
von: Alekszejenkó, Levente, et al.
Veröffentlicht: (2024)
von: Alekszejenkó, Levente, et al.
Veröffentlicht: (2024)
Multi-Agent Object Detection Framework Based on Raspberry Pi YOLO Detector and Slack-Ollama Natural Language Interface
von: Kalušev, Vladimir, et al.
Veröffentlicht: (2026)
von: Kalušev, Vladimir, et al.
Veröffentlicht: (2026)
The Provenance Paradox in Multi-Agent LLM Routing: Delegation Contracts and Attested Identity in LDP
von: Prakash, Sunil
Veröffentlicht: (2026)
von: Prakash, Sunil
Veröffentlicht: (2026)
FundaPod: A Multi-Persona Agent Pod Platform with Knowledge Graph Memory for AI-Assisted Fundamental Investment Research
von: Zhu, Di, et al.
Veröffentlicht: (2026)
von: Zhu, Di, et al.
Veröffentlicht: (2026)
LSTM-MAS: A Long Short-Term Memory Inspired Multi-Agent System for Long-Context Understanding
von: Jiang, Yichen, et al.
Veröffentlicht: (2026)
von: Jiang, Yichen, et al.
Veröffentlicht: (2026)
Project Synapse: A Hierarchical Multi-Agent Framework with Hybrid Memory for Autonomous Resolution of Last-Mile Delivery Disruptions
von: Yadav, Arin Gopalan, et al.
Veröffentlicht: (2026)
von: Yadav, Arin Gopalan, et al.
Veröffentlicht: (2026)
DeTrigger: A Gradient-Centric Approach to Backdoor Attack Mitigation in Federated Learning
von: Lee, Kichang, et al.
Veröffentlicht: (2024)
von: Lee, Kichang, et al.
Veröffentlicht: (2024)
Adaptive Minds: Empowering Agents with LoRA-as-Tools
von: Shekar, Pavan C, et al.
Veröffentlicht: (2025)
von: Shekar, Pavan C, et al.
Veröffentlicht: (2025)
Systematic Capability Benchmarking of Frontier Large Language Models for Offensive Cyber Tasks
von: Merves, Tyler H., et al.
Veröffentlicht: (2026)
von: Merves, Tyler H., et al.
Veröffentlicht: (2026)
Multi-Agent Synergy-Driven Iterative Visual Narrative Synthesis
von: Xi, Wang, et al.
Veröffentlicht: (2025)
von: Xi, Wang, et al.
Veröffentlicht: (2025)
Blind Gods and Broken Screens: Architecting a Secure, Intent-Centric Mobile Agent Operating System
von: Zou, Zhenhua, et al.
Veröffentlicht: (2026)
von: Zou, Zhenhua, et al.
Veröffentlicht: (2026)
Agent Skills for Large Language Models: Architecture, Acquisition, Security, and the Path Forward
von: Xu, Renjun, et al.
Veröffentlicht: (2026)
von: Xu, Renjun, et al.
Veröffentlicht: (2026)
AURA: Agent for Understanding, Reasoning, and Automated Tool Use in Voice-Driven Tasks
von: Maben, Leander Melroy, et al.
Veröffentlicht: (2025)
von: Maben, Leander Melroy, et al.
Veröffentlicht: (2025)
Agentic Discovery of Neural Architectures: AIRA-Compose and AIRA-Design
von: Pepe, Alberto, et al.
Veröffentlicht: (2026)
von: Pepe, Alberto, et al.
Veröffentlicht: (2026)
AdaptOrch: Task-Adaptive Multi-Agent Orchestration in the Era of LLM Performance Convergence
von: Yu, Geunbin
Veröffentlicht: (2026)
von: Yu, Geunbin
Veröffentlicht: (2026)
Emergent Coordination in Multi-Agent Systems via Pressure Fields and Temporal Decay
von: Rodriguez, Roland
Veröffentlicht: (2026)
von: Rodriguez, Roland
Veröffentlicht: (2026)
LookAhead: Preventing DeFi Attacks via Unveiling Adversarial Contracts
von: Ren, Shoupeng, et al.
Veröffentlicht: (2024)
von: Ren, Shoupeng, et al.
Veröffentlicht: (2024)
Hardening x402: PII-Safe Agentic Payments via Pre-Execution Metadata Filtering
von: Stantchev, Vladimir
Veröffentlicht: (2026)
von: Stantchev, Vladimir
Veröffentlicht: (2026)
SafetyDrift: Predicting When AI Agents Cross the Line Before They Actually Do
von: Dhodapkar, Aditya, et al.
Veröffentlicht: (2026)
von: Dhodapkar, Aditya, et al.
Veröffentlicht: (2026)
Geist in the Machine: Simulating Recognition and Inner Dialogue in AI-Mediated Teaching and Research
von: Magee, Liam
Veröffentlicht: (2026)
von: Magee, Liam
Veröffentlicht: (2026)
MDIA: A Multi-Agent Diagnostic Intelligence Pipeline on HealthBench Professional
von: Cruz, Roberto, et al.
Veröffentlicht: (2026)
von: Cruz, Roberto, et al.
Veröffentlicht: (2026)
ART: Adaptive Response Tuning Framework -- A Multi-Agent Tournament-Based Approach to LLM Response Optimization
von: Khan, Omer Jauhar
Veröffentlicht: (2025)
von: Khan, Omer Jauhar
Veröffentlicht: (2025)
Tool Receipts, Not Zero-Knowledge Proofs: Practical Hallucination Detection for AI Agents
von: Basu, Abhinaba
Veröffentlicht: (2026)
von: Basu, Abhinaba
Veröffentlicht: (2026)
Territory Paint Wars: Diagnosing and Mitigating Failure Modes in Competitive Multi-Agent PPO
von: Singh, Diyansha
Veröffentlicht: (2026)
von: Singh, Diyansha
Veröffentlicht: (2026)
CyberAId: AI-Driven Cybersecurity for Financial Service Providers
von: Fatouros, George, et al.
Veröffentlicht: (2026)
von: Fatouros, George, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Portable Agent Memory: A Protocol for Cryptographically-Verified Memory Transfer Across Heterogeneous AI Agents
von: Ravindran, Santhosh Kumar
Veröffentlicht: (2026) -
When the Agent Is the Adversary: Architectural Requirements for Agentic AI Containment After the April 2026 Frontier Model Escape
von: Mitchell, Richard Joseph
Veröffentlicht: (2026) -
Reservoir Computing with Evolved Critical Neural Cellular Automata
von: Pontes-Filho, Sidney, et al.
Veröffentlicht: (2025) -
Gradient Atoms: Unsupervised Discovery, Attribution and Steering of Model Behaviors via Sparse Decomposition of Training Gradients
von: Rosser, J
Veröffentlicht: (2026) -
Evolutionary Algorithms Approach For Search Based On Semantic Document Similarity
von: Muniyappa, Chandrashekar, et al.
Veröffentlicht: (2025)