Gespeichert in:
| Hauptverfasser: | Koaik, Fatima, Gupta, Aayush, Sheikh, Farahan Raza |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2601.06111 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
TaxAgent: How Large Language Model Designs Fiscal Policy
von: Wang, Jizhou, et al.
Veröffentlicht: (2025)
von: Wang, Jizhou, et al.
Veröffentlicht: (2025)
Safe and Policy-Compliant Multi-Agent Orchestration for Enterprise AI
von: Pasupuleti, Vinil, et al.
Veröffentlicht: (2026)
von: Pasupuleti, Vinil, et al.
Veröffentlicht: (2026)
From Idea to CAD: A Language Model-Driven Multi-Agent System for Collaborative Design
von: Ocker, Felix, et al.
Veröffentlicht: (2025)
von: Ocker, Felix, et al.
Veröffentlicht: (2025)
Knowledge Equivalence in Digital Twins of Intelligent Systems
von: Zhang, Nan, et al.
Veröffentlicht: (2022)
von: Zhang, Nan, et al.
Veröffentlicht: (2022)
LLM Scalability Risk for Agentic-AI and Model Supply Chain Security
von: Ahi, Kiarash, et al.
Veröffentlicht: (2026)
von: Ahi, Kiarash, et al.
Veröffentlicht: (2026)
TwinLoop: Simulation-in-the-Loop Digital Twins for Online Multi-Agent Reinforcement Learning
von: Zhang, Nan, et al.
Veröffentlicht: (2026)
von: Zhang, Nan, et al.
Veröffentlicht: (2026)
Bounded Autonomy for Enterprise AI: Typed Action Contracts and Consumer-Side Execution
von: Sohail, Sarmad, et al.
Veröffentlicht: (2026)
von: Sohail, Sarmad, et al.
Veröffentlicht: (2026)
TinyTroupe: An LLM-powered Multiagent Persona Simulation Toolkit
von: Salem, Paulo, et al.
Veröffentlicht: (2025)
von: Salem, Paulo, et al.
Veröffentlicht: (2025)
The Stochastic Gap: A Markovian Framework for Pre-Deployment Reliability and Oversight-Cost Auditing in Agentic Artificial Intelligence
von: Pal, Biplab, et al.
Veröffentlicht: (2026)
von: Pal, Biplab, et al.
Veröffentlicht: (2026)
Rewarding Beliefs, Not Actions: Consistency-Guided Credit Assignment for Long-Horizon Agents
von: Tang, Wenjie, et al.
Veröffentlicht: (2026)
von: Tang, Wenjie, et al.
Veröffentlicht: (2026)
The Importance of Out-of-Band Metadata for Safe Autonomous Agents: The Redpanda Agentic Data Plane
von: Akidau, Tyler, et al.
Veröffentlicht: (2026)
von: Akidau, Tyler, et al.
Veröffentlicht: (2026)
Research on Security Enhancement Methods for Adversarial Robust Large Language Model Intelligent Agents for Medical Decision-Making Tasks
von: Hu, Saisai
Veröffentlicht: (2026)
von: Hu, Saisai
Veröffentlicht: (2026)
The Power of Stories: Narrative Priming Shapes How LLM Agents Collaborate and Compete
von: Großmann, Gerrit, et al.
Veröffentlicht: (2025)
von: Großmann, Gerrit, et al.
Veröffentlicht: (2025)
Beyond Benchmark Islands: Toward Representative Trustworthiness Evaluation for Agentic AI
von: Qi, Jinhu, et al.
Veröffentlicht: (2026)
von: Qi, Jinhu, et al.
Veröffentlicht: (2026)
AI Agentic workflows and Enterprise APIs: Adapting API architectures for the age of AI agents
von: Tupe, Vaibhav, et al.
Veröffentlicht: (2025)
von: Tupe, Vaibhav, et al.
Veröffentlicht: (2025)
Security Attack and Defense Strategies for Autonomous Agent Frameworks: A Layered Review with OpenClaw as a Case Study
von: Xu, Luyao, et al.
Veröffentlicht: (2026)
von: Xu, Luyao, et al.
Veröffentlicht: (2026)
Governance Architecture for Autonomous Agent Systems: Threats, Framework, and Engineering Practice
von: Ge, Yuxu
Veröffentlicht: (2026)
von: Ge, Yuxu
Veröffentlicht: (2026)
HBEE: Human Behavioral Entropy Engine -- Pre-Registered Multi-Agent LLM Simulation of Peer-Suspicion-Based Detection Inversion
von: Ferrel, Vickson
Veröffentlicht: (2026)
von: Ferrel, Vickson
Veröffentlicht: (2026)
Adversarial Feeds Steer LLM Agent Decisions Against Their Defaults
von: Usman, Rana Muhammad
Veröffentlicht: (2026)
von: Usman, Rana Muhammad
Veröffentlicht: (2026)
PARNESS: A Paper Harness for End-to-End Automated Scientific Research with Dynamic Workflows, Full-Text Indexing, and Cross-Run Knowledge Accumulation
von: Wang, Yuchen, et al.
Veröffentlicht: (2026)
von: Wang, Yuchen, et al.
Veröffentlicht: (2026)
Depth-Dependent Indirect Prompt Injection in Tool-Calling ReAct Agents: Injection Depth, Payload Framing, and Turn-Budget Sensitivity
von: Rashidi, Mohammadreza
Veröffentlicht: (2026)
von: Rashidi, Mohammadreza
Veröffentlicht: (2026)
Go Big or Go Home: Simulating Mobbing Behavior with Braitenbergian Robots
von: Sanoubari, Elaheh
Veröffentlicht: (2026)
von: Sanoubari, Elaheh
Veröffentlicht: (2026)
Robust and Diverse Multi-Agent Learning via Rational Policy Gradient
von: Lauffer, Niklas, et al.
Veröffentlicht: (2025)
von: Lauffer, Niklas, et al.
Veröffentlicht: (2025)
PilotBench: A Benchmark for General Aviation Agents with Safety Constraints
von: Wu, Yalun, et al.
Veröffentlicht: (2026)
von: Wu, Yalun, et al.
Veröffentlicht: (2026)
Design Principles for the Construction of a Benchmark Evaluating Security Operation Capabilities of Multi-agent AI Systems
von: Cai, Yicheng, et al.
Veröffentlicht: (2026)
von: Cai, Yicheng, et al.
Veröffentlicht: (2026)
An Agentic Multi-Agent Architecture for Cybersecurity Risk Management
von: Gupta, Ravish, et al.
Veröffentlicht: (2026)
von: Gupta, Ravish, et al.
Veröffentlicht: (2026)
Real-Time In Silico Modeling of Postprandial Macronutrient Kinetics: A Validated Computational Engine for Nutrition Research and Digital Health
von: Calderone, Alberto
Veröffentlicht: (2026)
von: Calderone, Alberto
Veröffentlicht: (2026)
One Policy, Infinite NPCs: Persona-Traceable Shared RL Policies for Scalable Game Agents
von: Hong, Yoosung
Veröffentlicht: (2026)
von: Hong, Yoosung
Veröffentlicht: (2026)
CPEMH: An Agentic Framework for Prompt-Driven Behavior Evaluation and Assurance in Foundation-Model Systems for Mental Health Screening
von: Lorenzoni, Giuliano, et al.
Veröffentlicht: (2026)
von: Lorenzoni, Giuliano, et al.
Veröffentlicht: (2026)
Integrating Anomaly Detection into Agentic AI for Proactive Risk Management in Human Activity
von: Zorriassatine, Farbod, et al.
Veröffentlicht: (2026)
von: Zorriassatine, Farbod, et al.
Veröffentlicht: (2026)
When Outcome Looks Right But Discipline Fails: Trace-Based Evaluation Under Hidden Competitor State
von: Zhu, Peiying, et al.
Veröffentlicht: (2026)
von: Zhu, Peiying, et al.
Veröffentlicht: (2026)
HybridVFL: Disentangled Feature Learning for Edge-Enabled Vertical Federated Multimodal Classification
von: Anoosha, Mostafa, et al.
Veröffentlicht: (2025)
von: Anoosha, Mostafa, et al.
Veröffentlicht: (2025)
When the Agent Is the Adversary: Architectural Requirements for Agentic AI Containment After the April 2026 Frontier Model Escape
von: Mitchell, Richard Joseph
Veröffentlicht: (2026)
von: Mitchell, Richard Joseph
Veröffentlicht: (2026)
The Effect of State Representation on LLM Agent Behavior in Dynamic Routing Games
von: Goodyear, Lyle, et al.
Veröffentlicht: (2025)
von: Goodyear, Lyle, et al.
Veröffentlicht: (2025)
Difference Rewards Policy Gradients
von: Castellini, Jacopo, et al.
Veröffentlicht: (2020)
von: Castellini, Jacopo, et al.
Veröffentlicht: (2020)
Collaborative On-Sensor Array Cameras
von: Sun, Jipeng, et al.
Veröffentlicht: (2025)
von: Sun, Jipeng, et al.
Veröffentlicht: (2025)
Super-additive Cooperation in Language Model Agents
von: Tonini, Filippo, et al.
Veröffentlicht: (2025)
von: Tonini, Filippo, et al.
Veröffentlicht: (2025)
Instruction-Level Weight Shaping: A Framework for Self-Improving AI Agents
von: Costa, Rimom
Veröffentlicht: (2025)
von: Costa, Rimom
Veröffentlicht: (2025)
ILION: Deterministic Pre-Execution Safety Gates for Agentic AI Systems
von: Chitan, Florin Adrian
Veröffentlicht: (2026)
von: Chitan, Florin Adrian
Veröffentlicht: (2026)
Session Risk Memory (SRM): Temporal Authorization for Deterministic Pre-Execution Safety Gates
von: Chitan, Florin Adrian
Veröffentlicht: (2026)
von: Chitan, Florin Adrian
Veröffentlicht: (2026)
Ähnliche Einträge
-
TaxAgent: How Large Language Model Designs Fiscal Policy
von: Wang, Jizhou, et al.
Veröffentlicht: (2025) -
Safe and Policy-Compliant Multi-Agent Orchestration for Enterprise AI
von: Pasupuleti, Vinil, et al.
Veröffentlicht: (2026) -
From Idea to CAD: A Language Model-Driven Multi-Agent System for Collaborative Design
von: Ocker, Felix, et al.
Veröffentlicht: (2025) -
Knowledge Equivalence in Digital Twins of Intelligent Systems
von: Zhang, Nan, et al.
Veröffentlicht: (2022) -
LLM Scalability Risk for Agentic-AI and Model Supply Chain Security
von: Ahi, Kiarash, et al.
Veröffentlicht: (2026)