Guardado en:
| Autores principales: | Koaik, Fatima, Gupta, Aayush, Sheikh, Farahan Raza |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | https://arxiv.org/abs/2601.06111 |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
TaxAgent: How Large Language Model Designs Fiscal Policy
por: Wang, Jizhou, et al.
Publicado: (2025)
por: Wang, Jizhou, et al.
Publicado: (2025)
Safe and Policy-Compliant Multi-Agent Orchestration for Enterprise AI
por: Pasupuleti, Vinil, et al.
Publicado: (2026)
por: Pasupuleti, Vinil, et al.
Publicado: (2026)
From Idea to CAD: A Language Model-Driven Multi-Agent System for Collaborative Design
por: Ocker, Felix, et al.
Publicado: (2025)
por: Ocker, Felix, et al.
Publicado: (2025)
Knowledge Equivalence in Digital Twins of Intelligent Systems
por: Zhang, Nan, et al.
Publicado: (2022)
por: Zhang, Nan, et al.
Publicado: (2022)
LLM Scalability Risk for Agentic-AI and Model Supply Chain Security
por: Ahi, Kiarash, et al.
Publicado: (2026)
por: Ahi, Kiarash, et al.
Publicado: (2026)
TwinLoop: Simulation-in-the-Loop Digital Twins for Online Multi-Agent Reinforcement Learning
por: Zhang, Nan, et al.
Publicado: (2026)
por: Zhang, Nan, et al.
Publicado: (2026)
Bounded Autonomy for Enterprise AI: Typed Action Contracts and Consumer-Side Execution
por: Sohail, Sarmad, et al.
Publicado: (2026)
por: Sohail, Sarmad, et al.
Publicado: (2026)
TinyTroupe: An LLM-powered Multiagent Persona Simulation Toolkit
por: Salem, Paulo, et al.
Publicado: (2025)
por: Salem, Paulo, et al.
Publicado: (2025)
The Stochastic Gap: A Markovian Framework for Pre-Deployment Reliability and Oversight-Cost Auditing in Agentic Artificial Intelligence
por: Pal, Biplab, et al.
Publicado: (2026)
por: Pal, Biplab, et al.
Publicado: (2026)
Rewarding Beliefs, Not Actions: Consistency-Guided Credit Assignment for Long-Horizon Agents
por: Tang, Wenjie, et al.
Publicado: (2026)
por: Tang, Wenjie, et al.
Publicado: (2026)
The Importance of Out-of-Band Metadata for Safe Autonomous Agents: The Redpanda Agentic Data Plane
por: Akidau, Tyler, et al.
Publicado: (2026)
por: Akidau, Tyler, et al.
Publicado: (2026)
Research on Security Enhancement Methods for Adversarial Robust Large Language Model Intelligent Agents for Medical Decision-Making Tasks
por: Hu, Saisai
Publicado: (2026)
por: Hu, Saisai
Publicado: (2026)
The Power of Stories: Narrative Priming Shapes How LLM Agents Collaborate and Compete
por: Großmann, Gerrit, et al.
Publicado: (2025)
por: Großmann, Gerrit, et al.
Publicado: (2025)
Beyond Benchmark Islands: Toward Representative Trustworthiness Evaluation for Agentic AI
por: Qi, Jinhu, et al.
Publicado: (2026)
por: Qi, Jinhu, et al.
Publicado: (2026)
AI Agentic workflows and Enterprise APIs: Adapting API architectures for the age of AI agents
por: Tupe, Vaibhav, et al.
Publicado: (2025)
por: Tupe, Vaibhav, et al.
Publicado: (2025)
Security Attack and Defense Strategies for Autonomous Agent Frameworks: A Layered Review with OpenClaw as a Case Study
por: Xu, Luyao, et al.
Publicado: (2026)
por: Xu, Luyao, et al.
Publicado: (2026)
Governance Architecture for Autonomous Agent Systems: Threats, Framework, and Engineering Practice
por: Ge, Yuxu
Publicado: (2026)
por: Ge, Yuxu
Publicado: (2026)
HBEE: Human Behavioral Entropy Engine -- Pre-Registered Multi-Agent LLM Simulation of Peer-Suspicion-Based Detection Inversion
por: Ferrel, Vickson
Publicado: (2026)
por: Ferrel, Vickson
Publicado: (2026)
Adversarial Feeds Steer LLM Agent Decisions Against Their Defaults
por: Usman, Rana Muhammad
Publicado: (2026)
por: Usman, Rana Muhammad
Publicado: (2026)
PARNESS: A Paper Harness for End-to-End Automated Scientific Research with Dynamic Workflows, Full-Text Indexing, and Cross-Run Knowledge Accumulation
por: Wang, Yuchen, et al.
Publicado: (2026)
por: Wang, Yuchen, et al.
Publicado: (2026)
Depth-Dependent Indirect Prompt Injection in Tool-Calling ReAct Agents: Injection Depth, Payload Framing, and Turn-Budget Sensitivity
por: Rashidi, Mohammadreza
Publicado: (2026)
por: Rashidi, Mohammadreza
Publicado: (2026)
Go Big or Go Home: Simulating Mobbing Behavior with Braitenbergian Robots
por: Sanoubari, Elaheh
Publicado: (2026)
por: Sanoubari, Elaheh
Publicado: (2026)
Robust and Diverse Multi-Agent Learning via Rational Policy Gradient
por: Lauffer, Niklas, et al.
Publicado: (2025)
por: Lauffer, Niklas, et al.
Publicado: (2025)
PilotBench: A Benchmark for General Aviation Agents with Safety Constraints
por: Wu, Yalun, et al.
Publicado: (2026)
por: Wu, Yalun, et al.
Publicado: (2026)
Design Principles for the Construction of a Benchmark Evaluating Security Operation Capabilities of Multi-agent AI Systems
por: Cai, Yicheng, et al.
Publicado: (2026)
por: Cai, Yicheng, et al.
Publicado: (2026)
An Agentic Multi-Agent Architecture for Cybersecurity Risk Management
por: Gupta, Ravish, et al.
Publicado: (2026)
por: Gupta, Ravish, et al.
Publicado: (2026)
Real-Time In Silico Modeling of Postprandial Macronutrient Kinetics: A Validated Computational Engine for Nutrition Research and Digital Health
por: Calderone, Alberto
Publicado: (2026)
por: Calderone, Alberto
Publicado: (2026)
One Policy, Infinite NPCs: Persona-Traceable Shared RL Policies for Scalable Game Agents
por: Hong, Yoosung
Publicado: (2026)
por: Hong, Yoosung
Publicado: (2026)
CPEMH: An Agentic Framework for Prompt-Driven Behavior Evaluation and Assurance in Foundation-Model Systems for Mental Health Screening
por: Lorenzoni, Giuliano, et al.
Publicado: (2026)
por: Lorenzoni, Giuliano, et al.
Publicado: (2026)
Integrating Anomaly Detection into Agentic AI for Proactive Risk Management in Human Activity
por: Zorriassatine, Farbod, et al.
Publicado: (2026)
por: Zorriassatine, Farbod, et al.
Publicado: (2026)
When Outcome Looks Right But Discipline Fails: Trace-Based Evaluation Under Hidden Competitor State
por: Zhu, Peiying, et al.
Publicado: (2026)
por: Zhu, Peiying, et al.
Publicado: (2026)
HybridVFL: Disentangled Feature Learning for Edge-Enabled Vertical Federated Multimodal Classification
por: Anoosha, Mostafa, et al.
Publicado: (2025)
por: Anoosha, Mostafa, et al.
Publicado: (2025)
When the Agent Is the Adversary: Architectural Requirements for Agentic AI Containment After the April 2026 Frontier Model Escape
por: Mitchell, Richard Joseph
Publicado: (2026)
por: Mitchell, Richard Joseph
Publicado: (2026)
The Effect of State Representation on LLM Agent Behavior in Dynamic Routing Games
por: Goodyear, Lyle, et al.
Publicado: (2025)
por: Goodyear, Lyle, et al.
Publicado: (2025)
Difference Rewards Policy Gradients
por: Castellini, Jacopo, et al.
Publicado: (2020)
por: Castellini, Jacopo, et al.
Publicado: (2020)
Collaborative On-Sensor Array Cameras
por: Sun, Jipeng, et al.
Publicado: (2025)
por: Sun, Jipeng, et al.
Publicado: (2025)
Super-additive Cooperation in Language Model Agents
por: Tonini, Filippo, et al.
Publicado: (2025)
por: Tonini, Filippo, et al.
Publicado: (2025)
Instruction-Level Weight Shaping: A Framework for Self-Improving AI Agents
por: Costa, Rimom
Publicado: (2025)
por: Costa, Rimom
Publicado: (2025)
ILION: Deterministic Pre-Execution Safety Gates for Agentic AI Systems
por: Chitan, Florin Adrian
Publicado: (2026)
por: Chitan, Florin Adrian
Publicado: (2026)
Session Risk Memory (SRM): Temporal Authorization for Deterministic Pre-Execution Safety Gates
por: Chitan, Florin Adrian
Publicado: (2026)
por: Chitan, Florin Adrian
Publicado: (2026)
Ejemplares similares
-
TaxAgent: How Large Language Model Designs Fiscal Policy
por: Wang, Jizhou, et al.
Publicado: (2025) -
Safe and Policy-Compliant Multi-Agent Orchestration for Enterprise AI
por: Pasupuleti, Vinil, et al.
Publicado: (2026) -
From Idea to CAD: A Language Model-Driven Multi-Agent System for Collaborative Design
por: Ocker, Felix, et al.
Publicado: (2025) -
Knowledge Equivalence in Digital Twins of Intelligent Systems
por: Zhang, Nan, et al.
Publicado: (2022) -
LLM Scalability Risk for Agentic-AI and Model Supply Chain Security
por: Ahi, Kiarash, et al.
Publicado: (2026)