Robust Agent Compensation (RAC): Teaching AI Agents to Compensate
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Perera, Srinath, Hapuarachchi, Kaviru, Leymann, Frank, Khalaf, Rania |
|---|---|
| Format: | Preprint |
| Publié: |
2026
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Test-Driven AI Agent Definition (TDAD): Compiling Tool-Using Agents from Behavioral Specifications
par: Rehan, Tzafrir
Publié: (2026)
par: Rehan, Tzafrir
Publié: (2026)
Social, Legal, Ethical, Empathetic and Cultural Norm Operationalisation for AI Agents
par: Calinescu, Radu, et autres
Publié: (2026)
par: Calinescu, Radu, et autres
Publié: (2026)
AgentAssay: Token-Efficient Regression Testing for Non-Deterministic AI Agent Workflows
par: Bhardwaj, Varun Pratap
Publié: (2026)
par: Bhardwaj, Varun Pratap
Publié: (2026)
TML-Bench: Benchmark for Data Science Agents on Tabular ML Tasks
par: Pinchuk, Mykola
Publié: (2026)
par: Pinchuk, Mykola
Publié: (2026)
N-Version Assessment and Enhancement of Generative AI
par: Kessel, Marcus, et autres
Publié: (2024)
par: Kessel, Marcus, et autres
Publié: (2024)
Behavioral Fingerprints for LLM Endpoint Stability and Identity
par: Leshin, Jonah, et autres
Publié: (2026)
par: Leshin, Jonah, et autres
Publié: (2026)
Who is Introducing the Failure? Automatically Attributing Failures of Multi-Agent Systems via Spectrum Analysis
par: Ge, Yu, et autres
Publié: (2025)
par: Ge, Yu, et autres
Publié: (2025)
Comparative Analysis of AI Agent Architectures for Entity Relationship Classification
par: Berijanian, Maryam, et autres
Publié: (2025)
par: Berijanian, Maryam, et autres
Publié: (2025)
AI-Assisted Engineering Should Track the Epistemic Status and Temporal Validity of Architectural Decisions
par: Gilda, Sankalp, et autres
Publié: (2026)
par: Gilda, Sankalp, et autres
Publié: (2026)
Memory Management and Contextual Consistency for Long-Running Low-Code Agents
par: Xu, Jiexi
Publié: (2025)
par: Xu, Jiexi
Publié: (2025)
You Don't Need Public Tests to Generate Correct Code
par: Silva, Kaushitha, et autres
Publié: (2026)
par: Silva, Kaushitha, et autres
Publié: (2026)
Nemobot Games: Crafting Strategic AI Gaming Agents for Interactive Learning with Large Language Models
par: Tan, Chee Wei, et autres
Publié: (2026)
par: Tan, Chee Wei, et autres
Publié: (2026)
CVE-Bench: A Benchmark for AI Agents' Ability to Exploit Real-World Web Application Vulnerabilities
par: Zhu, Yuxuan, et autres
Publié: (2025)
par: Zhu, Yuxuan, et autres
Publié: (2025)
N-Agent Ad Hoc Teamwork
par: Wang, Caroline, et autres
Publié: (2024)
par: Wang, Caroline, et autres
Publié: (2024)
Deciphering Digital Detectives: Understanding LLM Behaviors and Capabilities in Multi-Agent Mystery Games
par: Wu, Dekun, et autres
Publié: (2023)
par: Wu, Dekun, et autres
Publié: (2023)
Structural Quality Gaps in Practitioner AI Governance Prompts: An Empirical Study Using a Five-Principle Evaluation Framework
par: Zietsman, Christo
Publié: (2026)
par: Zietsman, Christo
Publié: (2026)
One Policy, Infinite NPCs: Persona-Traceable Shared RL Policies for Scalable Game Agents
par: Hong, Yoosung
Publié: (2026)
par: Hong, Yoosung
Publié: (2026)
Reasoning Provenance for Autonomous AI Agents: Structured Behavioral Analytics Beyond State Checkpoints and Execution Traces
par: Vispute, Neelmani, et autres
Publié: (2026)
par: Vispute, Neelmani, et autres
Publié: (2026)
Multi-Agent Pathfinding with Non-Unit Integer Edge Costs via Enhanced Conflict-Based Search and Graph Discretization
par: Fan, Hongkai, et autres
Publié: (2026)
par: Fan, Hongkai, et autres
Publié: (2026)
Test-driven Software Experimentation with LASSO: an LLM Prompt Benchmarking Example
par: Kessel, Marcus
Publié: (2024)
par: Kessel, Marcus
Publié: (2024)
Morescient GAI for Software Engineering (Extended Version)
par: Kessel, Marcus, et autres
Publié: (2024)
par: Kessel, Marcus, et autres
Publié: (2024)
Introducing Brain-like Concepts to Embodied Hand-crafted Dialog Management System
par: Joublin, Frank, et autres
Publié: (2024)
par: Joublin, Frank, et autres
Publié: (2024)
Collaborative AI Enhances Image Understanding in Materials Science
par: Yin, Ruoyan Avery, et autres
Publié: (2025)
par: Yin, Ruoyan Avery, et autres
Publié: (2025)
Benchmarking AI for low-resource contexts: Thinking beyond leaderboards
par: Pant, Aakash, et autres
Publié: (2026)
par: Pant, Aakash, et autres
Publié: (2026)
PestMA: LLM-based Multi-Agent System for Informed Pest Management
par: Shi, Hongrui, et autres
Publié: (2025)
par: Shi, Hongrui, et autres
Publié: (2025)
Towards Single-System Illusion in Software-Defined Vehicles -- Automated, AI-Powered Workflow
par: Lebioda, Krzysztof, et autres
Publié: (2024)
par: Lebioda, Krzysztof, et autres
Publié: (2024)
SLEGO: A Collaborative Data Analytics System with LLM Recommender for Diverse Users
par: Ng, Siu Lung, et autres
Publié: (2024)
par: Ng, Siu Lung, et autres
Publié: (2024)
Agent-Aided Design for Dynamic CAD Models
par: Adler, Mitch, et autres
Publié: (2026)
par: Adler, Mitch, et autres
Publié: (2026)
Federated Learning and AI Regulation in the European Union: Who is Responsible? -- An Interdisciplinary Analysis
par: Woisetschläger, Herbert, et autres
Publié: (2024)
par: Woisetschläger, Herbert, et autres
Publié: (2024)
AI Playing Business Games: Benchmarking Large Language Models on Managerial Decision-Making in Dynamic Simulations
par: Ovezmyradov, Berdymyrat
Publié: (2025)
par: Ovezmyradov, Berdymyrat
Publié: (2025)
Reconsidering Requirements Engineering: Human-AI Collaboration in AI-Native Software Development
par: Abbasi, Mateen Ahmed, et autres
Publié: (2025)
par: Abbasi, Mateen Ahmed, et autres
Publié: (2025)
Bounded Autonomy for Enterprise AI: Typed Action Contracts and Consumer-Side Execution
par: Sohail, Sarmad, et autres
Publié: (2026)
par: Sohail, Sarmad, et autres
Publié: (2026)
Augmenting deep neural networks with symbolic knowledge: Towards trustworthy and interpretable AI for education
par: Hooshyar, Danial, et autres
Publié: (2023)
par: Hooshyar, Danial, et autres
Publié: (2023)
EvoPat: A Multi-LLM-based Patents Summarization and Analysis Agent
par: Wang, Suyuan, et autres
Publié: (2024)
par: Wang, Suyuan, et autres
Publié: (2024)
Smart Expansion Techniques for ASP-based Interactive Configuration
par: Balážová, Lucia, et autres
Publié: (2025)
par: Balážová, Lucia, et autres
Publié: (2025)
TelePlanNet: An AI-Driven Framework for Efficient Telecom Network Planning
par: Deng, Zongyuan, et autres
Publié: (2025)
par: Deng, Zongyuan, et autres
Publié: (2025)
Reinforcement of Explainability of ChatGPT Prompts by Embedding Breast Cancer Self-Screening Rules into AI Responses
par: Khan, Yousef, et autres
Publié: (2024)
par: Khan, Yousef, et autres
Publié: (2024)
Exact Synthetic Populations for Scalable Societal and Market Modeling
par: Petit, Thierry, et autres
Publié: (2025)
par: Petit, Thierry, et autres
Publié: (2025)
Multi-Agent Design Assistant for the Simulation of Inertial Fusion Energy
par: Shachar, Meir H., et autres
Publié: (2025)
par: Shachar, Meir H., et autres
Publié: (2025)
MedMemoryBench: Benchmarking Agent Memory in Personalized Healthcare
par: Wang, Yihao, et autres
Publié: (2026)
par: Wang, Yihao, et autres
Publié: (2026)
Documents similaires
-
Test-Driven AI Agent Definition (TDAD): Compiling Tool-Using Agents from Behavioral Specifications
par: Rehan, Tzafrir
Publié: (2026) -
Social, Legal, Ethical, Empathetic and Cultural Norm Operationalisation for AI Agents
par: Calinescu, Radu, et autres
Publié: (2026) -
AgentAssay: Token-Efficient Regression Testing for Non-Deterministic AI Agent Workflows
par: Bhardwaj, Varun Pratap
Publié: (2026) -
TML-Bench: Benchmark for Data Science Agents on Tabular ML Tasks
par: Pinchuk, Mykola
Publié: (2026) -
N-Version Assessment and Enhancement of Generative AI
par: Kessel, Marcus, et autres
Publié: (2024)