Replayable Financial Agents: A Determinism-Faithfulness Assurance Harness for Tool-Using LLM Agents
Fuente:
arXiv
Gespeichert in:
| 1. Verfasser: | Khatchadourian, Raffi |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Agent Capsules: Quality-Gated Granularity Control for Multi-Agent LLM Pipelines
von: Ray, Aninda
Veröffentlicht: (2026)
von: Ray, Aninda
Veröffentlicht: (2026)
Tool Forge: A Validation-Carrying Toolchain for Governed Agentic Execution
von: Rao, Swanand
Veröffentlicht: (2026)
von: Rao, Swanand
Veröffentlicht: (2026)
A Framework for Assessing AI Agent Decisions and Outcomes in AutoML Pipelines
von: Du, Gaoyuan, et al.
Veröffentlicht: (2026)
von: Du, Gaoyuan, et al.
Veröffentlicht: (2026)
Test-Driven AI Agent Definition (TDAD): Compiling Tool-Using Agents from Behavioral Specifications
von: Rehan, Tzafrir
Veröffentlicht: (2026)
von: Rehan, Tzafrir
Veröffentlicht: (2026)
Iterative Audit Convergence in LLM-Managed Multi-Agent Systems: A Case Study in Prompt Engineering Quality Assurance
von: Calboreanu, Elias
Veröffentlicht: (2026)
von: Calboreanu, Elias
Veröffentlicht: (2026)
CodeTracer: Towards Traceable Agent States
von: Li, Han, et al.
Veröffentlicht: (2026)
von: Li, Han, et al.
Veröffentlicht: (2026)
An Organization-Scoped LLM Agent Runtime Architecture for Regulated Cybersecurity Operations
von: Fatouros, George, et al.
Veröffentlicht: (2026)
von: Fatouros, George, et al.
Veröffentlicht: (2026)
Ontology-Constrained Neural Reasoning in Enterprise Agentic Systems: A Neurosymbolic Architecture for Domain-Grounded AI Agents
von: Tuan, Thanh Luong, et al.
Veröffentlicht: (2026)
von: Tuan, Thanh Luong, et al.
Veröffentlicht: (2026)
Exploring Robust Multi-Agent Workflows for Environmental Data Management
von: Guan, Boyuan, et al.
Veröffentlicht: (2026)
von: Guan, Boyuan, et al.
Veröffentlicht: (2026)
SmellBench: Evaluating LLM Agents on Architectural Code Smell Repair
von: Dinu, Ion George, et al.
Veröffentlicht: (2026)
von: Dinu, Ion George, et al.
Veröffentlicht: (2026)
Refute-or-Promote: An Adversarial Stage-Gated Multi-Agent Review Methodology for High-Precision LLM-Assisted Defect Discovery
von: Agarwal, Abhinav
Veröffentlicht: (2026)
von: Agarwal, Abhinav
Veröffentlicht: (2026)
Automated structural testing of LLM-based agents: methods, framework, and case studies
von: Kohl, Jens, et al.
Veröffentlicht: (2026)
von: Kohl, Jens, et al.
Veröffentlicht: (2026)
CUJBench: Benchmarking LLM-Agent on Cross-Modal Failure Diagnosis from Browser to Backend
von: Meng, Haoming
Veröffentlicht: (2026)
von: Meng, Haoming
Veröffentlicht: (2026)
Monitoring Agentic Systems Before They're Reliable
von: Boston, Marisa Ferrara, et al.
Veröffentlicht: (2026)
von: Boston, Marisa Ferrara, et al.
Veröffentlicht: (2026)
ClawVM: Harness-Managed Virtual Memory for Stateful Tool-Using LLM Agents
von: Rafique, Mofasshara, et al.
Veröffentlicht: (2026)
von: Rafique, Mofasshara, et al.
Veröffentlicht: (2026)
PRIMA: Operational Patterns for Resilient Multi-Agent Research with Verifiable Identity and Convergent Feedback
von: Annapureddy, Sasank
Veröffentlicht: (2026)
von: Annapureddy, Sasank
Veröffentlicht: (2026)
CodeCRDT: Observation-Driven Coordination for Multi-Agent LLM Code Generation
von: Pugachev, Sergey
Veröffentlicht: (2025)
von: Pugachev, Sergey
Veröffentlicht: (2025)
Adversarial Feeds Steer LLM Agent Decisions Against Their Defaults
von: Usman, Rana Muhammad
Veröffentlicht: (2026)
von: Usman, Rana Muhammad
Veröffentlicht: (2026)
Bounded Autonomy for Enterprise AI: Typed Action Contracts and Consumer-Side Execution
von: Sohail, Sarmad, et al.
Veröffentlicht: (2026)
von: Sohail, Sarmad, et al.
Veröffentlicht: (2026)
AI Agentic workflows and Enterprise APIs: Adapting API architectures for the age of AI agents
von: Tupe, Vaibhav, et al.
Veröffentlicht: (2025)
von: Tupe, Vaibhav, et al.
Veröffentlicht: (2025)
elsciRL: Integrating Language Solutions into Reinforcement Learning Problem Settings
von: Osborne, Philip, et al.
Veröffentlicht: (2025)
von: Osborne, Philip, et al.
Veröffentlicht: (2025)
Council Mode: A Heterogeneous Multi-Agent Consensus Framework for Reducing LLM Hallucination and Bias
von: Wu, Shuai, et al.
Veröffentlicht: (2026)
von: Wu, Shuai, et al.
Veröffentlicht: (2026)
ClawHub Security Signals: When VirusTotal, Static Analysis, and SkillSpector Disagree
von: Koc, Vincent, et al.
Veröffentlicht: (2026)
von: Koc, Vincent, et al.
Veröffentlicht: (2026)
When the Agent Is the Adversary: Architectural Requirements for Agentic AI Containment After the April 2026 Frontier Model Escape
von: Mitchell, Richard Joseph
Veröffentlicht: (2026)
von: Mitchell, Richard Joseph
Veröffentlicht: (2026)
Owner-Harm: A Missing Threat Model for AI Agent Safety
von: Zhang, Dongcheng, et al.
Veröffentlicht: (2026)
von: Zhang, Dongcheng, et al.
Veröffentlicht: (2026)
Which LLM Multi-Agent Protocol to Choose?
von: Du, Hongyi, et al.
Veröffentlicht: (2025)
von: Du, Hongyi, et al.
Veröffentlicht: (2025)
Building a Stable Planner: An Extended Finite State Machine Based Planning Module for Mobile GUI Agent
von: Mo, Fanglin, et al.
Veröffentlicht: (2025)
von: Mo, Fanglin, et al.
Veröffentlicht: (2025)
XARP Tools: An Extended Reality Platform for Humans and AI Agents
von: Caetano, Arthur, et al.
Veröffentlicht: (2025)
von: Caetano, Arthur, et al.
Veröffentlicht: (2025)
Toolsuite for Implementing Multiagent Systems Based on Communication Protocols
von: Chopra, Amit K., et al.
Veröffentlicht: (2025)
von: Chopra, Amit K., et al.
Veröffentlicht: (2025)
ART: Adaptive Response Tuning Framework -- A Multi-Agent Tournament-Based Approach to LLM Response Optimization
von: Khan, Omer Jauhar
Veröffentlicht: (2025)
von: Khan, Omer Jauhar
Veröffentlicht: (2025)
Rewarding Beliefs, Not Actions: Consistency-Guided Credit Assignment for Long-Horizon Agents
von: Tang, Wenjie, et al.
Veröffentlicht: (2026)
von: Tang, Wenjie, et al.
Veröffentlicht: (2026)
Umwelt Engineering: Designing the Cognitive Worlds of Linguistic Agents
von: Jehu-Appiah, Rodney
Veröffentlicht: (2026)
von: Jehu-Appiah, Rodney
Veröffentlicht: (2026)
Task-Level AI Readiness Assessment for Business Process Management:The T-IPO Model and LARA Matrix in Financial-Services IT Operations
von: Li, Mingjun, et al.
Veröffentlicht: (2026)
von: Li, Mingjun, et al.
Veröffentlicht: (2026)
SIA: Self Improving AI with Harness & Weight Updates
von: Hebbar, Prannay, et al.
Veröffentlicht: (2026)
von: Hebbar, Prannay, et al.
Veröffentlicht: (2026)
Applying Cognitive Design Patterns to General LLM Agents
von: Wray, Robert E., et al.
Veröffentlicht: (2025)
von: Wray, Robert E., et al.
Veröffentlicht: (2025)
From Helpfulness to Toxic Proactivity: Diagnosing Behavioral Misalignment in LLM Agents
von: Wang, Xinyue, et al.
Veröffentlicht: (2026)
von: Wang, Xinyue, et al.
Veröffentlicht: (2026)
ChromaFlow: A Negative Ablation Study of Orchestration Overhead in Tool-Augmented Agent Evaluation
von: Mittal, Tarun
Veröffentlicht: (2026)
von: Mittal, Tarun
Veröffentlicht: (2026)
Governance-Aware Vector Subscriptions for Multi-Agent Knowledge Ecosystems
von: Johnson, Steven
Veröffentlicht: (2026)
von: Johnson, Steven
Veröffentlicht: (2026)
Retrieval-Conditioned Topology Selection with Provable Budget Conservation for Multi-Agent Code Generation
von: Talluri, Abhijit, et al.
Veröffentlicht: (2026)
von: Talluri, Abhijit, et al.
Veröffentlicht: (2026)
A Pattern Language for Resilient Visual Agents
von: Gidey, Habtom Kahsay, et al.
Veröffentlicht: (2026)
von: Gidey, Habtom Kahsay, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Agent Capsules: Quality-Gated Granularity Control for Multi-Agent LLM Pipelines
von: Ray, Aninda
Veröffentlicht: (2026) -
Tool Forge: A Validation-Carrying Toolchain for Governed Agentic Execution
von: Rao, Swanand
Veröffentlicht: (2026) -
A Framework for Assessing AI Agent Decisions and Outcomes in AutoML Pipelines
von: Du, Gaoyuan, et al.
Veröffentlicht: (2026) -
Test-Driven AI Agent Definition (TDAD): Compiling Tool-Using Agents from Behavioral Specifications
von: Rehan, Tzafrir
Veröffentlicht: (2026) -
Iterative Audit Convergence in LLM-Managed Multi-Agent Systems: A Case Study in Prompt Engineering Quality Assurance
von: Calboreanu, Elias
Veröffentlicht: (2026)