StatePlane: A Cognitive State Plane for Long-Horizon AI Systems Under Bounded Context
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Annapureddy, Sasank, Mulcahy, John, Thamatani, Anjaneya Prasad |
|---|---|
| Format: | Preprint |
| Publié: |
2026
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
PRIMA: Operational Patterns for Resilient Multi-Agent Research with Verifiable Identity and Convergent Feedback
par: Annapureddy, Sasank
Publié: (2026)
par: Annapureddy, Sasank
Publié: (2026)
Rewarding Beliefs, Not Actions: Consistency-Guided Credit Assignment for Long-Horizon Agents
par: Tang, Wenjie, et autres
Publié: (2026)
par: Tang, Wenjie, et autres
Publié: (2026)
Robust and Diverse Multi-Agent Learning via Rational Policy Gradient
par: Lauffer, Niklas, et autres
Publié: (2025)
par: Lauffer, Niklas, et autres
Publié: (2025)
ChromaFlow: A Negative Ablation Study of Orchestration Overhead in Tool-Augmented Agent Evaluation
par: Mittal, Tarun
Publié: (2026)
par: Mittal, Tarun
Publié: (2026)
Advancing Multimodal Agent Reasoning with Long-Term Neuro-Symbolic Memory
par: Jiang, Rongjie, et autres
Publié: (2026)
par: Jiang, Rongjie, et autres
Publié: (2026)
When Can Human-AI Teams Outperform Individuals? Tight Bounds with Impossibility Guarantees
par: Guo, Dongxin, et autres
Publié: (2026)
par: Guo, Dongxin, et autres
Publié: (2026)
Instruction-Level Weight Shaping: A Framework for Self-Improving AI Agents
par: Costa, Rimom
Publié: (2025)
par: Costa, Rimom
Publié: (2025)
Safe and Policy-Compliant Multi-Agent Orchestration for Enterprise AI
par: Pasupuleti, Vinil, et autres
Publié: (2026)
par: Pasupuleti, Vinil, et autres
Publié: (2026)
Bimanual Robot Manipulation via Multi-Agent In-Context Learning
par: Palma, Alessio, et autres
Publié: (2026)
par: Palma, Alessio, et autres
Publié: (2026)
One Policy, Infinite NPCs: Persona-Traceable Shared RL Policies for Scalable Game Agents
par: Hong, Yoosung
Publié: (2026)
par: Hong, Yoosung
Publié: (2026)
Umwelt Engineering: Designing the Cognitive Worlds of Linguistic Agents
par: Jehu-Appiah, Rodney
Publié: (2026)
par: Jehu-Appiah, Rodney
Publié: (2026)
PARNESS: A Paper Harness for End-to-End Automated Scientific Research with Dynamic Workflows, Full-Text Indexing, and Cross-Run Knowledge Accumulation
par: Wang, Yuchen, et autres
Publié: (2026)
par: Wang, Yuchen, et autres
Publié: (2026)
Fuzzy, Symbolic, and Contextual: Enhancing LLM Instruction via Cognitive Scaffolding
par: Figueiredo, Vanessa
Publié: (2025)
par: Figueiredo, Vanessa
Publié: (2025)
When Outcome Looks Right But Discipline Fails: Trace-Based Evaluation Under Hidden Competitor State
par: Zhu, Peiying, et autres
Publié: (2026)
par: Zhu, Peiying, et autres
Publié: (2026)
A Framework for Assessing AI Agent Decisions and Outcomes in AutoML Pipelines
par: Du, Gaoyuan, et autres
Publié: (2026)
par: Du, Gaoyuan, et autres
Publié: (2026)
Knowledge Equivalence in Digital Twins of Intelligent Systems
par: Zhang, Nan, et autres
Publié: (2022)
par: Zhang, Nan, et autres
Publié: (2022)
SIA: Self Improving AI with Harness & Weight Updates
par: Hebbar, Prannay, et autres
Publié: (2026)
par: Hebbar, Prannay, et autres
Publié: (2026)
When Actions Disappear: Adversarial Action Removal in Self-Play Reinforcement Learning
par: Kujur, Arahan
Publié: (2026)
par: Kujur, Arahan
Publié: (2026)
Dynamic Dual-Granularity Skill Bank for Agentic RL
par: Tu, Songjun, et autres
Publié: (2026)
par: Tu, Songjun, et autres
Publié: (2026)
FlowSteer: Towards Agents Designing Agentic Workflows via Reinforced Progressive Canvas Editing
par: Zhang, Mingda, et autres
Publié: (2026)
par: Zhang, Mingda, et autres
Publié: (2026)
How Good is ChatGPT in Giving Adaptive Guidance Using Knowledge Graphs in E-Learning Environments?
par: Ocheja, Patrick, et autres
Publié: (2024)
par: Ocheja, Patrick, et autres
Publié: (2024)
When Agents Disagree: The Selection Bottleneck in Multi-Agent LLM Pipelines
par: Maryanskyy, Artem
Publié: (2026)
par: Maryanskyy, Artem
Publié: (2026)
Extending NGU to Multi-Agent RL: A Preliminary Study
par: Hernandez, Juan, et autres
Publié: (2025)
par: Hernandez, Juan, et autres
Publié: (2025)
The Stochastic Gap: A Markovian Framework for Pre-Deployment Reliability and Oversight-Cost Auditing in Agentic Artificial Intelligence
par: Pal, Biplab, et autres
Publié: (2026)
par: Pal, Biplab, et autres
Publié: (2026)
QTypeMix: Enhancing Multi-Agent Cooperative Strategies through Heterogeneous and Homogeneous Value Decomposition
par: Fu, Songchen, et autres
Publié: (2024)
par: Fu, Songchen, et autres
Publié: (2024)
Learning To Help: Training Models to Assist Legacy Devices
par: Wu, Yu, et autres
Publié: (2024)
par: Wu, Yu, et autres
Publié: (2024)
ME-IGM: Individual-Global-Max in Maximum Entropy Multi-Agent Reinforcement Learning
par: Chen, Wen-Tse, et autres
Publié: (2024)
par: Chen, Wen-Tse, et autres
Publié: (2024)
Analysing Factorizations of Action-Value Networks for Cooperative Multi-Agent Reinforcement Learning
par: Castellini, Jacopo, et autres
Publié: (2019)
par: Castellini, Jacopo, et autres
Publié: (2019)
Advancing Transformer Architecture in Long-Context Large Language Models: A Comprehensive Survey
par: Huang, Yunpeng, et autres
Publié: (2023)
par: Huang, Yunpeng, et autres
Publié: (2023)
N-Agent Ad Hoc Teamwork
par: Wang, Caroline, et autres
Publié: (2024)
par: Wang, Caroline, et autres
Publié: (2024)
Rethinking AI Hardware: A Three-Layer Cognitive Architecture for Autonomous Agents
par: Chen, Li
Publié: (2026)
par: Chen, Li
Publié: (2026)
Dynamic Attentional Context Scoping: Agent-Triggered Focus Sessions for Isolated Per-Agent Steering in Multi-Agent LLM Orchestration
par: Patel, Nickson
Publié: (2026)
par: Patel, Nickson
Publié: (2026)
Procedural Game Level Design with Deep Reinforcement Learning
par: Özkan, Miraç Buğra
Publié: (2025)
par: Özkan, Miraç Buğra
Publié: (2025)
PillagerBench: Benchmarking LLM-Based Agents in Competitive Minecraft Team Environments
par: Schipper, Olivier, et autres
Publié: (2025)
par: Schipper, Olivier, et autres
Publié: (2025)
Council Mode: A Heterogeneous Multi-Agent Consensus Framework for Reducing LLM Hallucination and Bias
par: Wu, Shuai, et autres
Publié: (2026)
par: Wu, Shuai, et autres
Publié: (2026)
Evolved Developmental Artificial Neural Networks for Multitasking with Advanced Activity Dependence
par: Zhang, Yintong, et autres
Publié: (2024)
par: Zhang, Yintong, et autres
Publié: (2024)
When to Forget: A Memory Governance Primitive
par: Simsek, Baris
Publié: (2026)
par: Simsek, Baris
Publié: (2026)
Portable Agent Memory: A Protocol for Cryptographically-Verified Memory Transfer Across Heterogeneous AI Agents
par: Ravindran, Santhosh Kumar
Publié: (2026)
par: Ravindran, Santhosh Kumar
Publié: (2026)
MindGames Arena Generalization Track: In2AI Solution with Delayed Per-Step Reward Attribution
par: Korshuk, Aliaksei, et autres
Publié: (2026)
par: Korshuk, Aliaksei, et autres
Publié: (2026)
Policy Search, Retrieval, and Composition via Task Similarity in Collaborative Agentic Systems
par: Nath, Saptarshi, et autres
Publié: (2025)
par: Nath, Saptarshi, et autres
Publié: (2025)
Documents similaires
-
PRIMA: Operational Patterns for Resilient Multi-Agent Research with Verifiable Identity and Convergent Feedback
par: Annapureddy, Sasank
Publié: (2026) -
Rewarding Beliefs, Not Actions: Consistency-Guided Credit Assignment for Long-Horizon Agents
par: Tang, Wenjie, et autres
Publié: (2026) -
Robust and Diverse Multi-Agent Learning via Rational Policy Gradient
par: Lauffer, Niklas, et autres
Publié: (2025) -
ChromaFlow: A Negative Ablation Study of Orchestration Overhead in Tool-Augmented Agent Evaluation
par: Mittal, Tarun
Publié: (2026) -
Advancing Multimodal Agent Reasoning with Long-Term Neuro-Symbolic Memory
par: Jiang, Rongjie, et autres
Publié: (2026)