One Policy, Infinite NPCs: Persona-Traceable Shared RL Policies for Scalable Game Agents
Fuente:
arXiv
Saved in:
| Main Author: | Hong, Yoosung |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
N-Agent Ad Hoc Teamwork
by: Wang, Caroline, et al.
Published: (2024)
by: Wang, Caroline, et al.
Published: (2024)
Robust and Diverse Multi-Agent Learning via Rational Policy Gradient
by: Lauffer, Niklas, et al.
Published: (2025)
by: Lauffer, Niklas, et al.
Published: (2025)
ROTATE: Regret-driven Open-ended Training for Ad Hoc Teamwork
by: Wang, Caroline, et al.
Published: (2025)
by: Wang, Caroline, et al.
Published: (2025)
Rewarding Beliefs, Not Actions: Consistency-Guided Credit Assignment for Long-Horizon Agents
by: Tang, Wenjie, et al.
Published: (2026)
by: Tang, Wenjie, et al.
Published: (2026)
TML-Bench: Benchmark for Data Science Agents on Tabular ML Tasks
by: Pinchuk, Mykola
Published: (2026)
by: Pinchuk, Mykola
Published: (2026)
Safe and Policy-Compliant Multi-Agent Orchestration for Enterprise AI
by: Pasupuleti, Vinil, et al.
Published: (2026)
by: Pasupuleti, Vinil, et al.
Published: (2026)
ChromaFlow: A Negative Ablation Study of Orchestration Overhead in Tool-Augmented Agent Evaluation
by: Mittal, Tarun
Published: (2026)
by: Mittal, Tarun
Published: (2026)
Centrally Coordinated Multi-Agent Reinforcement Learning for Power Grid Topology Control
by: de Mol, Barbera, et al.
Published: (2025)
by: de Mol, Barbera, et al.
Published: (2025)
Extending NGU to Multi-Agent RL: A Preliminary Study
by: Hernandez, Juan, et al.
Published: (2025)
by: Hernandez, Juan, et al.
Published: (2025)
Advancing Multimodal Agent Reasoning with Long-Term Neuro-Symbolic Memory
by: Jiang, Rongjie, et al.
Published: (2026)
by: Jiang, Rongjie, et al.
Published: (2026)
Instruction-Level Weight Shaping: A Framework for Self-Improving AI Agents
by: Costa, Rimom
Published: (2025)
by: Costa, Rimom
Published: (2025)
StatePlane: A Cognitive State Plane for Long-Horizon AI Systems Under Bounded Context
by: Annapureddy, Sasank, et al.
Published: (2026)
by: Annapureddy, Sasank, et al.
Published: (2026)
Dynamic Dual-Granularity Skill Bank for Agentic RL
by: Tu, Songjun, et al.
Published: (2026)
by: Tu, Songjun, et al.
Published: (2026)
The Stochastic Gap: A Markovian Framework for Pre-Deployment Reliability and Oversight-Cost Auditing in Agentic Artificial Intelligence
by: Pal, Biplab, et al.
Published: (2026)
by: Pal, Biplab, et al.
Published: (2026)
QTypeMix: Enhancing Multi-Agent Cooperative Strategies through Heterogeneous and Homogeneous Value Decomposition
by: Fu, Songchen, et al.
Published: (2024)
by: Fu, Songchen, et al.
Published: (2024)
When Can Human-AI Teams Outperform Individuals? Tight Bounds with Impossibility Guarantees
by: Guo, Dongxin, et al.
Published: (2026)
by: Guo, Dongxin, et al.
Published: (2026)
PilotBench: A Benchmark for General Aviation Agents with Safety Constraints
by: Wu, Yalun, et al.
Published: (2026)
by: Wu, Yalun, et al.
Published: (2026)
Procedural Game Level Design with Deep Reinforcement Learning
by: Özkan, Miraç Buğra
Published: (2025)
by: Özkan, Miraç Buğra
Published: (2025)
When Agents Disagree: The Selection Bottleneck in Multi-Agent LLM Pipelines
by: Maryanskyy, Artem
Published: (2026)
by: Maryanskyy, Artem
Published: (2026)
Collaboration Promotes Group Resilience in Multi-Agent RL
by: Shraga, Ilai, et al.
Published: (2021)
by: Shraga, Ilai, et al.
Published: (2021)
FlowSteer: Towards Agents Designing Agentic Workflows via Reinforced Progressive Canvas Editing
by: Zhang, Mingda, et al.
Published: (2026)
by: Zhang, Mingda, et al.
Published: (2026)
AI-Assisted Engineering Should Track the Epistemic Status and Temporal Validity of Architectural Decisions
by: Gilda, Sankalp, et al.
Published: (2026)
by: Gilda, Sankalp, et al.
Published: (2026)
elsciRL: Integrating Language Solutions into Reinforcement Learning Problem Settings
by: Osborne, Philip, et al.
Published: (2025)
by: Osborne, Philip, et al.
Published: (2025)
A Framework for Assessing AI Agent Decisions and Outcomes in AutoML Pipelines
by: Du, Gaoyuan, et al.
Published: (2026)
by: Du, Gaoyuan, et al.
Published: (2026)
Umwelt Engineering: Designing the Cognitive Worlds of Linguistic Agents
by: Jehu-Appiah, Rodney
Published: (2026)
by: Jehu-Appiah, Rodney
Published: (2026)
PillagerBench: Benchmarking LLM-Based Agents in Competitive Minecraft Team Environments
by: Schipper, Olivier, et al.
Published: (2025)
by: Schipper, Olivier, et al.
Published: (2025)
Multi-Agent Design Assistant for the Simulation of Inertial Fusion Energy
by: Shachar, Meir H., et al.
Published: (2025)
by: Shachar, Meir H., et al.
Published: (2025)
Council Mode: A Heterogeneous Multi-Agent Consensus Framework for Reducing LLM Hallucination and Bias
by: Wu, Shuai, et al.
Published: (2026)
by: Wu, Shuai, et al.
Published: (2026)
Learning Approximate Nash Equilibria in Cooperative Multi-Agent Reinforcement Learning via Mean-Field Subsampling
by: Anand, Emile, et al.
Published: (2026)
by: Anand, Emile, et al.
Published: (2026)
Risk-Sensitive Multi-Agent Reinforcement Learning in Network Aggregative Markov Games
by: Ghaemi, Hafez, et al.
Published: (2024)
by: Ghaemi, Hafez, et al.
Published: (2024)
Policy Search, Retrieval, and Composition via Task Similarity in Collaborative Agentic Systems
by: Nath, Saptarshi, et al.
Published: (2025)
by: Nath, Saptarshi, et al.
Published: (2025)
RL-LLM-DT: An Automatic Decision Tree Generation Method Based on RL Evaluation and LLM Enhancement
by: Lin, Junjie, et al.
Published: (2024)
by: Lin, Junjie, et al.
Published: (2024)
Bounded Autonomy for Enterprise AI: Typed Action Contracts and Consumer-Side Execution
by: Sohail, Sarmad, et al.
Published: (2026)
by: Sohail, Sarmad, et al.
Published: (2026)
When Actions Disappear: Adversarial Action Removal in Self-Play Reinforcement Learning
by: Kujur, Arahan
Published: (2026)
by: Kujur, Arahan
Published: (2026)
How Good is ChatGPT in Giving Adaptive Guidance Using Knowledge Graphs in E-Learning Environments?
by: Ocheja, Patrick, et al.
Published: (2024)
by: Ocheja, Patrick, et al.
Published: (2024)
From Idea to CAD: A Language Model-Driven Multi-Agent System for Collaborative Design
by: Ocker, Felix, et al.
Published: (2025)
by: Ocker, Felix, et al.
Published: (2025)
PRISM-Consult: A Panel-of-Experts Architecture for Clinician-Aligned Diagnosis
by: Levine, Lionel, et al.
Published: (2025)
by: Levine, Lionel, et al.
Published: (2025)
Bimanual Robot Manipulation via Multi-Agent In-Context Learning
by: Palma, Alessio, et al.
Published: (2026)
by: Palma, Alessio, et al.
Published: (2026)
Decentralized Aerial Manipulation of a Cable-Suspended Load using Multi-Agent Reinforcement Learning
by: Zeng, Jack, et al.
Published: (2025)
by: Zeng, Jack, et al.
Published: (2025)
Enhancing Heterogeneous Multi-Agent Cooperation in Decentralized MARL via GNN-driven Intrinsic Rewards
by: Monon, Jahir Sadik, et al.
Published: (2024)
by: Monon, Jahir Sadik, et al.
Published: (2024)
Similar Items
-
N-Agent Ad Hoc Teamwork
by: Wang, Caroline, et al.
Published: (2024) -
Robust and Diverse Multi-Agent Learning via Rational Policy Gradient
by: Lauffer, Niklas, et al.
Published: (2025) -
ROTATE: Regret-driven Open-ended Training for Ad Hoc Teamwork
by: Wang, Caroline, et al.
Published: (2025) -
Rewarding Beliefs, Not Actions: Consistency-Guided Credit Assignment for Long-Horizon Agents
by: Tang, Wenjie, et al.
Published: (2026) -
TML-Bench: Benchmark for Data Science Agents on Tabular ML Tasks
by: Pinchuk, Mykola
Published: (2026)