Style-Preserving Policy Optimization for Game Agents
Fuente:
arXiv
Saved in:
| Main Authors: | Li, Lingfeng, Lu, Yunlong, Wang, Yongyi, Li, Wenxin |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Mxplainer: Explain and Learn Insights by Imitating Mahjong Agents
by: Li, Lingfeng, et al.
Published: (2025)
by: Li, Lingfeng, et al.
Published: (2025)
Adapting Rules of Official International Mahjong for Online Players
by: Wang, Chucai, et al.
Published: (2026)
by: Wang, Chucai, et al.
Published: (2026)
Constructing Non-Markovian Decision Process via History Aggregator
by: Wang, Yongyi, et al.
Published: (2025)
by: Wang, Yongyi, et al.
Published: (2025)
BotzoneBench: Scalable LLM Evaluation via Graded AI Anchors
by: Li, Lingfeng, et al.
Published: (2026)
by: Li, Lingfeng, et al.
Published: (2026)
Synthetic POMDPs to Challenge Memory-Augmented RL: Memory Demand Structure Modeling
by: Wang, Yongyi, et al.
Published: (2025)
by: Wang, Yongyi, et al.
Published: (2025)
Decoupling Return-to-Go for Efficient Decision Transformer
by: Wang, Yongyi, et al.
Published: (2026)
by: Wang, Yongyi, et al.
Published: (2026)
ShuttleEnv: An Interactive Data-Driven RL Environment for Badminton Strategy Modeling
by: Li, Ang, et al.
Published: (2026)
by: Li, Ang, et al.
Published: (2026)
Beyond Autoregressive RTG: Conditioning via Injection Outside Sequential Modeling in Decision Transformer
by: Wang, Yongyi, et al.
Published: (2026)
by: Wang, Yongyi, et al.
Published: (2026)
From Multimodal Perception to Strategic Reasoning: A Survey on AI-Generated Game Commentary
by: Zheng, Qirui, et al.
Published: (2025)
by: Zheng, Qirui, et al.
Published: (2025)
Credit-Budgeted ICPC-Style Coding: When Agents Must Pay for Every Decision
by: Zhou, Lingfeng, et al.
Published: (2026)
by: Zhou, Lingfeng, et al.
Published: (2026)
MemPO: Self-Memory Policy Optimization for Long-Horizon Agents
by: Li, Ruoran, et al.
Published: (2026)
by: Li, Ruoran, et al.
Published: (2026)
Multi-Agent Guided Policy Optimization
by: Li, Yueheng, et al.
Published: (2025)
by: Li, Yueheng, et al.
Published: (2025)
Learning Strategic Language Agents in the Werewolf Game with Iterative Latent Space Policy Optimization
by: Xu, Zelai, et al.
Published: (2025)
by: Xu, Zelai, et al.
Published: (2025)
Learnable Game-theoretic Policy Optimization for Data-centric Self-explanation Rationalization
by: Zhao, Yunxiao, et al.
Published: (2025)
by: Zhao, Yunxiao, et al.
Published: (2025)
SetPO: Set-Level Policy Optimization for Diversity-Preserving LLM Reasoning
by: Li, Chenyi, et al.
Published: (2026)
by: Li, Chenyi, et al.
Published: (2026)
Style-Preserving Lip Sync via Audio-Aware Style Reference
by: Zhong, Weizhi, et al.
Published: (2024)
by: Zhong, Weizhi, et al.
Published: (2024)
Youtu-Agent: Scaling Agent Productivity with Automated Generation and Hybrid Policy Optimization
by: Shi, Yuchen, et al.
Published: (2025)
by: Shi, Yuchen, et al.
Published: (2025)
Game Generation via Large Language Models
by: Hu, Chengpeng, et al.
Published: (2024)
by: Hu, Chengpeng, et al.
Published: (2024)
Game-TARS: Pretrained Foundation Models for Scalable Generalist Multimodal Game Agents
by: Wang, Zihao, et al.
Published: (2025)
by: Wang, Zihao, et al.
Published: (2025)
Pareto-guided Pipeline for Distilling Featherweight AI Agents in Mobile MOBA Games
by: Yang, Xionghui, et al.
Published: (2026)
by: Yang, Xionghui, et al.
Published: (2026)
Counterfactual Credit Policy Optimization for Multi-Agent Collaboration
by: Li, Zhongyi, et al.
Published: (2026)
by: Li, Zhongyi, et al.
Published: (2026)
Perceptual Similarity for Measuring Decision-Making Style and Policy Diversity in Games
by: Lin, Chiu-Chou, et al.
Published: (2024)
by: Lin, Chiu-Chou, et al.
Published: (2024)
Agent-Pro: Learning to Evolve via Policy-Level Reflection and Optimization
by: Zhang, Wenqi, et al.
Published: (2024)
by: Zhang, Wenqi, et al.
Published: (2024)
AgenticRec: End-to-End Tool-Integrated Policy Optimization for Ranking-Oriented Recommender Agents
by: Li, Tianyi, et al.
Published: (2026)
by: Li, Tianyi, et al.
Published: (2026)
Games for Artificial Intelligence Research: A Review and Perspectives
by: Hu, Chengpeng, et al.
Published: (2023)
by: Hu, Chengpeng, et al.
Published: (2023)
Consolidation via Policy Information Regularization in Deep RL for Multi-Agent Games
by: Malloy, Tailia, et al.
Published: (2020)
by: Malloy, Tailia, et al.
Published: (2020)
From Off-Policy to On-Policy: Enhancing GUI Agents via Bi-level Expert-to-Policy Assimilation
by: Wang, Zezhou, et al.
Published: (2026)
by: Wang, Zezhou, et al.
Published: (2026)
D2PPO: Diffusion Policy Policy Optimization with Dispersive Loss
by: Zou, Guowei, et al.
Published: (2025)
by: Zou, Guowei, et al.
Published: (2025)
LARP: Language-Agent Role Play for Open-World Games
by: Yan, Ming, et al.
Published: (2023)
by: Yan, Ming, et al.
Published: (2023)
Guided Policy Optimization under Partial Observability
by: Li, Yueheng, et al.
Published: (2025)
by: Li, Yueheng, et al.
Published: (2025)
Safe Equilibrium Policy Optimization for Strategic Agent Policies
by: Arumugam, Karthika, et al.
Published: (2026)
by: Arumugam, Karthika, et al.
Published: (2026)
Graph-Enhanced Policy Optimization in LLM Agent Training
by: Yuan, Jiazhen, et al.
Published: (2025)
by: Yuan, Jiazhen, et al.
Published: (2025)
Bayesian Optimization-based Search for Agent Control in Automated Game Testing
by: Celemin, Carlos
Published: (2025)
by: Celemin, Carlos
Published: (2025)
PaperScout: An Autonomous Agent for Academic Paper Search with Process-Aware Sequence-Level Policy Optimization
by: Pan, Tingyue, et al.
Published: (2026)
by: Pan, Tingyue, et al.
Published: (2026)
COMAP: Co-Evolving World Models and Agent Policies for LLM Agents
by: Liu, Youwei, et al.
Published: (2026)
by: Liu, Youwei, et al.
Published: (2026)
Reference-guided Policy Optimization for Molecular Optimization via LLM Reasoning
by: Li, Xuan, et al.
Published: (2026)
by: Li, Xuan, et al.
Published: (2026)
Imitation Learning of Correlated Policies in Stackelberg Games
by: Wang, Kuang-Da, et al.
Published: (2025)
by: Wang, Kuang-Da, et al.
Published: (2025)
Analyzing and Internalizing Complex Policy Documents for LLM Agents
by: Liu, Jiateng, et al.
Published: (2025)
by: Liu, Jiateng, et al.
Published: (2025)
AI's Euclid's Elements Moment: From Language Models to Computable Thought
by: Fang, Xinmin, et al.
Published: (2025)
by: Fang, Xinmin, et al.
Published: (2025)
Closer to Language than Steam: AI as the Cognitive Engine of a New Productivity Revolution
by: Fang, Xinmin, et al.
Published: (2025)
by: Fang, Xinmin, et al.
Published: (2025)
Similar Items
-
Mxplainer: Explain and Learn Insights by Imitating Mahjong Agents
by: Li, Lingfeng, et al.
Published: (2025) -
Adapting Rules of Official International Mahjong for Online Players
by: Wang, Chucai, et al.
Published: (2026) -
Constructing Non-Markovian Decision Process via History Aggregator
by: Wang, Yongyi, et al.
Published: (2025) -
BotzoneBench: Scalable LLM Evaluation via Graded AI Anchors
by: Li, Lingfeng, et al.
Published: (2026) -
Synthetic POMDPs to Challenge Memory-Augmented RL: Memory Demand Structure Modeling
by: Wang, Yongyi, et al.
Published: (2025)