Balancing the AI Strength of Roles in Self-Play Training with Regret Matching+
Fuente:
arXiv
Gespeichert in:
| 1. Verfasser: | Wang, Xiaoxi |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Playing to Vision Foundation Model's Strengths in Stereo Matching
von: Liu, Chuang-Wei, et al.
Veröffentlicht: (2024)
von: Liu, Chuang-Wei, et al.
Veröffentlicht: (2024)
Self-Improving AI Agents through Self-Play
von: Chojecki, Przemyslaw
Veröffentlicht: (2025)
von: Chojecki, Przemyslaw
Veröffentlicht: (2025)
Disentangling Intent from Role: Adversarial Self-Play for Persona-Invariant Safety Alignment
von: Li, Jiajia, et al.
Veröffentlicht: (2026)
von: Li, Jiajia, et al.
Veröffentlicht: (2026)
The Oscars of AI Theater: A Survey on Role-Playing with Language Models
von: Chen, Nuo, et al.
Veröffentlicht: (2024)
von: Chen, Nuo, et al.
Veröffentlicht: (2024)
TriPlay-RL: Tri-Role Self-Play Reinforcement Learning for LLM Safety Alignment
von: Tan, Zhewen, et al.
Veröffentlicht: (2026)
von: Tan, Zhewen, et al.
Veröffentlicht: (2026)
Identity-Driven Hierarchical Role-Playing Agents
von: Sun, Libo, et al.
Veröffentlicht: (2024)
von: Sun, Libo, et al.
Veröffentlicht: (2024)
Minibal: Balanced Game-Playing Without Opponent Modeling
von: Cohen-Solal, Quentin, et al.
Veröffentlicht: (2026)
von: Cohen-Solal, Quentin, et al.
Veröffentlicht: (2026)
RoleCDE:Benchmarking and Mitigating Role-Alignment Trade-offs in Role-Playing Agents
von: Lai, Huayi, et al.
Veröffentlicht: (2026)
von: Lai, Huayi, et al.
Veröffentlicht: (2026)
Language Self-Play For Data-Free Training
von: Kuba, Jakub Grudzien, et al.
Veröffentlicht: (2025)
von: Kuba, Jakub Grudzien, et al.
Veröffentlicht: (2025)
TRACER: Turn-level Regret Matching with Inner Reinforcement Credit for Cooperative Multi-LLM Reasoning
von: Li, Chusen, et al.
Veröffentlicht: (2026)
von: Li, Chusen, et al.
Veröffentlicht: (2026)
Can Large Language Models Play Games? A Case Study of A Self-Play Approach
von: Guo, Hongyi, et al.
Veröffentlicht: (2024)
von: Guo, Hongyi, et al.
Veröffentlicht: (2024)
Robust Deep Monte Carlo Counterfactual Regret Minimization: Addressing Theoretical Risks in Neural Fictitious Self-Play
von: Jaafari, Zakaria El
Veröffentlicht: (2025)
von: Jaafari, Zakaria El
Veröffentlicht: (2025)
Stay in Character, Stay Safe: Dual-Cycle Adversarial Self-Evolution for Safety Role-Playing Agents
von: Liao, Mingyang, et al.
Veröffentlicht: (2026)
von: Liao, Mingyang, et al.
Veröffentlicht: (2026)
CoSER: A Comprehensive Literary Dataset and Framework for Training and Evaluating LLM Role-Playing and Persona Simulation
von: Wang, Xintao, et al.
Veröffentlicht: (2025)
von: Wang, Xintao, et al.
Veröffentlicht: (2025)
LARP: Language-Agent Role Play for Open-World Games
von: Yan, Ming, et al.
Veröffentlicht: (2023)
von: Yan, Ming, et al.
Veröffentlicht: (2023)
No-Regret Strategy Solving in Imperfect-Information Games via Pre-Trained Embedding
von: Fu, Yanchang, et al.
Veröffentlicht: (2025)
von: Fu, Yanchang, et al.
Veröffentlicht: (2025)
MINDECHO: Role-Playing Language Agents for Key Opinion Leaders
von: Xu, Rui, et al.
Veröffentlicht: (2024)
von: Xu, Rui, et al.
Veröffentlicht: (2024)
Reward-Decomposed Reinforcement Learning for Immersive Video Role-Playing
von: Wang, Miao, et al.
Veröffentlicht: (2026)
von: Wang, Miao, et al.
Veröffentlicht: (2026)
Leveraging AI Predicted and Expert Revised Annotations in Interactive Segmentation: Continual Tuning or Full Training?
von: Zhang, Tiezheng, et al.
Veröffentlicht: (2024)
von: Zhang, Tiezheng, et al.
Veröffentlicht: (2024)
Multi-Agent Training for Pommerman: Curriculum Learning and Population-based Self-Play Approach
von: Huynh, Nhat-Minh, et al.
Veröffentlicht: (2024)
von: Huynh, Nhat-Minh, et al.
Veröffentlicht: (2024)
Open Role-Playing with Delta-Engines
von: Wu, Hongqiu, et al.
Veröffentlicht: (2024)
von: Wu, Hongqiu, et al.
Veröffentlicht: (2024)
Post-Training LLMs as Better Decision-Making Agents: A Regret-Minimization Approach
von: Park, Chanwoo, et al.
Veröffentlicht: (2025)
von: Park, Chanwoo, et al.
Veröffentlicht: (2025)
Two-Sided Time-Independent Regret for Matching Markets with Limited Interviews
von: Mirfakhar, Amirmahdi, et al.
Veröffentlicht: (2026)
von: Mirfakhar, Amirmahdi, et al.
Veröffentlicht: (2026)
High-Performance Self-Supervised Learning by Joint Training of Flow Matching
von: Ukita, Kosuke, et al.
Veröffentlicht: (2025)
von: Ukita, Kosuke, et al.
Veröffentlicht: (2025)
Improving LLM Reasoning through Interpretable Role-Playing Steering
von: Wang, Anyi, et al.
Veröffentlicht: (2025)
von: Wang, Anyi, et al.
Veröffentlicht: (2025)
Boosting the Power of Small Multimodal Reasoning Models to Match Larger Models with Self-Consistency Training
von: Tan, Cheng, et al.
Veröffentlicht: (2023)
von: Tan, Cheng, et al.
Veröffentlicht: (2023)
Can VLMs Play Action Role-Playing Games? Take Black Myth Wukong as a Study Case
von: Chen, Peng, et al.
Veröffentlicht: (2024)
von: Chen, Peng, et al.
Veröffentlicht: (2024)
MMRole: A Comprehensive Framework for Developing and Evaluating Multimodal Role-Playing Agents
von: Dai, Yanqi, et al.
Veröffentlicht: (2024)
von: Dai, Yanqi, et al.
Veröffentlicht: (2024)
Superhuman AI for Stratego Using Self-Play Reinforcement Learning and Test-Time Search
von: Sokota, Samuel, et al.
Veröffentlicht: (2025)
von: Sokota, Samuel, et al.
Veröffentlicht: (2025)
Static Vs. Agentic Game Master AI for Facilitating Solo Role-Playing Experiences
von: Jørgensen, Nicolai Hejlesen, et al.
Veröffentlicht: (2025)
von: Jørgensen, Nicolai Hejlesen, et al.
Veröffentlicht: (2025)
Toward Training Superintelligent Software Agents through Self-Play SWE-RL
von: Wei, Yuxiang, et al.
Veröffentlicht: (2025)
von: Wei, Yuxiang, et al.
Veröffentlicht: (2025)
KokoroChat: A Japanese Psychological Counseling Dialogue Dataset Collected via Role-Playing by Trained Counselors
von: Qi, Zhiyang, et al.
Veröffentlicht: (2025)
von: Qi, Zhiyang, et al.
Veröffentlicht: (2025)
Role-Playing Evaluation for Large Language Models
von: Boudouri, Yassine El, et al.
Veröffentlicht: (2025)
von: Boudouri, Yassine El, et al.
Veröffentlicht: (2025)
Explaining Arguments' Strength: Unveiling the Role of Attacks and Supports (Technical Report)
von: Yin, Xiang, et al.
Veröffentlicht: (2024)
von: Yin, Xiang, et al.
Veröffentlicht: (2024)
Emotional RAG: Enhancing Role-Playing Agents through Emotional Retrieval
von: Huang, Le, et al.
Veröffentlicht: (2024)
von: Huang, Le, et al.
Veröffentlicht: (2024)
Character is Destiny: Can Role-Playing Language Agents Make Persona-Driven Decisions?
von: Xu, Rui, et al.
Veröffentlicht: (2024)
von: Xu, Rui, et al.
Veröffentlicht: (2024)
SPARK: Self-Play with Asymmetric Reward from Knowledge Graphs
von: Park, Hyobin, et al.
Veröffentlicht: (2026)
von: Park, Hyobin, et al.
Veröffentlicht: (2026)
Seirênes: Adversarial Self-Play with Evolving Distractions for LLM Reasoning
von: Zhang, Chi, et al.
Veröffentlicht: (2026)
von: Zhang, Chi, et al.
Veröffentlicht: (2026)
Propose, Solve, Verify: Self-Play Through Formal Verification
von: Wilf, Alex, et al.
Veröffentlicht: (2025)
von: Wilf, Alex, et al.
Veröffentlicht: (2025)
NGM: A Plug-and-Play Training-Free Memory Module for LLMs
von: Qu, Yuwen, et al.
Veröffentlicht: (2026)
von: Qu, Yuwen, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Playing to Vision Foundation Model's Strengths in Stereo Matching
von: Liu, Chuang-Wei, et al.
Veröffentlicht: (2024) -
Self-Improving AI Agents through Self-Play
von: Chojecki, Przemyslaw
Veröffentlicht: (2025) -
Disentangling Intent from Role: Adversarial Self-Play for Persona-Invariant Safety Alignment
von: Li, Jiajia, et al.
Veröffentlicht: (2026) -
The Oscars of AI Theater: A Survey on Role-Playing with Language Models
von: Chen, Nuo, et al.
Veröffentlicht: (2024) -
TriPlay-RL: Tri-Role Self-Play Reinforcement Learning for LLM Safety Alignment
von: Tan, Zhewen, et al.
Veröffentlicht: (2026)