Mxplainer: Explain and Learn Insights by Imitating Mahjong Agents
Fuente:
arXiv
Salvato in:
| Autori principali: | Li, Lingfeng, Lu, Yunlong, Wang, Yongyi, Zheng, Qifan, Li, Wenxin |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Adapting Rules of Official International Mahjong for Online Players
di: Wang, Chucai, et al.
Pubblicazione: (2026)
di: Wang, Chucai, et al.
Pubblicazione: (2026)
Style-Preserving Policy Optimization for Game Agents
di: Li, Lingfeng, et al.
Pubblicazione: (2025)
di: Li, Lingfeng, et al.
Pubblicazione: (2025)
Constructing Non-Markovian Decision Process via History Aggregator
di: Wang, Yongyi, et al.
Pubblicazione: (2025)
di: Wang, Yongyi, et al.
Pubblicazione: (2025)
BotzoneBench: Scalable LLM Evaluation via Graded AI Anchors
di: Li, Lingfeng, et al.
Pubblicazione: (2026)
di: Li, Lingfeng, et al.
Pubblicazione: (2026)
Synthetic POMDPs to Challenge Memory-Augmented RL: Memory Demand Structure Modeling
di: Wang, Yongyi, et al.
Pubblicazione: (2025)
di: Wang, Yongyi, et al.
Pubblicazione: (2025)
Decoupling Return-to-Go for Efficient Decision Transformer
di: Wang, Yongyi, et al.
Pubblicazione: (2026)
di: Wang, Yongyi, et al.
Pubblicazione: (2026)
Beyond Autoregressive RTG: Conditioning via Injection Outside Sequential Modeling in Decision Transformer
di: Wang, Yongyi, et al.
Pubblicazione: (2026)
di: Wang, Yongyi, et al.
Pubblicazione: (2026)
ShuttleEnv: An Interactive Data-Driven RL Environment for Badminton Strategy Modeling
di: Li, Ang, et al.
Pubblicazione: (2026)
di: Li, Ang, et al.
Pubblicazione: (2026)
Mahjax: A GPU-Accelerated Mahjong Simulator for Reinforcement Learning in JAX
di: Nishimori, Soichiro, et al.
Pubblicazione: (2026)
di: Nishimori, Soichiro, et al.
Pubblicazione: (2026)
Explaining Decentralized Multi-Agent Reinforcement Learning Policies
di: Boggess, Kayla, et al.
Pubblicazione: (2025)
di: Boggess, Kayla, et al.
Pubblicazione: (2025)
From Multimodal Perception to Strategic Reasoning: A Survey on AI-Generated Game Commentary
di: Zheng, Qirui, et al.
Pubblicazione: (2025)
di: Zheng, Qirui, et al.
Pubblicazione: (2025)
Learning Strategy Representation for Imitation Learning in Multi-Agent Games
di: Lei, Shiqi, et al.
Pubblicazione: (2024)
di: Lei, Shiqi, et al.
Pubblicazione: (2024)
KOI: Accelerating Online Imitation Learning via Hybrid Key-state Guidance
di: Lu, Jingxian, et al.
Pubblicazione: (2024)
di: Lu, Jingxian, et al.
Pubblicazione: (2024)
InterReal: A Unified Physics-Based Imitation Framework for Learning Human-Object Interaction Skills
di: Liang, Dayang, et al.
Pubblicazione: (2026)
di: Liang, Dayang, et al.
Pubblicazione: (2026)
Beyond Imitation: Reinforcement Learning for Active Latent Planning
di: Zheng, Zhi, et al.
Pubblicazione: (2026)
di: Zheng, Zhi, et al.
Pubblicazione: (2026)
Imitation Learning for Autonomous Driving: Insights from Real-World Testing
di: Dursun, Hidayet Ersin, et al.
Pubblicazione: (2025)
di: Dursun, Hidayet Ersin, et al.
Pubblicazione: (2025)
Learning to Adapt: Self-Improving Web Agent via Cognitive-Aware Exploration
di: Chen, Weile, et al.
Pubblicazione: (2026)
di: Chen, Weile, et al.
Pubblicazione: (2026)
Towards Autonomous Business Intelligence via Data-to-Insight Discovery Agent
di: Wu, Dongming, et al.
Pubblicazione: (2026)
di: Wu, Dongming, et al.
Pubblicazione: (2026)
Scaling Laws for Imitation Learning in Single-Agent Games
di: Tuyls, Jens, et al.
Pubblicazione: (2023)
di: Tuyls, Jens, et al.
Pubblicazione: (2023)
TriFlow: A Progressive Multi-Agent Framework for Intelligent Trip Planning
di: Chen, Yuxing, et al.
Pubblicazione: (2025)
di: Chen, Yuxing, et al.
Pubblicazione: (2025)
Learning Representations in Video Game Agents with Supervised Contrastive Imitation Learning
di: Celemin, Carlos, et al.
Pubblicazione: (2025)
di: Celemin, Carlos, et al.
Pubblicazione: (2025)
Reinforced Imitative Trajectory Planning for Urban Automated Driving
di: Zeng, Di, et al.
Pubblicazione: (2024)
di: Zeng, Di, et al.
Pubblicazione: (2024)
Avenir-Web: Human-Experience-Imitating Multimodal Web Agents with Mixture of Grounding Experts
di: Li, Aiden Yiliu, et al.
Pubblicazione: (2026)
di: Li, Aiden Yiliu, et al.
Pubblicazione: (2026)
SEABO: A Simple Search-Based Method for Offline Imitation Learning
di: Lyu, Jiafei, et al.
Pubblicazione: (2024)
di: Lyu, Jiafei, et al.
Pubblicazione: (2024)
LsAc ∗‐MJ: A Low‐Resource Consumption Reinforcement Learning Model for Mahjong Game
di: Xiali Li, et al.
Pubblicazione: (2024)
di: Xiali Li, et al.
Pubblicazione: (2024)
Credit-Budgeted ICPC-Style Coding: When Agents Must Pay for Every Decision
di: Zhou, Lingfeng, et al.
Pubblicazione: (2026)
di: Zhou, Lingfeng, et al.
Pubblicazione: (2026)
Imitation Game: A Model-based and Imitation Learning Deep Reinforcement Learning Hybrid
di: Veith, Eric MSP, et al.
Pubblicazione: (2024)
di: Veith, Eric MSP, et al.
Pubblicazione: (2024)
SKIL: Semantic Keypoint Imitation Learning for Generalizable Data-efficient Manipulation
di: Wang, Shengjie, et al.
Pubblicazione: (2025)
di: Wang, Shengjie, et al.
Pubblicazione: (2025)
Explaining Decisions of Agents in Mixed-Motive Games
di: Orner, Maayan, et al.
Pubblicazione: (2024)
di: Orner, Maayan, et al.
Pubblicazione: (2024)
Imitating Mistakes in a Learning Companion AI Agent for Online Peer Learning
di: Moribe, Sosui, et al.
Pubblicazione: (2025)
di: Moribe, Sosui, et al.
Pubblicazione: (2025)
Imitation Learning of Correlated Policies in Stackelberg Games
di: Wang, Kuang-Da, et al.
Pubblicazione: (2025)
di: Wang, Kuang-Da, et al.
Pubblicazione: (2025)
Explaining the Contributing Factors for Vulnerability Detection in Machine Learning
di: Mouine, Esma, et al.
Pubblicazione: (2024)
di: Mouine, Esma, et al.
Pubblicazione: (2024)
Generalization Capability for Imitation Learning
di: Wang, Yixiao
Pubblicazione: (2025)
di: Wang, Yixiao
Pubblicazione: (2025)
Deploying Ten Thousand Robots: Scalable Imitation Learning for Lifelong Multi-Agent Path Finding
di: Jiang, He, et al.
Pubblicazione: (2024)
di: Jiang, He, et al.
Pubblicazione: (2024)
A Multi-Task Role-Playing Agent Capable of Imitating Character Linguistic Styles
di: Chen, Siyuan, et al.
Pubblicazione: (2024)
di: Chen, Siyuan, et al.
Pubblicazione: (2024)
R2BC: Multi-Agent Imitation Learning from Single-Agent Demonstrations
di: Mattson, Connor, et al.
Pubblicazione: (2025)
di: Mattson, Connor, et al.
Pubblicazione: (2025)
HannesImitation: Grasping with the Hannes Prosthetic Hand via Imitation Learning
di: Alessi, Carlo, et al.
Pubblicazione: (2025)
di: Alessi, Carlo, et al.
Pubblicazione: (2025)
Self-evolved Imitation Learning in Simulated World
di: Ye, Yifan, et al.
Pubblicazione: (2025)
di: Ye, Yifan, et al.
Pubblicazione: (2025)
AI's Euclid's Elements Moment: From Language Models to Computable Thought
di: Fang, Xinmin, et al.
Pubblicazione: (2025)
di: Fang, Xinmin, et al.
Pubblicazione: (2025)
Closer to Language than Steam: AI as the Cognitive Engine of a New Productivity Revolution
di: Fang, Xinmin, et al.
Pubblicazione: (2025)
di: Fang, Xinmin, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Adapting Rules of Official International Mahjong for Online Players
di: Wang, Chucai, et al.
Pubblicazione: (2026) -
Style-Preserving Policy Optimization for Game Agents
di: Li, Lingfeng, et al.
Pubblicazione: (2025) -
Constructing Non-Markovian Decision Process via History Aggregator
di: Wang, Yongyi, et al.
Pubblicazione: (2025) -
BotzoneBench: Scalable LLM Evaluation via Graded AI Anchors
di: Li, Lingfeng, et al.
Pubblicazione: (2026) -
Synthetic POMDPs to Challenge Memory-Augmented RL: Memory Demand Structure Modeling
di: Wang, Yongyi, et al.
Pubblicazione: (2025)