Adapting Rules of Official International Mahjong for Online Players
Fuente:
arXiv
Saved in:
| Main Authors: | Wang, Chucai, Li, Lingfeng, Lu, Yunlong, Li, Wenxin |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Mxplainer: Explain and Learn Insights by Imitating Mahjong Agents
by: Li, Lingfeng, et al.
Published: (2025)
by: Li, Lingfeng, et al.
Published: (2025)
Style-Preserving Policy Optimization for Game Agents
by: Li, Lingfeng, et al.
Published: (2025)
by: Li, Lingfeng, et al.
Published: (2025)
Beyond Autoregressive RTG: Conditioning via Injection Outside Sequential Modeling in Decision Transformer
by: Wang, Yongyi, et al.
Published: (2026)
by: Wang, Yongyi, et al.
Published: (2026)
BotzoneBench: Scalable LLM Evaluation via Graded AI Anchors
by: Li, Lingfeng, et al.
Published: (2026)
by: Li, Lingfeng, et al.
Published: (2026)
Mahjax: A GPU-Accelerated Mahjong Simulator for Reinforcement Learning in JAX
by: Nishimori, Soichiro, et al.
Published: (2026)
by: Nishimori, Soichiro, et al.
Published: (2026)
Decoupling Return-to-Go for Efficient Decision Transformer
by: Wang, Yongyi, et al.
Published: (2026)
by: Wang, Yongyi, et al.
Published: (2026)
Synthetic POMDPs to Challenge Memory-Augmented RL: Memory Demand Structure Modeling
by: Wang, Yongyi, et al.
Published: (2025)
by: Wang, Yongyi, et al.
Published: (2025)
Is Artificial Intelligence Reshaping the Landscape of the International Academic Community of Geosciences?
by: Li, Liang, et al.
Published: (2025)
by: Li, Liang, et al.
Published: (2025)
ShuttleEnv: An Interactive Data-Driven RL Environment for Badminton Strategy Modeling
by: Li, Ang, et al.
Published: (2026)
by: Li, Ang, et al.
Published: (2026)
Constructing Non-Markovian Decision Process via History Aggregator
by: Wang, Yongyi, et al.
Published: (2025)
by: Wang, Yongyi, et al.
Published: (2025)
From Multimodal Perception to Strategic Reasoning: A Survey on AI-Generated Game Commentary
by: Zheng, Qirui, et al.
Published: (2025)
by: Zheng, Qirui, et al.
Published: (2025)
Provably Efficient Algorithm for Best Scoring Rule Identification in Online Principal-Agent Information Acquisition
by: Wang, Zichen, et al.
Published: (2025)
by: Wang, Zichen, et al.
Published: (2025)
The Epistemic Planning Domain Definition Language: Official Guideline
by: Burigana, Alessandro, et al.
Published: (2026)
by: Burigana, Alessandro, et al.
Published: (2026)
AI's Euclid's Elements Moment: From Language Models to Computable Thought
by: Fang, Xinmin, et al.
Published: (2025)
by: Fang, Xinmin, et al.
Published: (2025)
Closer to Language than Steam: AI as the Cognitive Engine of a New Productivity Revolution
by: Fang, Xinmin, et al.
Published: (2025)
by: Fang, Xinmin, et al.
Published: (2025)
Who is a Better Player: LLM against LLM
by: Zhou, Yingjie, et al.
Published: (2025)
by: Zhou, Yingjie, et al.
Published: (2025)
RuleSafe-VL: Evaluating Rule-Conditioned Decision Reasoning in Vision-Language Content Moderation
by: Lu, Zhifeng, et al.
Published: (2026)
by: Lu, Zhifeng, et al.
Published: (2026)
Complementary Learning System Empowers Online Continual Learning of Vehicle Motion Forecasting in Smart Cities
by: Li, Zirui, et al.
Published: (2025)
by: Li, Zirui, et al.
Published: (2025)
Digital Twin Channel-Enabled Online Resource Allocation for 6G: Principle, Architecture and Application
by: Li, Tongjie, et al.
Published: (2025)
by: Li, Tongjie, et al.
Published: (2025)
MPD$^2$-Router: Mask-aware Multi-expert Prior-regularized Dual-head Deferral Router in Glaucoma Screening and Diagnosis
by: Zhan, Wenxin
Published: (2026)
by: Zhan, Wenxin
Published: (2026)
GRAPPA: Generalizing and Adapting Robot Policies via Online Agentic Guidance
by: Bucker, Arthur, et al.
Published: (2024)
by: Bucker, Arthur, et al.
Published: (2024)
AI Pangaea: Unifying Intelligence Islands for Adapting Myriad Tasks
by: Chang, Jianlong, et al.
Published: (2025)
by: Chang, Jianlong, et al.
Published: (2025)
TRON: Targeted Rule-Verifiable Online Environments for Visual Reasoning RL
by: Yang, Tianze, et al.
Published: (2026)
by: Yang, Tianze, et al.
Published: (2026)
Official-NV: An LLM-Generated News Video Dataset for Multimodal Fake News Detection
by: Wang, Yihao, et al.
Published: (2024)
by: Wang, Yihao, et al.
Published: (2024)
LsAc ∗‐MJ: A Low‐Resource Consumption Reinforcement Learning Model for Mahjong Game
by: Xiali Li, et al.
Published: (2024)
by: Xiali Li, et al.
Published: (2024)
ADRD: LLM-Driven Autonomous Driving Based on Rule-based Decision Systems
by: Zeng, Fanzhi, et al.
Published: (2025)
by: Zeng, Fanzhi, et al.
Published: (2025)
DIRCR: Dual-Inference Rule-Contrastive Reasoning for Solving RAVENs
by: Zhang, Jiachen, et al.
Published: (2026)
by: Zhang, Jiachen, et al.
Published: (2026)
A Study of Rule Omission in Raven's Progressive Matrices
by: Li, Binze
Published: (2025)
by: Li, Binze
Published: (2025)
Anchoring AI Capabilities in Market Valuations: The Capability Realization Rate Model and Valuation Misalignment Risk
by: Fang, Xinmin, et al.
Published: (2025)
by: Fang, Xinmin, et al.
Published: (2025)
Observer, Not Player: Simulating Theory of Mind in LLMs through Game Observation
by: Wang, Jerry, et al.
Published: (2025)
by: Wang, Jerry, et al.
Published: (2025)
Align-Then-stEer: Adapting the Vision-Language Action Models through Unified Latent Guidance
by: Zhang, Yang, et al.
Published: (2025)
by: Zhang, Yang, et al.
Published: (2025)
Toward Modeling Player-Specific Chess Behaviors
by: Sogliuzzo, Loris, et al.
Published: (2026)
by: Sogliuzzo, Loris, et al.
Published: (2026)
Global Policy-Space Response Oracles for Two-Player Zero-Sum Games
by: Zhang, Junyu, et al.
Published: (2026)
by: Zhang, Junyu, et al.
Published: (2026)
Adapting Rule Representation With Four-Parameter Beta Distribution for Learning Classifier Systems
by: Shiraishi, Hiroki, et al.
Published: (2025)
by: Shiraishi, Hiroki, et al.
Published: (2025)
RuleR: Improving LLM Controllability by Rule-based Data Recycling
by: Li, Ming, et al.
Published: (2024)
by: Li, Ming, et al.
Published: (2024)
Adapting the Interface, Not the Model: Runtime Harness Adaptation for Deterministic LLM Agents
by: Xu, Tianshi, et al.
Published: (2026)
by: Xu, Tianshi, et al.
Published: (2026)
Adapting LLM Agents with Universal Feedback in Communication
by: Wang, Kuan, et al.
Published: (2023)
by: Wang, Kuan, et al.
Published: (2023)
Traceable Cross-Source RAG for Chinese Tibetan Medicine Question Answering
by: Chen, Fengxian, et al.
Published: (2026)
by: Chen, Fengxian, et al.
Published: (2026)
AdaptStress: Online Adaptive Learning for Interpretable and Personalized Stress Prediction Using Multivariate and Sparse Physiological Signals
by: Wang, Xueyi, et al.
Published: (2026)
by: Wang, Xueyi, et al.
Published: (2026)
Adapting Like Humans: A Metacognitive Agent with Test-time Reasoning
by: Li, Yang, et al.
Published: (2025)
by: Li, Yang, et al.
Published: (2025)
Similar Items
-
Mxplainer: Explain and Learn Insights by Imitating Mahjong Agents
by: Li, Lingfeng, et al.
Published: (2025) -
Style-Preserving Policy Optimization for Game Agents
by: Li, Lingfeng, et al.
Published: (2025) -
Beyond Autoregressive RTG: Conditioning via Injection Outside Sequential Modeling in Decision Transformer
by: Wang, Yongyi, et al.
Published: (2026) -
BotzoneBench: Scalable LLM Evaluation via Graded AI Anchors
by: Li, Lingfeng, et al.
Published: (2026) -
Mahjax: A GPU-Accelerated Mahjong Simulator for Reinforcement Learning in JAX
by: Nishimori, Soichiro, et al.
Published: (2026)