Mixture of Masters: Sparse Chess Language Models with Player Routing
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Frisoni, Giacomo, Molfetta, Lorenzo, Freddi, Davide, Moro, Gianluca |
|---|---|
| Format: | Preprint |
| Publié: |
2026
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Graph-of-Mark: Promote Spatial Reasoning in Multimodal Language Models with Graph-Based Visual Prompting
par: Frisoni, Giacomo, et autres
Publié: (2026)
par: Frisoni, Giacomo, et autres
Publié: (2026)
FEAST: Retrieval-Augmented Multi-Hierarchical Food Classification for the FoodEx2 System
par: Molfetta, Lorenzo, et autres
Publié: (2026)
par: Molfetta, Lorenzo, et autres
Publié: (2026)
ChessQA: Evaluating Large Language Models for Chess Understanding
par: Wen, Qianfeng, et autres
Publié: (2025)
par: Wen, Qianfeng, et autres
Publié: (2025)
Neuro-Symbolic Artificial Intelligence: A Task-Directed Survey in the Black-Box Models Era
par: Delvecchio, Giovanni Pio, et autres
Publié: (2026)
par: Delvecchio, Giovanni Pio, et autres
Publié: (2026)
Complete Chess Games Enable LLM Become A Chess Master
par: Zhang, Yinqi, et autres
Publié: (2025)
par: Zhang, Yinqi, et autres
Publié: (2025)
Soft-to-Hard Routing in Sparse Mixture-of-Experts Models
par: Rastegar, Reza
Publié: (2026)
par: Rastegar, Reza
Publié: (2026)
A Behavior-Based Knowledge Representation Improves Prediction of Players' Moves in Chess by 25%
par: Skidanov, Benny, et autres
Publié: (2025)
par: Skidanov, Benny, et autres
Publié: (2025)
Mastering Chinese Chess AI (Xiangqi) Without Search
par: Chen, Yu, et autres
Publié: (2024)
par: Chen, Yu, et autres
Publié: (2024)
ChessArena: A Chess Testbed for Evaluating Strategic Reasoning Capabilities of Large Language Models
par: Liu, Jincheng, et autres
Publié: (2025)
par: Liu, Jincheng, et autres
Publié: (2025)
Sparse Models, Sparse Safety: Unsafe Routes in Mixture-of-Experts LLMs
par: Jiang, Yukun, et autres
Publié: (2026)
par: Jiang, Yukun, et autres
Publié: (2026)
Generating Creative Chess Puzzles
par: Feng, Xidong, et autres
Publié: (2025)
par: Feng, Xidong, et autres
Publié: (2025)
Eradicating Negative Transfer in Multi-Physics Foundation Models via Sparse Mixture-of-Experts Routing
par: Sharma, Ellwil, et autres
Publié: (2026)
par: Sharma, Ellwil, et autres
Publié: (2026)
RouteHijack: Routing-Aware Attack on Mixture-of-Experts LLMs
par: Xu, Zhiyuan, et autres
Publié: (2026)
par: Xu, Zhiyuan, et autres
Publié: (2026)
Can Large Language Models Develop Strategic Reasoning? Post-training Insights from Learning Chess
par: Hwang, Dongyoon, et autres
Publié: (2025)
par: Hwang, Dongyoon, et autres
Publié: (2025)
Towards Generalization-Oriented Models for Vehicle Routing Problems with Mixture-of-Experts
par: Miao, Changhao, et autres
Publié: (2026)
par: Miao, Changhao, et autres
Publié: (2026)
Routing Mamba: Scaling State Space Models with Mixture-of-Experts Projection
par: Zhan, Zheng, et autres
Publié: (2025)
par: Zhan, Zheng, et autres
Publié: (2025)
PERFT: Parameter-Efficient Routed Fine-Tuning for Mixture-of-Expert Model
par: Liu, Yilun, et autres
Publié: (2024)
par: Liu, Yilun, et autres
Publié: (2024)
Human-aligned Chess with a Bit of Search
par: Zhang, Yiming, et autres
Publié: (2024)
par: Zhang, Yiming, et autres
Publié: (2024)
Enhancing Chess Reinforcement Learning with Graph Representation
par: Rigaux, Tomas, et autres
Publié: (2024)
par: Rigaux, Tomas, et autres
Publié: (2024)
Mixture of Raytraced Experts
par: Perin, Andrea, et autres
Publié: (2025)
par: Perin, Andrea, et autres
Publié: (2025)
Not All Models Suit Expert Offloading: On Local Routing Consistency of Mixture-of-Expert Models
par: Liang, Jingcong, et autres
Publié: (2025)
par: Liang, Jingcong, et autres
Publié: (2025)
Towards Piece-by-Piece Explanations for Chess Positions with SHAP
par: Spinnato, Francesco
Publié: (2025)
par: Spinnato, Francesco
Publié: (2025)
Iterative Inference in a Chess-Playing Neural Network
par: Sandmann, Elias, et autres
Publié: (2025)
par: Sandmann, Elias, et autres
Publié: (2025)
Diversifying AI: Towards Creative Chess with AlphaZero
par: Zahavy, Tom, et autres
Publié: (2023)
par: Zahavy, Tom, et autres
Publié: (2023)
Neural Inhibition Improves Dynamic Routing and Mixture of Experts
par: Zou, Will Y., et autres
Publié: (2025)
par: Zou, Will Y., et autres
Publié: (2025)
Modeling Expert Interactions in Sparse Mixture of Experts via Graph Structures
par: Nguyen-Nhat, Minh-Khoi, et autres
Publié: (2025)
par: Nguyen-Nhat, Minh-Khoi, et autres
Publié: (2025)
Dense Training, Sparse Inference: Rethinking Training of Mixture-of-Experts Language Models
par: Pan, Bowen, et autres
Publié: (2024)
par: Pan, Bowen, et autres
Publié: (2024)
Routing-Free Mixture-of-Experts
par: Liu, Yilun, et autres
Publié: (2026)
par: Liu, Yilun, et autres
Publié: (2026)
Multilingual Routing in Mixture-of-Experts
par: Bandarkar, Lucas, et autres
Publié: (2025)
par: Bandarkar, Lucas, et autres
Publié: (2025)
Probing Semantic Routing in Large Mixture-of-Expert Models
par: Olson, Matthew Lyle, et autres
Publié: (2025)
par: Olson, Matthew Lyle, et autres
Publié: (2025)
To Generate or to Retrieve? On the Effectiveness of Artificial Contexts for Medical Open-Domain Question Answering
par: Frisoni, Giacomo, et autres
Publié: (2024)
par: Frisoni, Giacomo, et autres
Publié: (2024)
Oracle-Guided Soft Shielding for Safe Move Prediction in Chess
par: Rajendran, Prajit T, et autres
Publié: (2026)
par: Rajendran, Prajit T, et autres
Publié: (2026)
Evaluating In Silico Creativity: An Expert Review of AI Chess Compositions
par: Veeriah, Vivek, et autres
Publié: (2025)
par: Veeriah, Vivek, et autres
Publié: (2025)
Implicit Search via Discrete Diffusion: A Study on Chess
par: Ye, Jiacheng, et autres
Publié: (2025)
par: Ye, Jiacheng, et autres
Publié: (2025)
UniMaia: Steering Chess Policies with Language for Human-like Play
par: Siu, Sherman, et autres
Publié: (2026)
par: Siu, Sherman, et autres
Publié: (2026)
Bridging the Gap between Expert and Language Models: Concept-guided Chess Commentary Generation and Evaluation
par: Kim, Jaechang, et autres
Publié: (2024)
par: Kim, Jaechang, et autres
Publié: (2024)
ProbMoE: Differentiable Probabilistic Routing for Mixture-of-Experts
par: Zhao, Heng, et autres
Publié: (2026)
par: Zhao, Heng, et autres
Publié: (2026)
MoxE: Mixture of xLSTM Experts with Entropy-Aware Routing for Efficient Language Modeling
par: Thiombiano, Abdoul Majid O., et autres
Publié: (2025)
par: Thiombiano, Abdoul Majid O., et autres
Publié: (2025)
Task-Conditioned Routing Signatures in Sparse Mixture-of-Experts Transformers
par: Avinash, Mynampati Sri Ranganadha
Publié: (2026)
par: Avinash, Mynampati Sri Ranganadha
Publié: (2026)
Modeling Matches as Language: A Generative Transformer Approach for Counterfactual Player Valuation in Football
par: Hong, Miru, et autres
Publié: (2026)
par: Hong, Miru, et autres
Publié: (2026)
Documents similaires
-
Graph-of-Mark: Promote Spatial Reasoning in Multimodal Language Models with Graph-Based Visual Prompting
par: Frisoni, Giacomo, et autres
Publié: (2026) -
FEAST: Retrieval-Augmented Multi-Hierarchical Food Classification for the FoodEx2 System
par: Molfetta, Lorenzo, et autres
Publié: (2026) -
ChessQA: Evaluating Large Language Models for Chess Understanding
par: Wen, Qianfeng, et autres
Publié: (2025) -
Neuro-Symbolic Artificial Intelligence: A Task-Directed Survey in the Black-Box Models Era
par: Delvecchio, Giovanni Pio, et autres
Publié: (2026) -
Complete Chess Games Enable LLM Become A Chess Master
par: Zhang, Yinqi, et autres
Publié: (2025)