Mixture of Masters: Sparse Chess Language Models with Player Routing
Fuente:
arXiv
Salvato in:
| Autori principali: | Frisoni, Giacomo, Molfetta, Lorenzo, Freddi, Davide, Moro, Gianluca |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Graph-of-Mark: Promote Spatial Reasoning in Multimodal Language Models with Graph-Based Visual Prompting
di: Frisoni, Giacomo, et al.
Pubblicazione: (2026)
di: Frisoni, Giacomo, et al.
Pubblicazione: (2026)
FEAST: Retrieval-Augmented Multi-Hierarchical Food Classification for the FoodEx2 System
di: Molfetta, Lorenzo, et al.
Pubblicazione: (2026)
di: Molfetta, Lorenzo, et al.
Pubblicazione: (2026)
ChessQA: Evaluating Large Language Models for Chess Understanding
di: Wen, Qianfeng, et al.
Pubblicazione: (2025)
di: Wen, Qianfeng, et al.
Pubblicazione: (2025)
Neuro-Symbolic Artificial Intelligence: A Task-Directed Survey in the Black-Box Models Era
di: Delvecchio, Giovanni Pio, et al.
Pubblicazione: (2026)
di: Delvecchio, Giovanni Pio, et al.
Pubblicazione: (2026)
Complete Chess Games Enable LLM Become A Chess Master
di: Zhang, Yinqi, et al.
Pubblicazione: (2025)
di: Zhang, Yinqi, et al.
Pubblicazione: (2025)
Soft-to-Hard Routing in Sparse Mixture-of-Experts Models
di: Rastegar, Reza
Pubblicazione: (2026)
di: Rastegar, Reza
Pubblicazione: (2026)
A Behavior-Based Knowledge Representation Improves Prediction of Players' Moves in Chess by 25%
di: Skidanov, Benny, et al.
Pubblicazione: (2025)
di: Skidanov, Benny, et al.
Pubblicazione: (2025)
Mastering Chinese Chess AI (Xiangqi) Without Search
di: Chen, Yu, et al.
Pubblicazione: (2024)
di: Chen, Yu, et al.
Pubblicazione: (2024)
ChessArena: A Chess Testbed for Evaluating Strategic Reasoning Capabilities of Large Language Models
di: Liu, Jincheng, et al.
Pubblicazione: (2025)
di: Liu, Jincheng, et al.
Pubblicazione: (2025)
Sparse Models, Sparse Safety: Unsafe Routes in Mixture-of-Experts LLMs
di: Jiang, Yukun, et al.
Pubblicazione: (2026)
di: Jiang, Yukun, et al.
Pubblicazione: (2026)
Generating Creative Chess Puzzles
di: Feng, Xidong, et al.
Pubblicazione: (2025)
di: Feng, Xidong, et al.
Pubblicazione: (2025)
Eradicating Negative Transfer in Multi-Physics Foundation Models via Sparse Mixture-of-Experts Routing
di: Sharma, Ellwil, et al.
Pubblicazione: (2026)
di: Sharma, Ellwil, et al.
Pubblicazione: (2026)
RouteHijack: Routing-Aware Attack on Mixture-of-Experts LLMs
di: Xu, Zhiyuan, et al.
Pubblicazione: (2026)
di: Xu, Zhiyuan, et al.
Pubblicazione: (2026)
Can Large Language Models Develop Strategic Reasoning? Post-training Insights from Learning Chess
di: Hwang, Dongyoon, et al.
Pubblicazione: (2025)
di: Hwang, Dongyoon, et al.
Pubblicazione: (2025)
Towards Generalization-Oriented Models for Vehicle Routing Problems with Mixture-of-Experts
di: Miao, Changhao, et al.
Pubblicazione: (2026)
di: Miao, Changhao, et al.
Pubblicazione: (2026)
Routing Mamba: Scaling State Space Models with Mixture-of-Experts Projection
di: Zhan, Zheng, et al.
Pubblicazione: (2025)
di: Zhan, Zheng, et al.
Pubblicazione: (2025)
PERFT: Parameter-Efficient Routed Fine-Tuning for Mixture-of-Expert Model
di: Liu, Yilun, et al.
Pubblicazione: (2024)
di: Liu, Yilun, et al.
Pubblicazione: (2024)
Human-aligned Chess with a Bit of Search
di: Zhang, Yiming, et al.
Pubblicazione: (2024)
di: Zhang, Yiming, et al.
Pubblicazione: (2024)
Enhancing Chess Reinforcement Learning with Graph Representation
di: Rigaux, Tomas, et al.
Pubblicazione: (2024)
di: Rigaux, Tomas, et al.
Pubblicazione: (2024)
Mixture of Raytraced Experts
di: Perin, Andrea, et al.
Pubblicazione: (2025)
di: Perin, Andrea, et al.
Pubblicazione: (2025)
Not All Models Suit Expert Offloading: On Local Routing Consistency of Mixture-of-Expert Models
di: Liang, Jingcong, et al.
Pubblicazione: (2025)
di: Liang, Jingcong, et al.
Pubblicazione: (2025)
Towards Piece-by-Piece Explanations for Chess Positions with SHAP
di: Spinnato, Francesco
Pubblicazione: (2025)
di: Spinnato, Francesco
Pubblicazione: (2025)
Iterative Inference in a Chess-Playing Neural Network
di: Sandmann, Elias, et al.
Pubblicazione: (2025)
di: Sandmann, Elias, et al.
Pubblicazione: (2025)
Diversifying AI: Towards Creative Chess with AlphaZero
di: Zahavy, Tom, et al.
Pubblicazione: (2023)
di: Zahavy, Tom, et al.
Pubblicazione: (2023)
Neural Inhibition Improves Dynamic Routing and Mixture of Experts
di: Zou, Will Y., et al.
Pubblicazione: (2025)
di: Zou, Will Y., et al.
Pubblicazione: (2025)
Modeling Expert Interactions in Sparse Mixture of Experts via Graph Structures
di: Nguyen-Nhat, Minh-Khoi, et al.
Pubblicazione: (2025)
di: Nguyen-Nhat, Minh-Khoi, et al.
Pubblicazione: (2025)
Dense Training, Sparse Inference: Rethinking Training of Mixture-of-Experts Language Models
di: Pan, Bowen, et al.
Pubblicazione: (2024)
di: Pan, Bowen, et al.
Pubblicazione: (2024)
Routing-Free Mixture-of-Experts
di: Liu, Yilun, et al.
Pubblicazione: (2026)
di: Liu, Yilun, et al.
Pubblicazione: (2026)
Multilingual Routing in Mixture-of-Experts
di: Bandarkar, Lucas, et al.
Pubblicazione: (2025)
di: Bandarkar, Lucas, et al.
Pubblicazione: (2025)
Probing Semantic Routing in Large Mixture-of-Expert Models
di: Olson, Matthew Lyle, et al.
Pubblicazione: (2025)
di: Olson, Matthew Lyle, et al.
Pubblicazione: (2025)
To Generate or to Retrieve? On the Effectiveness of Artificial Contexts for Medical Open-Domain Question Answering
di: Frisoni, Giacomo, et al.
Pubblicazione: (2024)
di: Frisoni, Giacomo, et al.
Pubblicazione: (2024)
Oracle-Guided Soft Shielding for Safe Move Prediction in Chess
di: Rajendran, Prajit T, et al.
Pubblicazione: (2026)
di: Rajendran, Prajit T, et al.
Pubblicazione: (2026)
Evaluating In Silico Creativity: An Expert Review of AI Chess Compositions
di: Veeriah, Vivek, et al.
Pubblicazione: (2025)
di: Veeriah, Vivek, et al.
Pubblicazione: (2025)
Implicit Search via Discrete Diffusion: A Study on Chess
di: Ye, Jiacheng, et al.
Pubblicazione: (2025)
di: Ye, Jiacheng, et al.
Pubblicazione: (2025)
UniMaia: Steering Chess Policies with Language for Human-like Play
di: Siu, Sherman, et al.
Pubblicazione: (2026)
di: Siu, Sherman, et al.
Pubblicazione: (2026)
Bridging the Gap between Expert and Language Models: Concept-guided Chess Commentary Generation and Evaluation
di: Kim, Jaechang, et al.
Pubblicazione: (2024)
di: Kim, Jaechang, et al.
Pubblicazione: (2024)
ProbMoE: Differentiable Probabilistic Routing for Mixture-of-Experts
di: Zhao, Heng, et al.
Pubblicazione: (2026)
di: Zhao, Heng, et al.
Pubblicazione: (2026)
MoxE: Mixture of xLSTM Experts with Entropy-Aware Routing for Efficient Language Modeling
di: Thiombiano, Abdoul Majid O., et al.
Pubblicazione: (2025)
di: Thiombiano, Abdoul Majid O., et al.
Pubblicazione: (2025)
Task-Conditioned Routing Signatures in Sparse Mixture-of-Experts Transformers
di: Avinash, Mynampati Sri Ranganadha
Pubblicazione: (2026)
di: Avinash, Mynampati Sri Ranganadha
Pubblicazione: (2026)
Modeling Matches as Language: A Generative Transformer Approach for Counterfactual Player Valuation in Football
di: Hong, Miru, et al.
Pubblicazione: (2026)
di: Hong, Miru, et al.
Pubblicazione: (2026)
Documenti analoghi
-
Graph-of-Mark: Promote Spatial Reasoning in Multimodal Language Models with Graph-Based Visual Prompting
di: Frisoni, Giacomo, et al.
Pubblicazione: (2026) -
FEAST: Retrieval-Augmented Multi-Hierarchical Food Classification for the FoodEx2 System
di: Molfetta, Lorenzo, et al.
Pubblicazione: (2026) -
ChessQA: Evaluating Large Language Models for Chess Understanding
di: Wen, Qianfeng, et al.
Pubblicazione: (2025) -
Neuro-Symbolic Artificial Intelligence: A Task-Directed Survey in the Black-Box Models Era
di: Delvecchio, Giovanni Pio, et al.
Pubblicazione: (2026) -
Complete Chess Games Enable LLM Become A Chess Master
di: Zhang, Yinqi, et al.
Pubblicazione: (2025)