Checkmating One, by Using Many: Combining Mixture of Experts with MCTS to Improve in Chess
Fuente:
arXiv
Salvato in:
| Autori principali: | Helfenstein, Felix, Czech, Johannes, Blüml, Jannis, Eisel, Max, Kersting, Kristian |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Representation Matters for Mastering Chess: Improved Feature Representation in AlphaZero Outperforms Switching to Transformers
di: Czech, Johannes, et al.
Pubblicazione: (2023)
di: Czech, Johannes, et al.
Pubblicazione: (2023)
HackAtari: Atari Learning Environments for Robust and Continual Reinforcement Learning
di: Delfosse, Quentin, et al.
Pubblicazione: (2024)
di: Delfosse, Quentin, et al.
Pubblicazione: (2024)
Polynomial Regret Concentration of UCB for Non-Deterministic State Transitions
di: Cömer, Can, et al.
Pubblicazione: (2025)
di: Cömer, Can, et al.
Pubblicazione: (2025)
OCAtari: Object-Centric Atari 2600 Reinforcement Learning Environments
di: Delfosse, Quentin, et al.
Pubblicazione: (2023)
di: Delfosse, Quentin, et al.
Pubblicazione: (2023)
OCALM: Object-Centric Assessment with Language Models
di: Kaufmann, Timo, et al.
Pubblicazione: (2024)
di: Kaufmann, Timo, et al.
Pubblicazione: (2024)
Deep Reinforcement Learning via Object-Centric Attention
di: Blüml, Jannis, et al.
Pubblicazione: (2025)
di: Blüml, Jannis, et al.
Pubblicazione: (2025)
Better Decisions through the Right Causal World Model
di: Dillies, Elisabeth, et al.
Pubblicazione: (2025)
di: Dillies, Elisabeth, et al.
Pubblicazione: (2025)
Boosting deep Reinforcement Learning using pretraining with Logical Options
di: Ye, Zihan, et al.
Pubblicazione: (2026)
di: Ye, Zihan, et al.
Pubblicazione: (2026)
Kintsugi: Learning Policies by Repairing Executable Knowledge Bases
di: Cao, Teng, et al.
Pubblicazione: (2026)
di: Cao, Teng, et al.
Pubblicazione: (2026)
Deep Reinforcement Learning Agents are not even close to Human Intelligence
di: Delfosse, Quentin, et al.
Pubblicazione: (2025)
di: Delfosse, Quentin, et al.
Pubblicazione: (2025)
Amplifying Exploration in Monte-Carlo Tree Search by Focusing on the Unknown
di: Derstroff, Cedric, et al.
Pubblicazione: (2024)
di: Derstroff, Cedric, et al.
Pubblicazione: (2024)
How Many Experts Are Enough? Towards Optimal Semantic Specialization for Mixture-of-Experts
di: Park, Sumin, et al.
Pubblicazione: (2025)
di: Park, Sumin, et al.
Pubblicazione: (2025)
Learning to Intervene on Concept Bottlenecks
di: Steinmann, David, et al.
Pubblicazione: (2023)
di: Steinmann, David, et al.
Pubblicazione: (2023)
Adaptable Hindsight Experience Replay for Search-Based Learning
di: Vazaios, Alexandros, et al.
Pubblicazione: (2025)
di: Vazaios, Alexandros, et al.
Pubblicazione: (2025)
Mixture of Masters: Sparse Chess Language Models with Player Routing
di: Frisoni, Giacomo, et al.
Pubblicazione: (2026)
di: Frisoni, Giacomo, et al.
Pubblicazione: (2026)
Depth-Recurrent Attention Mixtures: Giving Latent Reasoning the Attention it Deserves
di: Knupp, Jonas, et al.
Pubblicazione: (2026)
di: Knupp, Jonas, et al.
Pubblicazione: (2026)
Evaluating In Silico Creativity: An Expert Review of AI Chess Compositions
di: Veeriah, Vivek, et al.
Pubblicazione: (2025)
di: Veeriah, Vivek, et al.
Pubblicazione: (2025)
Mixture of Many Zero-Compute Experts: A High-Rate Quantization Theory Perspective
di: Dar, Yehuda
Pubblicazione: (2025)
di: Dar, Yehuda
Pubblicazione: (2025)
A Typology for Exploring the Mitigation of Shortcut Behavior
di: Friedrich, Felix, et al.
Pubblicazione: (2022)
di: Friedrich, Felix, et al.
Pubblicazione: (2022)
United We Pretrain, Divided We Fail! Representation Learning for Time Series by Pretraining on 75 Datasets at Once
di: Kraus, Maurice, et al.
Pubblicazione: (2024)
di: Kraus, Maurice, et al.
Pubblicazione: (2024)
xLSTM-Mixer: Multivariate Time Series Forecasting by Mixing via Scalar Memories
di: Kraus, Maurice, et al.
Pubblicazione: (2024)
di: Kraus, Maurice, et al.
Pubblicazione: (2024)
Exploring Neural Granger Causality with xLSTMs: Unveiling Temporal Dependencies in Complex Data
di: Poonia, Harsh, et al.
Pubblicazione: (2025)
di: Poonia, Harsh, et al.
Pubblicazione: (2025)
Selective Sinkhorn Routing for Improved Sparse Mixture of Experts
di: Nguyen, Duc Anh, et al.
Pubblicazione: (2025)
di: Nguyen, Duc Anh, et al.
Pubblicazione: (2025)
Improving Routing in Sparse Mixture of Experts with Graph of Tokens
di: Nguyen, Tam, et al.
Pubblicazione: (2025)
di: Nguyen, Tam, et al.
Pubblicazione: (2025)
Deep Classifier Mimicry without Data Access
di: Braun, Steven, et al.
Pubblicazione: (2023)
di: Braun, Steven, et al.
Pubblicazione: (2023)
SEUF: Is Unlearning One Expert Enough for Mixture-of-Experts LLMs?
di: Zhuang, Haomin, et al.
Pubblicazione: (2024)
di: Zhuang, Haomin, et al.
Pubblicazione: (2024)
Efficient Mixture-of-Experts LLM Inference with Apple Silicon NPUs
di: Benazir, Afsara, et al.
Pubblicazione: (2026)
di: Benazir, Afsara, et al.
Pubblicazione: (2026)
Let the Experts Speak: Improving Survival Prediction & Calibration via Mixture-of-Experts Heads
di: Morrill, Todd, et al.
Pubblicazione: (2025)
di: Morrill, Todd, et al.
Pubblicazione: (2025)
Routing Manifold Alignment Improves Generalization of Mixture-of-Experts LLMs
di: Li, Zhongyang, et al.
Pubblicazione: (2025)
di: Li, Zhongyang, et al.
Pubblicazione: (2025)
AtManRL: Towards Faithful Reasoning via Differentiable Attention Saliency
di: Höth, Max Henning, et al.
Pubblicazione: (2026)
di: Höth, Max Henning, et al.
Pubblicazione: (2026)
No Safe Dose: How Training Data Drives Unsafe Image Generation
di: Friedrich, Felix, et al.
Pubblicazione: (2026)
di: Friedrich, Felix, et al.
Pubblicazione: (2026)
Dense Backpropagation Improves Training for Sparse Mixture-of-Experts
di: Panda, Ashwinee, et al.
Pubblicazione: (2025)
di: Panda, Ashwinee, et al.
Pubblicazione: (2025)
Neural Inhibition Improves Dynamic Routing and Mixture of Experts
di: Zou, Will Y., et al.
Pubblicazione: (2025)
di: Zou, Will Y., et al.
Pubblicazione: (2025)
Mixture of Concept Bottleneck Experts
di: De Santis, Francesco, et al.
Pubblicazione: (2026)
di: De Santis, Francesco, et al.
Pubblicazione: (2026)
ChessQA: Evaluating Large Language Models for Chess Understanding
di: Wen, Qianfeng, et al.
Pubblicazione: (2025)
di: Wen, Qianfeng, et al.
Pubblicazione: (2025)
Mixture of Latent Experts Using Tensor Products
di: Su, Zhan, et al.
Pubblicazione: (2024)
di: Su, Zhan, et al.
Pubblicazione: (2024)
One-Prompt Strikes Back: Sparse Mixture of Experts for Prompt-based Continual Learning
di: Le, Minh, et al.
Pubblicazione: (2025)
di: Le, Minh, et al.
Pubblicazione: (2025)
Navigating Shortcuts, Spurious Correlations, and Confounders: From Origins via Detection to Mitigation
di: Steinmann, David, et al.
Pubblicazione: (2024)
di: Steinmann, David, et al.
Pubblicazione: (2024)
Learning by Self-Explaining
di: Stammer, Wolfgang, et al.
Pubblicazione: (2023)
di: Stammer, Wolfgang, et al.
Pubblicazione: (2023)
Learning More Generalized Experts by Merging Experts in Mixture-of-Experts
di: Park, Sejik
Pubblicazione: (2024)
di: Park, Sejik
Pubblicazione: (2024)
Documenti analoghi
-
Representation Matters for Mastering Chess: Improved Feature Representation in AlphaZero Outperforms Switching to Transformers
di: Czech, Johannes, et al.
Pubblicazione: (2023) -
HackAtari: Atari Learning Environments for Robust and Continual Reinforcement Learning
di: Delfosse, Quentin, et al.
Pubblicazione: (2024) -
Polynomial Regret Concentration of UCB for Non-Deterministic State Transitions
di: Cömer, Can, et al.
Pubblicazione: (2025) -
OCAtari: Object-Centric Atari 2600 Reinforcement Learning Environments
di: Delfosse, Quentin, et al.
Pubblicazione: (2023) -
OCALM: Object-Centric Assessment with Language Models
di: Kaufmann, Timo, et al.
Pubblicazione: (2024)