Mastering NIM and Impartial Games with Weak Neural Networks: An AlphaZero-inspired Multi-Frame Approach
Fuente:
arXiv
Saved in:
| Main Author: | Riis, Søren |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Separation and Collapse of Equilibria Inequalities on AND-OR Trees without Shape Constraints
by: Ito, Fuki, et al.
Published: (2024)
by: Ito, Fuki, et al.
Published: (2024)
Structured Radial Basis Function Network: Modelling Diversity for Multiple Hypotheses Prediction
by: Dominguez, Alejandro Rodriguez, et al.
Published: (2023)
by: Dominguez, Alejandro Rodriguez, et al.
Published: (2023)
Mitigating Catastrophic Forgetting in Streaming Generative and Predictive Learning via Stateful Replay
by: Du, Wenzhang
Published: (2025)
by: Du, Wenzhang
Published: (2025)
STACHE: Local Black-Box Explanations for Reinforcement Learning Policies
by: Elashkin, Andrew, et al.
Published: (2025)
by: Elashkin, Andrew, et al.
Published: (2025)
The Optimizer Quotient and the Certification Trilemma
by: Simas, Tristan
Published: (2026)
by: Simas, Tristan
Published: (2026)
Yanasse: Finding New Proofs from Deep Vision's Analogies, Part 1
by: Linhares, Alexandre
Published: (2026)
by: Linhares, Alexandre
Published: (2026)
Semantic-Constrained Federated Aggregation: Convergence Theory and Privacy-Utility Bounds for Knowledge-Enhanced Distributed Learning
by: Arafat, Jahidul
Published: (2025)
by: Arafat, Jahidul
Published: (2025)
The Serial Scaling Hypothesis
by: Liu, Yuxi, et al.
Published: (2025)
by: Liu, Yuxi, et al.
Published: (2025)
Optimistic Feasible Search for Closed-Loop Fair Threshold Decision-Making
by: Du, Wenzhang
Published: (2025)
by: Du, Wenzhang
Published: (2025)
OpCode-Based Malware Classification Using Machine Learning and Deep Learning Techniques
by: Saini, Varij, et al.
Published: (2025)
by: Saini, Varij, et al.
Published: (2025)
Resource-constrained Amazons chess decision framework integrating large language models and graph attention
by: Qian, Tianhao, et al.
Published: (2026)
by: Qian, Tianhao, et al.
Published: (2026)
CellARC: Measuring Intelligence with Cellular Automata
by: Lžičař, Miroslav
Published: (2025)
by: Lžičař, Miroslav
Published: (2025)
The Current and Future Perspectives of Zinc Oxide Nanoparticles in the Treatment of Diabetes Mellitus
by: Yousaf, Iqra
Published: (2024)
by: Yousaf, Iqra
Published: (2024)
STRIDE: A Self-Reflective Agent Framework for Reliable Automatic Equation Discovery
by: Su, Jiarui, et al.
Published: (2026)
by: Su, Jiarui, et al.
Published: (2026)
Connectivity-Aware Representations for Constrained Motion Planning via Multi-Scale Contrastive Learning
by: Jeon, Suhyun, et al.
Published: (2026)
by: Jeon, Suhyun, et al.
Published: (2026)
An arithmetic method algorithm optimizing k-nearest neighbors compared to regression algorithms and evaluated on real world data sources
by: Anagnostopoulos, Theodoros, et al.
Published: (2026)
by: Anagnostopoulos, Theodoros, et al.
Published: (2026)
Tao's Equational Proof Challenge Accepted (Technical Report)
by: Kondylidou, Lydia, et al.
Published: (2026)
by: Kondylidou, Lydia, et al.
Published: (2026)
Proving Olympiad Algebraic Inequalities without Human Demonstrations
by: Wei, Chenrui, et al.
Published: (2024)
by: Wei, Chenrui, et al.
Published: (2024)
Normalisation and Initialisation Strategies for Graph Neural Networks in Blockchain Anomaly Detection
by: Duy, Dang Sy, et al.
Published: (2026)
by: Duy, Dang Sy, et al.
Published: (2026)
Intersymbolic AI: Interlinking Symbolic AI and Subsymbolic AI
by: Platzer, André
Published: (2024)
by: Platzer, André
Published: (2024)
LAWS: Learning from Actual Workloads Symbolically -- A Self-Certifying Parametrized Cache Architecture for Neural Inference, Robotics, and Edge Deployment
by: Magarshak, Gregory
Published: (2026)
by: Magarshak, Gregory
Published: (2026)
Memory-Augmented State Machine Prompting: A Novel LLM Agent Framework for Real-Time Strategy Games
by: Qi, Runnan, et al.
Published: (2025)
by: Qi, Runnan, et al.
Published: (2025)
Semantic Variational Bayes Based on Semantic Information G Theory for Solving Latent Variables
by: Lu, Chenguang
Published: (2024)
by: Lu, Chenguang
Published: (2024)
Prediction-space knowledge markets for communication-efficient federated learning on multimedia tasks
by: Du, Wenzhang
Published: (2025)
by: Du, Wenzhang
Published: (2025)
Active perception and disentangled representations allow continual, episodic zero and few-shot learning
by: Rawlinson, David, et al.
Published: (2026)
by: Rawlinson, David, et al.
Published: (2026)
Adaptive Negative Scheduling for Graph Contrastive Learning
by: Ali, Adnan, et al.
Published: (2026)
by: Ali, Adnan, et al.
Published: (2026)
The Impact of Data Characteristics on GNN Evaluation for Detecting Fake News
by: Karn, Isha, et al.
Published: (2025)
by: Karn, Isha, et al.
Published: (2025)
A Structural Threshold in Decision Capacity Governs Collapse in Self-Play Reinforcement Learning
by: Kujur, Arahan
Published: (2026)
by: Kujur, Arahan
Published: (2026)
Murphys Laws of AI Alignment: Why the Gap Always Wins
by: Gaikwad, Madhava
Published: (2025)
by: Gaikwad, Madhava
Published: (2025)
Graded Transformers
by: Shaska Sr, Tony
Published: (2025)
by: Shaska Sr, Tony
Published: (2025)
Can a Bayesian Oracle Prevent Harm from an Agent?
by: Bengio, Yoshua, et al.
Published: (2024)
by: Bengio, Yoshua, et al.
Published: (2024)
FANAL -- Financial Activity News Alerting Language Modeling Framework
by: Patel, Urjitkumar, et al.
Published: (2024)
by: Patel, Urjitkumar, et al.
Published: (2024)
Aligning LLMs for Multilingual Consistency in Enterprise Applications
by: Agarwal, Amit, et al.
Published: (2025)
by: Agarwal, Amit, et al.
Published: (2025)
Tricks and Plug-ins for Gradient Boosting with Transformers
by: Fang, Biyi, et al.
Published: (2025)
by: Fang, Biyi, et al.
Published: (2025)
Formal Conjectures: An Open and Evolving Benchmark for Verified Discovery in Mathematics
by: Firsching, Moritz, et al.
Published: (2026)
by: Firsching, Moritz, et al.
Published: (2026)
On the Fundamental Limitations of Decentralized Learnable Reward Shaping in Cooperative Multi-Agent Reinforcement Learning
by: Akella, Aditya
Published: (2025)
by: Akella, Aditya
Published: (2025)
A Semantic Generalization of Shannon's Information Theory and Applications
by: Lu, Chenguang
Published: (2025)
by: Lu, Chenguang
Published: (2025)
When Are Two RLHF Objectives the Same?
by: Gaikwad, Madhava
Published: (2025)
by: Gaikwad, Madhava
Published: (2025)
Generating Realistic Safety-Critical Scenarios for Vehicle-Pedestrian Interactions
by: Pu, Qingwen, et al.
Published: (2026)
by: Pu, Qingwen, et al.
Published: (2026)
Quantum Deep Learning Still Needs a Quantum Leap
by: Gundlach, Hans, et al.
Published: (2025)
by: Gundlach, Hans, et al.
Published: (2025)
Similar Items
-
Separation and Collapse of Equilibria Inequalities on AND-OR Trees without Shape Constraints
by: Ito, Fuki, et al.
Published: (2024) -
Structured Radial Basis Function Network: Modelling Diversity for Multiple Hypotheses Prediction
by: Dominguez, Alejandro Rodriguez, et al.
Published: (2023) -
Mitigating Catastrophic Forgetting in Streaming Generative and Predictive Learning via Stateful Replay
by: Du, Wenzhang
Published: (2025) -
STACHE: Local Black-Box Explanations for Reinforcement Learning Policies
by: Elashkin, Andrew, et al.
Published: (2025) -
The Optimizer Quotient and the Certification Trilemma
by: Simas, Tristan
Published: (2026)