Oracle-Guided Soft Shielding for Safe Move Prediction in Chess
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Rajendran, Prajit T, Arnez, Fabio, Espinoza, Huascar, Delaborde, Agnes, Mraidha, Chokri |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Less Precise Can Be More Reliable: A Systematic Evaluation of Quantization's Impact on VLMs Beyond Accuracy
von: Bouguerra, Aymen, et al.
Veröffentlicht: (2025)
von: Bouguerra, Aymen, et al.
Veröffentlicht: (2025)
A Behavior-Based Knowledge Representation Improves Prediction of Players' Moves in Chess by 25%
von: Skidanov, Benny, et al.
Veröffentlicht: (2025)
von: Skidanov, Benny, et al.
Veröffentlicht: (2025)
Dynamic Model Predictive Shielding for Provably Safe Reinforcement Learning
von: Banerjee, Arko, et al.
Veröffentlicht: (2024)
von: Banerjee, Arko, et al.
Veröffentlicht: (2024)
The Map of Misbelief: Tracing Intrinsic and Extrinsic Hallucinations Through Attention Patterns
von: Hajji, Elyes, et al.
Veröffentlicht: (2025)
von: Hajji, Elyes, et al.
Veröffentlicht: (2025)
Probabilistic Shielding for Safe Reinforcement Learning
von: Court, Edwin Hamel-De le, et al.
Veröffentlicht: (2025)
von: Court, Edwin Hamel-De le, et al.
Veröffentlicht: (2025)
ChessQA: Evaluating Large Language Models for Chess Understanding
von: Wen, Qianfeng, et al.
Veröffentlicht: (2025)
von: Wen, Qianfeng, et al.
Veröffentlicht: (2025)
Transformers can do Bayesian Clustering
von: Bhaskaran, Prajit, et al.
Veröffentlicht: (2025)
von: Bhaskaran, Prajit, et al.
Veröffentlicht: (2025)
X-Node: Self-Explanation is All We Need
von: Sengupta, Prajit, et al.
Veröffentlicht: (2025)
von: Sengupta, Prajit, et al.
Veröffentlicht: (2025)
LLM Chemistry Estimation for Multi-LLM Recommendation
von: Sanchez, Huascar, et al.
Veröffentlicht: (2025)
von: Sanchez, Huascar, et al.
Veröffentlicht: (2025)
Robust Probabilistic Shielding for Safe Offline Reinforcement Learning
von: Galesloot, Maris F. L., et al.
Veröffentlicht: (2026)
von: Galesloot, Maris F. L., et al.
Veröffentlicht: (2026)
Generating Creative Chess Puzzles
von: Feng, Xidong, et al.
Veröffentlicht: (2025)
von: Feng, Xidong, et al.
Veröffentlicht: (2025)
ChessArena: A Chess Testbed for Evaluating Strategic Reasoning Capabilities of Large Language Models
von: Liu, Jincheng, et al.
Veröffentlicht: (2025)
von: Liu, Jincheng, et al.
Veröffentlicht: (2025)
Safe Reinforcement Learning in Black-Box Environments via Adaptive Shielding
von: Bethell, Daniel, et al.
Veröffentlicht: (2024)
von: Bethell, Daniel, et al.
Veröffentlicht: (2024)
Complete Chess Games Enable LLM Become A Chess Master
von: Zhang, Yinqi, et al.
Veröffentlicht: (2025)
von: Zhang, Yinqi, et al.
Veröffentlicht: (2025)
Human-aligned Chess with a Bit of Search
von: Zhang, Yiming, et al.
Veröffentlicht: (2024)
von: Zhang, Yiming, et al.
Veröffentlicht: (2024)
Enhancing Chess Reinforcement Learning with Graph Representation
von: Rigaux, Tomas, et al.
Veröffentlicht: (2024)
von: Rigaux, Tomas, et al.
Veröffentlicht: (2024)
Robust Shielding for Safe Reinforcement Learning
von: Court, Edwin Hamel-De le, et al.
Veröffentlicht: (2026)
von: Court, Edwin Hamel-De le, et al.
Veröffentlicht: (2026)
Adaptive Shielding for Safe Reinforcement Learning under Hidden-Parameter Dynamics Shifts
von: Kwon, Minjae, et al.
Veröffentlicht: (2025)
von: Kwon, Minjae, et al.
Veröffentlicht: (2025)
Superior Computer Chess with Model Predictive Control, Reinforcement Learning, and Rollout
von: Gundawar, Atharva, et al.
Veröffentlicht: (2024)
von: Gundawar, Atharva, et al.
Veröffentlicht: (2024)
Approximate Model-Based Shielding for Safe Reinforcement Learning
von: Goodall, Alexander W., et al.
Veröffentlicht: (2023)
von: Goodall, Alexander W., et al.
Veröffentlicht: (2023)
Towards Piece-by-Piece Explanations for Chess Positions with SHAP
von: Spinnato, Francesco
Veröffentlicht: (2025)
von: Spinnato, Francesco
Veröffentlicht: (2025)
Iterative Inference in a Chess-Playing Neural Network
von: Sandmann, Elias, et al.
Veröffentlicht: (2025)
von: Sandmann, Elias, et al.
Veröffentlicht: (2025)
Diversifying AI: Towards Creative Chess with AlphaZero
von: Zahavy, Tom, et al.
Veröffentlicht: (2023)
von: Zahavy, Tom, et al.
Veröffentlicht: (2023)
Mastering Chinese Chess AI (Xiangqi) Without Search
von: Chen, Yu, et al.
Veröffentlicht: (2024)
von: Chen, Yu, et al.
Veröffentlicht: (2024)
Safe Reinforcement Learning via Recovery-based Shielding with Gaussian Process Dynamics Models
von: Goodall, Alexander W., et al.
Veröffentlicht: (2026)
von: Goodall, Alexander W., et al.
Veröffentlicht: (2026)
Auto-Discovery-Bench: Diagnosing Structured State Tracking in Oracle-Guided Discovery
von: Chen, Tingting, et al.
Veröffentlicht: (2025)
von: Chen, Tingting, et al.
Veröffentlicht: (2025)
Mixture of Masters: Sparse Chess Language Models with Player Routing
von: Frisoni, Giacomo, et al.
Veröffentlicht: (2026)
von: Frisoni, Giacomo, et al.
Veröffentlicht: (2026)
Evaluating In Silico Creativity: An Expert Review of AI Chess Compositions
von: Veeriah, Vivek, et al.
Veröffentlicht: (2025)
von: Veeriah, Vivek, et al.
Veröffentlicht: (2025)
Implicit Search via Discrete Diffusion: A Study on Chess
von: Ye, Jiacheng, et al.
Veröffentlicht: (2025)
von: Ye, Jiacheng, et al.
Veröffentlicht: (2025)
Reinforcement Learning by Guided Safe Exploration
von: Yang, Qisong, et al.
Veröffentlicht: (2023)
von: Yang, Qisong, et al.
Veröffentlicht: (2023)
Amortized Planning with Large-Scale Transformers: A Case Study on Chess
von: Ruoss, Anian, et al.
Veröffentlicht: (2024)
von: Ruoss, Anian, et al.
Veröffentlicht: (2024)
Evidence of Learned Look-Ahead in a Chess-Playing Neural Network
von: Jenner, Erik, et al.
Veröffentlicht: (2024)
von: Jenner, Erik, et al.
Veröffentlicht: (2024)
Neural Network-based Information Set Weighting for Playing Reconnaissance Blind Chess
von: Bertram, Timo, et al.
Veröffentlicht: (2024)
von: Bertram, Timo, et al.
Veröffentlicht: (2024)
Multi-LLM Collaboration for Medication Recommendation
von: Sanchez, Huascar, et al.
Veröffentlicht: (2025)
von: Sanchez, Huascar, et al.
Veröffentlicht: (2025)
Satellite Connectivity Prediction for Fast-Moving Platforms
von: Yan, Chao, et al.
Veröffentlicht: (2025)
von: Yan, Chao, et al.
Veröffentlicht: (2025)
Predicting User Perception of Move Brilliance in Chess
von: Zaidi, Kamron, et al.
Veröffentlicht: (2024)
von: Zaidi, Kamron, et al.
Veröffentlicht: (2024)
How Reasoning Evolves from Post-Training Data: An Empirical Study Using Chess
von: Dionisopoulos, Lucas, et al.
Veröffentlicht: (2026)
von: Dionisopoulos, Lucas, et al.
Veröffentlicht: (2026)
Liquid Reasoning Transformers: A Sudoku-Based Prototype for Chess-Scale Algorithmic Tasks
von: Sahni, Shivansh, et al.
Veröffentlicht: (2025)
von: Sahni, Shivansh, et al.
Veröffentlicht: (2025)
Designing Skill-Compatible AI: Methodologies and Frameworks in Chess
von: Hamade, Karim, et al.
Veröffentlicht: (2024)
von: Hamade, Karim, et al.
Veröffentlicht: (2024)
VAM: Verbalized Action Masking for Controllable Exploration in RL Post-Training -- A Chess Case Study
von: Zhang, Zhicheng, et al.
Veröffentlicht: (2026)
von: Zhang, Zhicheng, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Less Precise Can Be More Reliable: A Systematic Evaluation of Quantization's Impact on VLMs Beyond Accuracy
von: Bouguerra, Aymen, et al.
Veröffentlicht: (2025) -
A Behavior-Based Knowledge Representation Improves Prediction of Players' Moves in Chess by 25%
von: Skidanov, Benny, et al.
Veröffentlicht: (2025) -
Dynamic Model Predictive Shielding for Provably Safe Reinforcement Learning
von: Banerjee, Arko, et al.
Veröffentlicht: (2024) -
The Map of Misbelief: Tracing Intrinsic and Extrinsic Hallucinations Through Attention Patterns
von: Hajji, Elyes, et al.
Veröffentlicht: (2025) -
Probabilistic Shielding for Safe Reinforcement Learning
von: Court, Edwin Hamel-De le, et al.
Veröffentlicht: (2025)