On some improvements to Unbounded Minimax
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Cohen-Solal, Quentin, Cazenave, Tristan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Minimax Strikes Back
von: Cohen-Solal, Quentin, et al.
Veröffentlicht: (2020)
von: Cohen-Solal, Quentin, et al.
Veröffentlicht: (2020)
Completeness of Unbounded Best-First Minimax and Descent Minimax
von: Cohen-Solal, Quentin
Veröffentlicht: (2026)
von: Cohen-Solal, Quentin
Veröffentlicht: (2026)
Minibal: Balanced Game-Playing Without Opponent Modeling
von: Cohen-Solal, Quentin, et al.
Veröffentlicht: (2026)
von: Cohen-Solal, Quentin, et al.
Veröffentlicht: (2026)
Study and improvement of search algorithms in two-players perfect information games
von: Cohen-Solal, Quentin
Veröffentlicht: (2025)
von: Cohen-Solal, Quentin
Veröffentlicht: (2025)
Learning to Play Two-Player Perfect-Information Games without Knowledge
von: Cohen-Solal, Quentin
Veröffentlicht: (2020)
von: Cohen-Solal, Quentin
Veröffentlicht: (2020)
Study and Improvement of Search Algorithms in Multi-Player Perfect-Information Games
von: Cohen-Solal, Quentin
Veröffentlicht: (2026)
von: Cohen-Solal, Quentin
Veröffentlicht: (2026)
Eterna is Solved
von: Cazenave, Tristan
Veröffentlicht: (2025)
von: Cazenave, Tristan
Veröffentlicht: (2025)
Monte Carlo Search Algorithms Discovering Monte Carlo Tree Search Exploration Terms
von: Cazenave, Tristan
Veröffentlicht: (2024)
von: Cazenave, Tristan
Veröffentlicht: (2024)
Generalized Nested Rollout Policy Adaptation with Limited Repetitions
von: Cazenave, Tristan
Veröffentlicht: (2024)
von: Cazenave, Tristan
Veröffentlicht: (2024)
Learning a Prior for Monte Carlo Search by Replaying Solutions to Combinatorial Problems
von: Cazenave, Tristan
Veröffentlicht: (2024)
von: Cazenave, Tristan
Veröffentlicht: (2024)
Monte Carlo Permutation Search
von: Cazenave, Tristan
Veröffentlicht: (2025)
von: Cazenave, Tristan
Veröffentlicht: (2025)
Deciding the Satisfiability of Combined Qualitative Constraint Networks
von: Cohen-Solal, Quentin, et al.
Veröffentlicht: (2026)
von: Cohen-Solal, Quentin, et al.
Veröffentlicht: (2026)
Monte Carlo Graph Coloring
von: Cazenave, Tristan, et al.
Veröffentlicht: (2025)
von: Cazenave, Tristan, et al.
Veröffentlicht: (2025)
Perfect Information Monte Carlo with Postponing Reasoning
von: Arjonilla, Jérôme, et al.
Veröffentlicht: (2024)
von: Arjonilla, Jérôme, et al.
Veröffentlicht: (2024)
Enhancing Reinforcement Learning Through Guided Search
von: Arjonilla, Jérôme, et al.
Veröffentlicht: (2024)
von: Arjonilla, Jérôme, et al.
Veröffentlicht: (2024)
LLMs can Schedule
von: Abgaryan, Henrik, et al.
Veröffentlicht: (2024)
von: Abgaryan, Henrik, et al.
Veröffentlicht: (2024)
Generalized Rapid Action Value Estimation in Memory-Constrained Environments
von: Rautureau, Aloïs, et al.
Veröffentlicht: (2026)
von: Rautureau, Aloïs, et al.
Veröffentlicht: (2026)
SpinGPT: A Large-Language-Model Approach to Playing Poker Correctly
von: Maugin, Narada, et al.
Veröffentlicht: (2025)
von: Maugin, Narada, et al.
Veröffentlicht: (2025)
BeeRNA: tertiary structure-based RNA inverse folding using Artificial Bee Colony
von: Mlaweh, Mehyar, et al.
Veröffentlicht: (2025)
von: Mlaweh, Mehyar, et al.
Veröffentlicht: (2025)
Starjob: Dataset for LLM-Driven Job Shop Scheduling
von: Abgaryan, Henrik, et al.
Veröffentlicht: (2025)
von: Abgaryan, Henrik, et al.
Veröffentlicht: (2025)
ACCORD: Autoregressive Constraint-satisfying Generation for COmbinatorial Optimization with Routing and Dynamic attention
von: Abgaryan, Henrik, et al.
Veröffentlicht: (2025)
von: Abgaryan, Henrik, et al.
Veröffentlicht: (2025)
Pareto-NRPA: A Novel Monte-Carlo Search Algorithm for Multi-Objective Optimization
von: Lallouet, Noé, et al.
Veröffentlicht: (2025)
von: Lallouet, Noé, et al.
Veröffentlicht: (2025)
Mixture of Public and Private Distributions in Imperfect Information Games
von: Arjonilla, Jérôme, et al.
Veröffentlicht: (2024)
von: Arjonilla, Jérôme, et al.
Veröffentlicht: (2024)
Deep Reinforcement Learning for 5*5 Multiplayer Go
von: Driss, Brahim, et al.
Veröffentlicht: (2024)
von: Driss, Brahim, et al.
Veröffentlicht: (2024)
Adaptive Bias Generalized Rollout Policy Adaptation on the Flexible Job-Shop Scheduling Problem
von: Kobrosly, Lotfi, et al.
Veröffentlicht: (2025)
von: Kobrosly, Lotfi, et al.
Veröffentlicht: (2025)
Refutation of Spectral Graph Theory Conjectures with Search Algorithms)
von: Roucairol, Milo, et al.
Veröffentlicht: (2024)
von: Roucairol, Milo, et al.
Veröffentlicht: (2024)
Exploration Unbound
von: Arumugam, Dilip, et al.
Veröffentlicht: (2024)
von: Arumugam, Dilip, et al.
Veröffentlicht: (2024)
Mitigating the reconstruction-detection trade-off in VAE-based unsupervised anomaly detection
von: Senellart, Agathe, et al.
Veröffentlicht: (2026)
von: Senellart, Agathe, et al.
Veröffentlicht: (2026)
A Minimax Approach to Ad Hoc Teamwork
von: Villin, Victor, et al.
Veröffentlicht: (2025)
von: Villin, Victor, et al.
Veröffentlicht: (2025)
Refining Minimax Regret for Unsupervised Environment Design
von: Beukman, Michael, et al.
Veröffentlicht: (2024)
von: Beukman, Michael, et al.
Veröffentlicht: (2024)
Minimax Rates and Spectral Distillation for Tree Ensembles
von: Vu, Binh Duc, et al.
Veröffentlicht: (2026)
von: Vu, Binh Duc, et al.
Veröffentlicht: (2026)
Unbounded Harms, Bounded Law: Liability in the Age of Borderless AI
von: Tran, Ha-Chi
Veröffentlicht: (2026)
von: Tran, Ha-Chi
Veröffentlicht: (2026)
Minimax Data Sanitization with Distortion Constraint and Adversarial Inference
von: Moatazedian, Amirarsalan, et al.
Veröffentlicht: (2025)
von: Moatazedian, Amirarsalan, et al.
Veröffentlicht: (2025)
On the Stability and Generalization of First-order Bilevel Minimax Optimization
von: Zhang, Xuelin, et al.
Veröffentlicht: (2026)
von: Zhang, Xuelin, et al.
Veröffentlicht: (2026)
Formal Verification of Minimax Algorithms
von: Wesselink, Wieger, et al.
Veröffentlicht: (2025)
von: Wesselink, Wieger, et al.
Veröffentlicht: (2025)
InfiniSST: Simultaneous Translation of Unbounded Speech with Large Language Model
von: Ouyang, Siqi, et al.
Veröffentlicht: (2025)
von: Ouyang, Siqi, et al.
Veröffentlicht: (2025)
Learning When Not to Learn: Risk-Sensitive Abstention in Bandits with Unbounded Rewards
von: Liaw, Sarah, et al.
Veröffentlicht: (2025)
von: Liaw, Sarah, et al.
Veröffentlicht: (2025)
InfiniteScienceGym: An Unbounded, Procedurally-Generated Benchmark for Scientific Analysis
von: Bentham, Oliver, et al.
Veröffentlicht: (2026)
von: Bentham, Oliver, et al.
Veröffentlicht: (2026)
Two-Fidelity Best-Action Identification for Stochastic Minimax Tree
von: Chen, Peter, et al.
Veröffentlicht: (2026)
von: Chen, Peter, et al.
Veröffentlicht: (2026)
Learning a Generic Value-Selection Heuristic Inside a Constraint Programming Solver
von: Marty, Tom, et al.
Veröffentlicht: (2023)
von: Marty, Tom, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
Minimax Strikes Back
von: Cohen-Solal, Quentin, et al.
Veröffentlicht: (2020) -
Completeness of Unbounded Best-First Minimax and Descent Minimax
von: Cohen-Solal, Quentin
Veröffentlicht: (2026) -
Minibal: Balanced Game-Playing Without Opponent Modeling
von: Cohen-Solal, Quentin, et al.
Veröffentlicht: (2026) -
Study and improvement of search algorithms in two-players perfect information games
von: Cohen-Solal, Quentin
Veröffentlicht: (2025) -
Learning to Play Two-Player Perfect-Information Games without Knowledge
von: Cohen-Solal, Quentin
Veröffentlicht: (2020)