Finding Increasingly Large Extremal Graphs with AlphaZero and Tabu Search
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Mehrabian, Abbas, Anand, Ankit, Kim, Hyunjik, Sonnerat, Nicolas, Balog, Matej, Comanici, Gheorghe, Berariu, Tudor, Lee, Andrew, Ruoss, Anian, Bulanova, Anna, Toyama, Daniel, Blackwell, Sam, Paredes, Bernardino Romera, Veličković, Petar, Orseau, Laurent, Lee, Joonkyung, Naredla, Anurag Murty, Precup, Doina, Wagner, Adam Zsolt |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2023
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Affordances Enable Partial World Modeling with LLMs
von: Khetarpal, Khimya, et al.
Veröffentlicht: (2026)
von: Khetarpal, Khimya, et al.
Veröffentlicht: (2026)
Understanding Prompt Tuning and In-Context Learning via Meta-Learning
von: Genewein, Tim, et al.
Veröffentlicht: (2025)
von: Genewein, Tim, et al.
Veröffentlicht: (2025)
Extremal numbers and Sidorenko's conjecture
von: Conlon, David, et al.
Veröffentlicht: (2023)
von: Conlon, David, et al.
Veröffentlicht: (2023)
Separating Two Points with Obstacles in the Plane: Improved Upper and Lower Bounds
von: Spalding-Jamieson, Jack, et al.
Veröffentlicht: (2025)
von: Spalding-Jamieson, Jack, et al.
Veröffentlicht: (2025)
Diversity-Enriched Option-Critic
von: Kamat, Anand, et al.
Veröffentlicht: (2020)
von: Kamat, Anand, et al.
Veröffentlicht: (2020)
Functional Acceleration for Policy Mirror Descent
von: Chelu, Veronica, et al.
Veröffentlicht: (2024)
von: Chelu, Veronica, et al.
Veröffentlicht: (2024)
A Look at Value-Based Decision-Time vs. Background Planning Methods Across Different Settings
von: Alver, Safa, et al.
Veröffentlicht: (2022)
von: Alver, Safa, et al.
Veröffentlicht: (2022)
The Art of Being Difficult: Combining Human and AI Strengths to Find Adversarial Instances for Heuristics
von: Nikoleit, Henri, et al.
Veröffentlicht: (2026)
von: Nikoleit, Henri, et al.
Veröffentlicht: (2026)
Compression via Pre-trained Transformers: A Study on Byte-Level Multimodal Data
von: Heurtel-Depeiges, David, et al.
Veröffentlicht: (2024)
von: Heurtel-Depeiges, David, et al.
Veröffentlicht: (2024)
From Characters to Tokens: Dynamic Grouping with Hierarchical BPE
von: Dolga, Rares, et al.
Veröffentlicht: (2025)
von: Dolga, Rares, et al.
Veröffentlicht: (2025)
Balancing Plasticity and Stability with Fast and Slow Successor Features
von: Chua, Raymond, et al.
Veröffentlicht: (2026)
von: Chua, Raymond, et al.
Veröffentlicht: (2026)
On the Privacy of Selection Mechanisms with Gaussian Noise
von: Lebensold, Jonathan, et al.
Veröffentlicht: (2024)
von: Lebensold, Jonathan, et al.
Veröffentlicht: (2024)
Domination inequalities and dominating graphs
von: Conlon, David, et al.
Veröffentlicht: (2023)
von: Conlon, David, et al.
Veröffentlicht: (2023)
Lower bounds for multivariate independence polynomials and their generalisations
von: Lee, Joonkyung, et al.
Veröffentlicht: (2026)
von: Lee, Joonkyung, et al.
Veröffentlicht: (2026)
Learning Universal Predictors
von: Grau-Moya, Jordi, et al.
Veröffentlicht: (2024)
von: Grau-Moya, Jordi, et al.
Veröffentlicht: (2024)
Why is prompting hard? Understanding prompts on binary sequence predictors
von: Wenliang, Li Kevin, et al.
Veröffentlicht: (2025)
von: Wenliang, Li Kevin, et al.
Veröffentlicht: (2025)
LMAct: A Benchmark for In-Context Imitation Learning with Long Multimodal Demonstrations
von: Ruoss, Anian, et al.
Veröffentlicht: (2024)
von: Ruoss, Anian, et al.
Veröffentlicht: (2024)
Conditions on Preference Relations that Guarantee the Existence of Optimal Policies
von: Carr, Jonathan Colaço, et al.
Veröffentlicht: (2023)
von: Carr, Jonathan Colaço, et al.
Veröffentlicht: (2023)
Sparse-Reg: Improving Sample Complexity in Offline Reinforcement Learning using Sparsity
von: Arnob, Samin Yeasar, et al.
Veröffentlicht: (2025)
von: Arnob, Samin Yeasar, et al.
Veröffentlicht: (2025)
Adaptive Exploration for Data-Efficient General Value Function Evaluations
von: Jain, Arushi, et al.
Veröffentlicht: (2024)
von: Jain, Arushi, et al.
Veröffentlicht: (2024)
Fluid-Agent Reinforcement Learning
von: Sharma, Shishir, et al.
Veröffentlicht: (2026)
von: Sharma, Shishir, et al.
Veröffentlicht: (2026)
Partial Models for Building Adaptive Model-Based Reinforcement Learning Agents
von: Alver, Safa, et al.
Veröffentlicht: (2024)
von: Alver, Safa, et al.
Veröffentlicht: (2024)
AlphaZero-Edu: Democratizing Access to AlphaZero
von: Li, Ruitong, et al.
Veröffentlicht: (2025)
von: Li, Ruitong, et al.
Veröffentlicht: (2025)
Language Modeling Is Compression
von: Delétang, Grégoire, et al.
Veröffentlicht: (2023)
von: Delétang, Grégoire, et al.
Veröffentlicht: (2023)
Parseval Regularization for Continual Reinforcement Learning
von: Chung, Wesley, et al.
Veröffentlicht: (2024)
von: Chung, Wesley, et al.
Veröffentlicht: (2024)
Relative Trajectory Balance is equivalent to Trust-PCL
von: Deleu, Tristan, et al.
Veröffentlicht: (2025)
von: Deleu, Tristan, et al.
Veröffentlicht: (2025)
Auf Pump
von: Ruoss, Matthias
Veröffentlicht: (2025)
von: Ruoss, Matthias
Veröffentlicht: (2025)
Schweizerdeutsch und Sprachbewusstsein
von: Ruoss, Emanuel
Veröffentlicht: (2020)
von: Ruoss, Emanuel
Veröffentlicht: (2020)
Non-invasive evaluation of electrical brain activity: effects of non-medicative treatment and subject’s state
von: Inga Griškova-Bulanova
Veröffentlicht: (2013)
von: Inga Griškova-Bulanova
Veröffentlicht: (2013)
Around the positive graph conjecture
von: Conlon, David, et al.
Veröffentlicht: (2024)
von: Conlon, David, et al.
Veröffentlicht: (2024)
Counting homomorphisms in antiferromagnetic graphs via Lorentzian polynomials
von: Lee, Joonkyung, et al.
Veröffentlicht: (2025)
von: Lee, Joonkyung, et al.
Veröffentlicht: (2025)
Chi-boundedness of graphs containing no cycles with $k$ chords
von: Lee, Joonkyung, et al.
Veröffentlicht: (2022)
von: Lee, Joonkyung, et al.
Veröffentlicht: (2022)
On the extremal number of incidence graphs
von: Baek, Jisun, et al.
Veröffentlicht: (2024)
von: Baek, Jisun, et al.
Veröffentlicht: (2024)
Distributional Bellman Operators over Mean Embeddings
von: Wenliang, Li Kevin, et al.
Veröffentlicht: (2023)
von: Wenliang, Li Kevin, et al.
Veröffentlicht: (2023)
Incorporating Spatial Information into Goal-Conditioned Hierarchical Reinforcement Learning via Graph Representations
von: Zhang, Shuyuan, et al.
Veröffentlicht: (2025)
von: Zhang, Shuyuan, et al.
Veröffentlicht: (2025)
SCAR: Shapley Credit Assignment for More Efficient RLHF
von: Cao, Meng, et al.
Veröffentlicht: (2025)
von: Cao, Meng, et al.
Veröffentlicht: (2025)
Capacity-Constrained Continual Learning
von: Wen, Zheng, et al.
Veröffentlicht: (2025)
von: Wen, Zheng, et al.
Veröffentlicht: (2025)
Finite time analysis of temporal difference learning with linear function approximation: Tail averaging and regularisation
von: Patil, Gandharv, et al.
Veröffentlicht: (2022)
von: Patil, Gandharv, et al.
Veröffentlicht: (2022)
Langevin Soft Actor-Critic: Efficient Exploration through Uncertainty-Driven Critic Learning
von: Ishfaq, Haque, et al.
Veröffentlicht: (2025)
von: Ishfaq, Haque, et al.
Veröffentlicht: (2025)
Theoretical Steps to Optimize Transportation in the Cubic Networks and the Congestion Paradox
von: Yoo, Joonkyung
Veröffentlicht: (2024)
von: Yoo, Joonkyung
Veröffentlicht: (2024)
Ähnliche Einträge
-
Affordances Enable Partial World Modeling with LLMs
von: Khetarpal, Khimya, et al.
Veröffentlicht: (2026) -
Understanding Prompt Tuning and In-Context Learning via Meta-Learning
von: Genewein, Tim, et al.
Veröffentlicht: (2025) -
Extremal numbers and Sidorenko's conjecture
von: Conlon, David, et al.
Veröffentlicht: (2023) -
Separating Two Points with Obstacles in the Plane: Improved Upper and Lower Bounds
von: Spalding-Jamieson, Jack, et al.
Veröffentlicht: (2025) -
Diversity-Enriched Option-Critic
von: Kamat, Anand, et al.
Veröffentlicht: (2020)