Self-Guiding Exploration for Combinatorial Problems
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Iklassov, Zangir, Du, Yali, Akimov, Farkhad, Takac, Martin |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
The AI Data Scientist
von: Akimov, Farkhad, et al.
Veröffentlicht: (2025)
von: Akimov, Farkhad, et al.
Veröffentlicht: (2025)
Reinforcement Learning for Solving Stochastic Vehicle Routing Problem with Time Windows
von: Iklassov, Zangir, et al.
Veröffentlicht: (2024)
von: Iklassov, Zangir, et al.
Veröffentlicht: (2024)
SVRPBench: A Realistic Benchmark for Stochastic Vehicle Routing Problem
von: Heakl, Ahmed, et al.
Veröffentlicht: (2025)
von: Heakl, Ahmed, et al.
Veröffentlicht: (2025)
LLM-BABYBENCH: Understanding and Evaluating Grounded Planning and Reasoning in LLMs
von: Choukrani, Omar, et al.
Veröffentlicht: (2025)
von: Choukrani, Omar, et al.
Veröffentlicht: (2025)
RECALL: Library-Like Behavior In Language Models is Enhanced by Self-Referencing Causal Cycles
von: Nwadike, Munachiso, et al.
Veröffentlicht: (2025)
von: Nwadike, Munachiso, et al.
Veröffentlicht: (2025)
Measuring AI Reasoning: A Guide for Researchers
von: Nwadike, Munachiso Samuel, et al.
Veröffentlicht: (2026)
von: Nwadike, Munachiso Samuel, et al.
Veröffentlicht: (2026)
A Decade of Deep Learning: A Survey on The Magnificent Seven
von: Azizov, Dilshod, et al.
Veröffentlicht: (2024)
von: Azizov, Dilshod, et al.
Veröffentlicht: (2024)
VAM: Verbalized Action Masking for Controllable Exploration in RL Post-Training -- A Chess Case Study
von: Zhang, Zhicheng, et al.
Veröffentlicht: (2026)
von: Zhang, Zhicheng, et al.
Veröffentlicht: (2026)
The Initial Exploration Problem in Knowledge Graph Exploration
von: McNamara, Claire, et al.
Veröffentlicht: (2026)
von: McNamara, Claire, et al.
Veröffentlicht: (2026)
LLM-First Search: Self-Guided Exploration of the Solution Space
von: Herr, Nathan, et al.
Veröffentlicht: (2025)
von: Herr, Nathan, et al.
Veröffentlicht: (2025)
XCSP3: An Integrated Format for Benchmarking Combinatorial Constrained Problems
von: Boussemart, Frederic, et al.
Veröffentlicht: (2016)
von: Boussemart, Frederic, et al.
Veröffentlicht: (2016)
PyCSP3: Modeling Combinatorial Constrained Problems in Python
von: Lecoutre, Christophe, et al.
Veröffentlicht: (2020)
von: Lecoutre, Christophe, et al.
Veröffentlicht: (2020)
Permutation Picture of Graph Combinatorial Optimization Problems
von: Min, Yimeng
Veröffentlicht: (2024)
von: Min, Yimeng
Veröffentlicht: (2024)
CORE: Collaborative Reasoning via Cross Teaching
von: Mishra, Kshitij, et al.
Veröffentlicht: (2026)
von: Mishra, Kshitij, et al.
Veröffentlicht: (2026)
Self-Play Only Evolves When Self-Synthetic Pipeline Ensures Learnable Information Gain
von: Liu, Wei, et al.
Veröffentlicht: (2026)
von: Liu, Wei, et al.
Veröffentlicht: (2026)
DISCO: Efficient Diffusion Solver for Large-Scale Combinatorial Optimization Problems
von: Zhao, Hang, et al.
Veröffentlicht: (2024)
von: Zhao, Hang, et al.
Veröffentlicht: (2024)
EchoTrail-GUI: Building Actionable Memory for GUI Agents via Critic-Guided Self-Exploration
von: Li, Runze, et al.
Veröffentlicht: (2025)
von: Li, Runze, et al.
Veröffentlicht: (2025)
Learning a Prior for Monte Carlo Search by Replaying Solutions to Combinatorial Problems
von: Cazenave, Tristan
Veröffentlicht: (2024)
von: Cazenave, Tristan
Veröffentlicht: (2024)
DCP-Bench-Open: Evaluating LLMs for Constraint Modelling of Discrete Combinatorial Problems
von: Michailidis, Kostis, et al.
Veröffentlicht: (2025)
von: Michailidis, Kostis, et al.
Veröffentlicht: (2025)
EALG: Evolutionary Adversarial Generation of Language Model-Guided Generators for Combinatorial Optimization
von: Duan, Ruibo, et al.
Veröffentlicht: (2025)
von: Duan, Ruibo, et al.
Veröffentlicht: (2025)
Eyeballing Combinatorial Problems: A Case Study of Using Multimodal Large Language Models to Solve Traveling Salesman Problems
von: Elhenawy, Mohammed, et al.
Veröffentlicht: (2024)
von: Elhenawy, Mohammed, et al.
Veröffentlicht: (2024)
Attention-based Reinforcement Learning for Combinatorial Optimization: Application to Job Shop Scheduling Problem
von: Lee, Jaejin, et al.
Veröffentlicht: (2024)
von: Lee, Jaejin, et al.
Veröffentlicht: (2024)
Post-Incorporating Code Structural Knowledge into Pretrained Models via ICL for Code Translation
von: Du, Yali, et al.
Veröffentlicht: (2025)
von: Du, Yali, et al.
Veröffentlicht: (2025)
Customized Exploration of Landscape Features Driving Multi-Objective Combinatorial Optimization Performance
von: Nikolikj, Ana, et al.
Veröffentlicht: (2025)
von: Nikolikj, Ana, et al.
Veröffentlicht: (2025)
Latent Reasoning in TRMs is Secretly a Policy Improvement Operator
von: Asadulaev, Arip, et al.
Veröffentlicht: (2025)
von: Asadulaev, Arip, et al.
Veröffentlicht: (2025)
Self-Improved Learning for Scalable Neural Combinatorial Optimization
von: Luo, Fu, et al.
Veröffentlicht: (2024)
von: Luo, Fu, et al.
Veröffentlicht: (2024)
Self-Predictive Representations for Combinatorial Generalization in Behavioral Cloning
von: Lawson, Daniel, et al.
Veröffentlicht: (2025)
von: Lawson, Daniel, et al.
Veröffentlicht: (2025)
ProSEA: Problem Solving via Exploration Agents
von: Nguyen, William, et al.
Veröffentlicht: (2025)
von: Nguyen, William, et al.
Veröffentlicht: (2025)
Neural Combinatorial Optimization Algorithms for Solving Vehicle Routing Problems: A Comprehensive Survey with Perspectives
von: Wu, Xuan, et al.
Veröffentlicht: (2024)
von: Wu, Xuan, et al.
Veröffentlicht: (2024)
Towards a Generic Representation of Combinatorial Problems for Learning-Based Approaches
von: Boisvert, Léo, et al.
Veröffentlicht: (2024)
von: Boisvert, Léo, et al.
Veröffentlicht: (2024)
MARGE: Improving Math Reasoning for LLMs with Guided Exploration
von: Gao, Jingyue, et al.
Veröffentlicht: (2025)
von: Gao, Jingyue, et al.
Veröffentlicht: (2025)
Preference-Based Gradient Estimation for ML-Guided Approximate Combinatorial Optimization
von: Mielke, Arman, et al.
Veröffentlicht: (2025)
von: Mielke, Arman, et al.
Veröffentlicht: (2025)
Edge-wise Topological Divergence Gaps: Guiding Search in Combinatorial Optimization
von: Trofimov, Ilya, et al.
Veröffentlicht: (2025)
von: Trofimov, Ilya, et al.
Veröffentlicht: (2025)
A Comparative User Evaluation of XRL Explanations using Goal Identification
von: Towers, Mark, et al.
Veröffentlicht: (2025)
von: Towers, Mark, et al.
Veröffentlicht: (2025)
Reinforcement Learning by Guided Safe Exploration
von: Yang, Qisong, et al.
Veröffentlicht: (2023)
von: Yang, Qisong, et al.
Veröffentlicht: (2023)
Temporal Contrastive Decoding: A Training-Free Method for Large Audio-Language Models
von: Li, Yanda, et al.
Veröffentlicht: (2026)
von: Li, Yanda, et al.
Veröffentlicht: (2026)
Boosting Inference with Guided Reasoning: Stochastic Exploration for Recursive Models
von: Corbett, Andrew, et al.
Veröffentlicht: (2026)
von: Corbett, Andrew, et al.
Veröffentlicht: (2026)
Uncovering the Spectral Bias in Diagonal State Space Models
von: Solozabal, Ruben, et al.
Veröffentlicht: (2025)
von: Solozabal, Ruben, et al.
Veröffentlicht: (2025)
NCO4CVRP: Neural Combinatorial Optimization for the Capacitated Vehicle Routing Problem
von: Dihan, Mahir Labib, et al.
Veröffentlicht: (2026)
von: Dihan, Mahir Labib, et al.
Veröffentlicht: (2026)
Breaking the Martingale Curse: Multi-Agent Debate via Asymmetric Cognitive Potential Energy
von: Liu, Yuhan, et al.
Veröffentlicht: (2026)
von: Liu, Yuhan, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
The AI Data Scientist
von: Akimov, Farkhad, et al.
Veröffentlicht: (2025) -
Reinforcement Learning for Solving Stochastic Vehicle Routing Problem with Time Windows
von: Iklassov, Zangir, et al.
Veröffentlicht: (2024) -
SVRPBench: A Realistic Benchmark for Stochastic Vehicle Routing Problem
von: Heakl, Ahmed, et al.
Veröffentlicht: (2025) -
LLM-BABYBENCH: Understanding and Evaluating Grounded Planning and Reasoning in LLMs
von: Choukrani, Omar, et al.
Veröffentlicht: (2025) -
RECALL: Library-Like Behavior In Language Models is Enhanced by Self-Referencing Causal Cycles
von: Nwadike, Munachiso, et al.
Veröffentlicht: (2025)