A* Search Without Expansions: Learning Heuristic Functions with Deep Q-Networks
Fuente:
arXiv
Salvato in:
| Autori principali: | Agostinelli, Forest, Shperberg, Shahaf S., Shmakov, Alexander, McAleer, Stephen, Fox, Roy, Baldi, Pierre |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2021
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Beyond Single-Step Updates: Reinforcement Learning of Heuristics with Limited-Horizon Search
di: Hadar, Gal, et al.
Pubblicazione: (2025)
di: Hadar, Gal, et al.
Pubblicazione: (2025)
The DeepXube Software Package for Solving Pathfinding Problems with Learned Heuristic Functions and Search
di: Agostinelli, Forest
Pubblicazione: (2026)
di: Agostinelli, Forest
Pubblicazione: (2026)
Bidirectional Bounded-Suboptimal Heuristic Search with Consistent Heuristics
di: Shperberg, Shahaf S., et al.
Pubblicazione: (2025)
di: Shperberg, Shahaf S., et al.
Pubblicazione: (2025)
Towards Learning Foundation Models for Heuristic Functions to Solve Pathfinding Problems
di: Khandelwal, Vedant, et al.
Pubblicazione: (2024)
di: Khandelwal, Vedant, et al.
Pubblicazione: (2024)
Tree Search for Language Model Agents
di: Koh, Jing Yu, et al.
Pubblicazione: (2024)
di: Koh, Jing Yu, et al.
Pubblicazione: (2024)
From Kinematics to Dynamics: Learning to Refine Hybrid Plans for Physically Feasible Execution
di: Erez, Lidor, et al.
Pubblicazione: (2026)
di: Erez, Lidor, et al.
Pubblicazione: (2026)
Plato's 'Republic'
di: McAleer, Sean
Pubblicazione: (2020)
di: McAleer, Sean
Pubblicazione: (2020)
On Parallel External-Memory Bidirectional Search
di: Siag, Lior, et al.
Pubblicazione: (2024)
di: Siag, Lior, et al.
Pubblicazione: (2024)
Integrating Reinforcement Learning, Action Model Learning, and Numeric Planning for Tackling Complex Tasks
di: Benyamin, Yarin, et al.
Pubblicazione: (2025)
di: Benyamin, Yarin, et al.
Pubblicazione: (2025)
RAMP: Hybrid DRL for Online Learning of Numeric Action Models
di: Benyamin, Yarin, et al.
Pubblicazione: (2026)
di: Benyamin, Yarin, et al.
Pubblicazione: (2026)
Learning Safe Numeric Planning Action Models
di: Mordoch, Argaman, et al.
Pubblicazione: (2023)
di: Mordoch, Argaman, et al.
Pubblicazione: (2023)
Toward PDDL Planning Copilot
di: Benyamin, Yarin, et al.
Pubblicazione: (2025)
di: Benyamin, Yarin, et al.
Pubblicazione: (2025)
Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training
di: Feng, Xidong, et al.
Pubblicazione: (2023)
di: Feng, Xidong, et al.
Pubblicazione: (2023)
Faster Game Solving via Hyperparameter Schedules
di: Zhang, Naifeng, et al.
Pubblicazione: (2024)
di: Zhang, Naifeng, et al.
Pubblicazione: (2024)
Game-Theoretic Multiagent Reinforcement Learning
di: Yang, Yaodong, et al.
Pubblicazione: (2020)
di: Yang, Yaodong, et al.
Pubblicazione: (2020)
Policy Space Response Oracles: A Survey
di: Bighashdel, Ariyan, et al.
Pubblicazione: (2024)
di: Bighashdel, Ariyan, et al.
Pubblicazione: (2024)
PDDLFuse: A Tool for Generating Diverse Planning Domains
di: Khandelwal, Vedant, et al.
Pubblicazione: (2024)
di: Khandelwal, Vedant, et al.
Pubblicazione: (2024)
Game-Theoretic Robust Reinforcement Learning Handles Temporally-Coupled Perturbations
di: Liang, Yongyuan, et al.
Pubblicazione: (2023)
di: Liang, Yongyuan, et al.
Pubblicazione: (2023)
A Systematic Review to Explore Antenatal Care From the Perspectives of Women With Intellectual Disabilities and Midwives
di: Weam Alhulaibi, et al.
Pubblicazione: (2024)
di: Weam Alhulaibi, et al.
Pubblicazione: (2024)
Ensemble Value Functions for Efficient Exploration in Multi-Agent Reinforcement Learning
di: Schäfer, Lukas, et al.
Pubblicazione: (2023)
di: Schäfer, Lukas, et al.
Pubblicazione: (2023)
Grasper: A Generalist Pursuer for Pursuit-Evasion Problems
di: Li, Pengdeng, et al.
Pubblicazione: (2024)
di: Li, Pengdeng, et al.
Pubblicazione: (2024)
Planning and Acting While the Clock Ticks
di: Coles, Andrew, et al.
Pubblicazione: (2024)
di: Coles, Andrew, et al.
Pubblicazione: (2024)
Scalable Mechanism Design for Multi-Agent Path Finding
di: Friedrich, Paul, et al.
Pubblicazione: (2024)
di: Friedrich, Paul, et al.
Pubblicazione: (2024)
Data-Augmented Game Starts for Accelerating Self-Play Exploration in Imperfect Information Games
di: Lanier, JB, et al.
Pubblicazione: (2026)
di: Lanier, JB, et al.
Pubblicazione: (2026)
Illusory Attacks: Information-Theoretic Detectability Matters in Adversarial Attacks
di: Franzmeyer, Tim, et al.
Pubblicazione: (2022)
di: Franzmeyer, Tim, et al.
Pubblicazione: (2022)
AgentKit: Structured LLM Reasoning with Dynamic Graphs
di: Wu, Yue, et al.
Pubblicazione: (2024)
di: Wu, Yue, et al.
Pubblicazione: (2024)
Full Event Particle-Level Unfolding with Variable-Length Latent Variational Diffusion
di: Shmakov, Alexander, et al.
Pubblicazione: (2024)
di: Shmakov, Alexander, et al.
Pubblicazione: (2024)
Orienting and Engaging New Faculty
di: Pamela MacRae, et al.
Pubblicazione: (2024)
di: Pamela MacRae, et al.
Pubblicazione: (2024)
Make the Pertinent Salient: Task-Relevant Reconstruction for Visual Control with Distractions
di: Kim, Kyungmin, et al.
Pubblicazione: (2024)
di: Kim, Kyungmin, et al.
Pubblicazione: (2024)
Llemma: An Open Language Model For Mathematics
di: Azerbayev, Zhangir, et al.
Pubblicazione: (2023)
di: Azerbayev, Zhangir, et al.
Pubblicazione: (2023)
Student Engagement in AI Assisted Complex Problem Solving: A Pilot Study of Human AI Rubik's Cube Collaboration
di: Vanacore, Kirk, et al.
Pubblicazione: (2025)
di: Vanacore, Kirk, et al.
Pubblicazione: (2025)
Effort Allocation for Deadline-Aware Task and Motion Planning: A Metareasoning Approach
di: Sung, Yoonchang, et al.
Pubblicazione: (2024)
di: Sung, Yoonchang, et al.
Pubblicazione: (2024)
Algorithms and Complexity for Computing Nash Equilibria in Adversarial Team Games
di: Anagnostides, Ioannis, et al.
Pubblicazione: (2023)
di: Anagnostides, Ioannis, et al.
Pubblicazione: (2023)
Sample-Efficient Regret-Minimizing Double Oracle in Extensive-Form Games
di: Tang, Xiaohang, et al.
Pubblicazione: (2024)
di: Tang, Xiaohang, et al.
Pubblicazione: (2024)
Enhancing Q-Learning with Large Language Model Heuristics
di: Wu, Xiefeng
Pubblicazione: (2024)
di: Wu, Xiefeng
Pubblicazione: (2024)
Subgoal-Guided Policy Heuristic Search with Learned Subgoals
di: Tuero, Jake, et al.
Pubblicazione: (2025)
di: Tuero, Jake, et al.
Pubblicazione: (2025)
Social Media Bot Policies: Evaluating Passive and Active Enforcement
di: Radivojevic, Kristina, et al.
Pubblicazione: (2024)
di: Radivojevic, Kristina, et al.
Pubblicazione: (2024)
An Extended Jump Functions Benchmark for the Analysis of Randomized Search Heuristics
di: Bambury, Henry, et al.
Pubblicazione: (2021)
di: Bambury, Henry, et al.
Pubblicazione: (2021)
Is Geometry Enough? An Evaluation of Landmark-Based Gaze Estimation
di: Agostinelli, Daniele, et al.
Pubblicazione: (2026)
di: Agostinelli, Daniele, et al.
Pubblicazione: (2026)
Learning Heuristics for Transit Network Design and Improvement with Deep Reinforcement Learning
di: Holliday, Andrew, et al.
Pubblicazione: (2024)
di: Holliday, Andrew, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Beyond Single-Step Updates: Reinforcement Learning of Heuristics with Limited-Horizon Search
di: Hadar, Gal, et al.
Pubblicazione: (2025) -
The DeepXube Software Package for Solving Pathfinding Problems with Learned Heuristic Functions and Search
di: Agostinelli, Forest
Pubblicazione: (2026) -
Bidirectional Bounded-Suboptimal Heuristic Search with Consistent Heuristics
di: Shperberg, Shahaf S., et al.
Pubblicazione: (2025) -
Towards Learning Foundation Models for Heuristic Functions to Solve Pathfinding Problems
di: Khandelwal, Vedant, et al.
Pubblicazione: (2024) -
Tree Search for Language Model Agents
di: Koh, Jing Yu, et al.
Pubblicazione: (2024)