Learning Generalized Policies for Fully Observable Non-Deterministic Planning Domains
Fuente:
arXiv
Guardado en:
| Autores principales: | Hofmann, Till, Geffner, Hector |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Learning More Expressive General Policies for Classical Planning Domains
por: Ståhlberg, Simon, et al.
Publicado: (2024)
por: Ståhlberg, Simon, et al.
Publicado: (2024)
Learning General Policies with Policy Gradient Methods
por: Ståhlberg, Simon, et al.
Publicado: (2025)
por: Ståhlberg, Simon, et al.
Publicado: (2025)
Differentiable Learning of Lifted Action Schemas for Classical Planning
por: Reiter, Jonas, et al.
Publicado: (2026)
por: Reiter, Jonas, et al.
Publicado: (2026)
Learning General Policies From Examples
por: Bonet, Blai, et al.
Publicado: (2025)
por: Bonet, Blai, et al.
Publicado: (2025)
Adversarial Latent-State Training for Robust Policies in Partially Observable Domains
por: Ahuja, Angad Singh
Publicado: (2026)
por: Ahuja, Angad Singh
Publicado: (2026)
Learning to Search and Searching to Learn for Generalization in Planning
por: Aichmüller, Michael, et al.
Publicado: (2026)
por: Aichmüller, Michael, et al.
Publicado: (2026)
Efficient Lookahead Encoding and Abstracted Width for Learning General Policies in Classical Planning
por: Aichmüller, Michael, et al.
Publicado: (2026)
por: Aichmüller, Michael, et al.
Publicado: (2026)
Provable Representation with Efficient Planning for Partial Observable Reinforcement Learning
por: Zhang, Hongming, et al.
Publicado: (2023)
por: Zhang, Hongming, et al.
Publicado: (2023)
Kernel Metric Learning for In-Sample Off-Policy Evaluation of Deterministic RL Policies
por: Lee, Haanvid, et al.
Publicado: (2024)
por: Lee, Haanvid, et al.
Publicado: (2024)
Benchmarking Partial Observability in Reinforcement Learning with a Suite of Memory-Improvable Domains
por: Tao, Ruo Yu, et al.
Publicado: (2025)
por: Tao, Ruo Yu, et al.
Publicado: (2025)
Learning Interpretable Policies in Hindsight-Observable POMDPs through Partially Supervised Reinforcement Learning
por: Lanier, Michael, et al.
Publicado: (2024)
por: Lanier, Michael, et al.
Publicado: (2024)
Soft Deterministic Policy Gradient with Gaussian Smoothing
por: Na, Hyunjun, et al.
Publicado: (2026)
por: Na, Hyunjun, et al.
Publicado: (2026)
On Policy Reuse: An Expressive Language for Representing and Executing General Policies that Call Other Policies
por: Bonet, Blai, et al.
Publicado: (2024)
por: Bonet, Blai, et al.
Publicado: (2024)
Guided Policy Optimization under Partial Observability
por: Li, Yueheng, et al.
Publicado: (2025)
por: Li, Yueheng, et al.
Publicado: (2025)
Quantum Reinforcement Learning by Adaptive Non-local Observables
por: Lin, Hsin-Yi, et al.
Publicado: (2025)
por: Lin, Hsin-Yi, et al.
Publicado: (2025)
Symmetries and Expressive Requirements for Learning General Policies
por: Drexler, Dominik, et al.
Publicado: (2024)
por: Drexler, Dominik, et al.
Publicado: (2024)
Deterministic Policy Gradient for Reinforcement Learning with Continuous Time and State
por: Cheng, Ziheng, et al.
Publicado: (2025)
por: Cheng, Ziheng, et al.
Publicado: (2025)
Deterministic Decomposition of Stochastic Generative Dynamics
por: Song, Xingyu, et al.
Publicado: (2026)
por: Song, Xingyu, et al.
Publicado: (2026)
Per-Domain Generalizing Policies: On Validation Instances and Scaling Behavior
por: Gros, Timo P., et al.
Publicado: (2025)
por: Gros, Timo P., et al.
Publicado: (2025)
Mitigating Suboptimality of Deterministic Policy Gradients in Complex Q-functions
por: Jain, Ayush, et al.
Publicado: (2024)
por: Jain, Ayush, et al.
Publicado: (2024)
Domain Adversarial Active Learning for Domain Generalization Classification
por: Chen, Jianting, et al.
Publicado: (2024)
por: Chen, Jianting, et al.
Publicado: (2024)
Per-Domain Generalizing Policies: On Learning Efficient and Robust Q-Value Functions (Extended Version with Technical Appendix)
por: Müller, Nicola J., et al.
Publicado: (2026)
por: Müller, Nicola J., et al.
Publicado: (2026)
Zero-Shot Reinforcement Learning Under Partial Observability
por: Jeen, Scott, et al.
Publicado: (2025)
por: Jeen, Scott, et al.
Publicado: (2025)
Multi-View Causal Representation Learning with Partial Observability
por: Yao, Dingling, et al.
Publicado: (2023)
por: Yao, Dingling, et al.
Publicado: (2023)
Limits of Actor-Critic Algorithms for Decision Tree Policies Learning in IBMDPs
por: Kohler, Hector, et al.
Publicado: (2023)
por: Kohler, Hector, et al.
Publicado: (2023)
Edge Delayed Deep Deterministic Policy Gradient: efficient continuous control for edge scenarios
por: Sinigaglia, Alberto, et al.
Publicado: (2024)
por: Sinigaglia, Alberto, et al.
Publicado: (2024)
Taming the Adversary: Stable Minimax Deep Deterministic Policy Gradient via Fractional Objectives
por: Lee, Taeho, et al.
Publicado: (2026)
por: Lee, Taeho, et al.
Publicado: (2026)
Recursive Backwards Q-Learning in Deterministic Environments
por: Diekhoff, Jan, et al.
Publicado: (2024)
por: Diekhoff, Jan, et al.
Publicado: (2024)
Policy Gradient Methods for Non-Markovian Reinforcement Learning
por: Kar, Avik, et al.
Publicado: (2026)
por: Kar, Avik, et al.
Publicado: (2026)
Planning with a Learned Policy Basis to Optimally Solve Complex Tasks
por: Infante, Guillermo, et al.
Publicado: (2024)
por: Infante, Guillermo, et al.
Publicado: (2024)
Interpretable and Editable Programmatic Tree Policies for Reinforcement Learning
por: Kohler, Hector, et al.
Publicado: (2024)
por: Kohler, Hector, et al.
Publicado: (2024)
Evaluating Interpretable Reinforcement Learning by Distilling Policies into Programs
por: Kohler, Hector, et al.
Publicado: (2025)
por: Kohler, Hector, et al.
Publicado: (2025)
A Sparsity Principle for Partially Observable Causal Representation Learning
por: Xu, Danru, et al.
Publicado: (2024)
por: Xu, Danru, et al.
Publicado: (2024)
Adaptive Non-local Observable on Quantum Neural Networks
por: Lin, Hsin-Yi, et al.
Publicado: (2025)
por: Lin, Hsin-Yi, et al.
Publicado: (2025)
Quantum Super-resolution by Adaptive Non-local Observables
por: Lin, Hsin-Yi, et al.
Publicado: (2026)
por: Lin, Hsin-Yi, et al.
Publicado: (2026)
Sketch Decompositions for Classical Planning via Deep Reinforcement Learning
por: Aichmüller, Michael, et al.
Publicado: (2024)
por: Aichmüller, Michael, et al.
Publicado: (2024)
One for All: A Non-Linear Transformer can Enable Cross-Domain Generalization for In-Context Reinforcement Learning
por: He, Bowen, et al.
Publicado: (2026)
por: He, Bowen, et al.
Publicado: (2026)
A Generic Machine Learning Framework for Fully-Unsupervised Anomaly Detection with Contaminated Data
por: Ulmer, Markus, et al.
Publicado: (2023)
por: Ulmer, Markus, et al.
Publicado: (2023)
Domain-Generalization to Improve Learning in Meta-Learning Algorithms
por: Anjum, Usman, et al.
Publicado: (2025)
por: Anjum, Usman, et al.
Publicado: (2025)
Off-Policy Evaluation and Learning for the Future under Non-Stationarity
por: Shimizu, Tatsuhiro, et al.
Publicado: (2025)
por: Shimizu, Tatsuhiro, et al.
Publicado: (2025)
Ejemplares similares
-
Learning More Expressive General Policies for Classical Planning Domains
por: Ståhlberg, Simon, et al.
Publicado: (2024) -
Learning General Policies with Policy Gradient Methods
por: Ståhlberg, Simon, et al.
Publicado: (2025) -
Differentiable Learning of Lifted Action Schemas for Classical Planning
por: Reiter, Jonas, et al.
Publicado: (2026) -
Learning General Policies From Examples
por: Bonet, Blai, et al.
Publicado: (2025) -
Adversarial Latent-State Training for Robust Policies in Partially Observable Domains
por: Ahuja, Angad Singh
Publicado: (2026)