On shallow planning under partial observability
Fuente:
arXiv
Saved in:
| Main Authors: | Lefebvre, Randy, Durand, Audrey |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
An intelligent tutor for planning in large partially observable environments
by: Heindrich, Lovis, et al.
Published: (2023)
by: Heindrich, Lovis, et al.
Published: (2023)
LLM-as-a-Judge: Toward World Models for Slate Recommendation Systems
by: Bonin, Baptiste, et al.
Published: (2025)
by: Bonin, Baptiste, et al.
Published: (2025)
Propagating the prior from shallow to deep with a pre-trained velocity-model Generative Transformer network
by: Harsuko, Randy, et al.
Published: (2024)
by: Harsuko, Randy, et al.
Published: (2024)
Nested-ReFT: Efficient Reinforcement Learning for Large Language Model Fine-Tuning via Off-Policy Rollouts
by: Heuillet, Maxime, et al.
Published: (2025)
by: Heuillet, Maxime, et al.
Published: (2025)
Artificial intelligence and machine learning generated conjectures with TxGraffiti
by: Davila, Randy
Published: (2024)
by: Davila, Randy
Published: (2024)
Automated conjecturing with \emph{TxGraffiti}
by: Davila, Randy
Published: (2024)
by: Davila, Randy
Published: (2024)
The \emph{Optimist}: Towards Fully Automated Graph Theory Research
by: Davila, Randy
Published: (2024)
by: Davila, Randy
Published: (2024)
Active Inference and Reinforcement Learning: A unified inference on continuous state and action spaces under partial observability
by: Malekzadeh, Parvin, et al.
Published: (2022)
by: Malekzadeh, Parvin, et al.
Published: (2022)
Intelligent prospector v2.0: exploration drill planning under epistemic model uncertainty
by: Mern, John, et al.
Published: (2024)
by: Mern, John, et al.
Published: (2024)
Adaptive mine planning under geological uncertainty: A POMDP framework for sequential decision-making
by: Khalifi, Hamza, et al.
Published: (2026)
by: Khalifi, Hamza, et al.
Published: (2026)
Forager: a lightweight testbed for continual learning with partial observability in RL
by: Tang, Steven, et al.
Published: (2026)
by: Tang, Steven, et al.
Published: (2026)
Balancing Sparse RNNs with Hyperparameterization Benefiting Meta-Learning
by: Hershey, Quincy, et al.
Published: (2025)
by: Hershey, Quincy, et al.
Published: (2025)
Automated planning with ontologies under coherence update semantics (Extended Version)
by: Borgwardt, Stefan, et al.
Published: (2025)
by: Borgwardt, Stefan, et al.
Published: (2025)
Optimal sensor deception in stochastic environments with partial observability to mislead a robot to a decoy goal
by: Rahmani, Hazhar, et al.
Published: (2025)
by: Rahmani, Hazhar, et al.
Published: (2025)
Safety Implications of Explainable Artificial Intelligence in End-to-End Autonomous Driving
by: Atakishiyev, Shahin, et al.
Published: (2024)
by: Atakishiyev, Shahin, et al.
Published: (2024)
Incorporating Explanations into Human-Machine Interfaces for Trust and Situation Awareness in Autonomous Vehicles
by: Atakishiyev, Shahin, et al.
Published: (2024)
by: Atakishiyev, Shahin, et al.
Published: (2024)
A generative foundation model for an all-in-one seismic processing framework
by: Cheng, Shijun, et al.
Published: (2025)
by: Cheng, Shijun, et al.
Published: (2025)
Graph Neural Networks vs Convolutional Neural Networks for Graph Domination Number Prediction
by: Davila, Randy, et al.
Published: (2025)
by: Davila, Randy, et al.
Published: (2025)
Principled Curriculum Learning using Parameter Continuation Methods
by: Pathak, Harsh Nilesh, et al.
Published: (2025)
by: Pathak, Harsh Nilesh, et al.
Published: (2025)
IMAGINE: An 8-to-1b 22nm FD-SOI Compute-In-Memory CNN Accelerator With an End-to-End Analog Charge-Based 0.15-8POPS/W Macro Featuring Distribution-Aware Data Reshaping
by: Kneip, Adrian, et al.
Published: (2024)
by: Kneip, Adrian, et al.
Published: (2024)
An unconditional distribution learning advantage with shallow quantum circuits
by: Pirnay, N., et al.
Published: (2024)
by: Pirnay, N., et al.
Published: (2024)
Experience-driven discovery of planning strategies
by: He, Ruiqi, et al.
Published: (2024)
by: He, Ruiqi, et al.
Published: (2024)
Synthesizing world models for bilevel planning
by: Ahmed, Zergham, et al.
Published: (2025)
by: Ahmed, Zergham, et al.
Published: (2025)
Getting SMARTER for Motion Planning in Autonomous Driving Systems
by: Alban, Montgomery, et al.
Published: (2025)
by: Alban, Montgomery, et al.
Published: (2025)
Ask before you Build: Rethinking AI-for-Good in Human Trafficking Interventions
by: Nair, Pratheeksha, et al.
Published: (2025)
by: Nair, Pratheeksha, et al.
Published: (2025)
LLM Robustness Leaderboard v1 --Technical report
by: Lefebvre, Pierre Peigné -, et al.
Published: (2025)
by: Lefebvre, Pierre Peigné -, et al.
Published: (2025)
Robust Fine-Tuning from Non-Robust Pretrained Models: Mitigating Suboptimal Transfer With Epsilon-Scheduling
by: Ngnawé, Jonas, et al.
Published: (2025)
by: Ngnawé, Jonas, et al.
Published: (2025)
CSM-H-R: A Context Modeling Framework in Supporting Reasoning Automation for Interoperable Intelligent Systems and Privacy Protection
by: Yue, Songhui, et al.
Published: (2023)
by: Yue, Songhui, et al.
Published: (2023)
Synthesis of timeline-based planning strategies avoiding determinization
by: Della Monica, Dario, et al.
Published: (2025)
by: Della Monica, Dario, et al.
Published: (2025)
ASKCOS: an open source software suite for synthesis planning
by: Tu, Zhengkai, et al.
Published: (2025)
by: Tu, Zhengkai, et al.
Published: (2025)
Individual differences in the cognitive mechanisms of planning strategy discovery
by: He, Ruiqi, et al.
Published: (2025)
by: He, Ruiqi, et al.
Published: (2025)
Estimating cognitive biases with attention-aware inverse planning
by: Banerjee, Sounak, et al.
Published: (2025)
by: Banerjee, Sounak, et al.
Published: (2025)
Developing Autonomous Robot-Mediated Behavior Coaching Sessions with Haru
by: Jelínek, Matouš, et al.
Published: (2024)
by: Jelínek, Matouš, et al.
Published: (2024)
In Reverie Together: Ten Years of Mathematical Discovery with a Machine Collaborator
by: Davila, Randy, et al.
Published: (2025)
by: Davila, Randy, et al.
Published: (2025)
Solo Connection: A Parameter Efficient Fine-Tuning Technique for Transformers
by: Pathak, Harsh Nilesh, et al.
Published: (2025)
by: Pathak, Harsh Nilesh, et al.
Published: (2025)
Proceedings of the First International Workshop on Next-Generation Language Models for Knowledge Representation and Reasoning (NeLaMKRR 2024)
by: Satoh, Ken, et al.
Published: (2024)
by: Satoh, Ken, et al.
Published: (2024)
Proceedings of the Second International Workshop on Next-Generation Language Models for Knowledge Representation and Reasoning (NeLaMKRR 2025)
by: Nguyen, Ha-Thanh, et al.
Published: (2025)
by: Nguyen, Ha-Thanh, et al.
Published: (2025)
Transformers perform adaptive partial pooling
by: Kapatsinski, Vsevolod
Published: (2026)
by: Kapatsinski, Vsevolod
Published: (2026)
Synthelite: Chemist-aligned and feasibility-aware synthesis planning with LLMs
by: Xuan-Vu, Nguyen, et al.
Published: (2025)
by: Xuan-Vu, Nguyen, et al.
Published: (2025)
Exploring the hierarchical structure of human plans via program generation
by: Correa, Carlos G., et al.
Published: (2023)
by: Correa, Carlos G., et al.
Published: (2023)
Similar Items
-
An intelligent tutor for planning in large partially observable environments
by: Heindrich, Lovis, et al.
Published: (2023) -
LLM-as-a-Judge: Toward World Models for Slate Recommendation Systems
by: Bonin, Baptiste, et al.
Published: (2025) -
Propagating the prior from shallow to deep with a pre-trained velocity-model Generative Transformer network
by: Harsuko, Randy, et al.
Published: (2024) -
Nested-ReFT: Efficient Reinforcement Learning for Large Language Model Fine-Tuning via Off-Policy Rollouts
by: Heuillet, Maxime, et al.
Published: (2025) -
Artificial intelligence and machine learning generated conjectures with TxGraffiti
by: Davila, Randy
Published: (2024)