On shallow planning under partial observability
Fuente:
arXiv
Guardado en:
| Autores principales: | Lefebvre, Randy, Durand, Audrey |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
An intelligent tutor for planning in large partially observable environments
por: Heindrich, Lovis, et al.
Publicado: (2023)
por: Heindrich, Lovis, et al.
Publicado: (2023)
LLM-as-a-Judge: Toward World Models for Slate Recommendation Systems
por: Bonin, Baptiste, et al.
Publicado: (2025)
por: Bonin, Baptiste, et al.
Publicado: (2025)
Propagating the prior from shallow to deep with a pre-trained velocity-model Generative Transformer network
por: Harsuko, Randy, et al.
Publicado: (2024)
por: Harsuko, Randy, et al.
Publicado: (2024)
Nested-ReFT: Efficient Reinforcement Learning for Large Language Model Fine-Tuning via Off-Policy Rollouts
por: Heuillet, Maxime, et al.
Publicado: (2025)
por: Heuillet, Maxime, et al.
Publicado: (2025)
Artificial intelligence and machine learning generated conjectures with TxGraffiti
por: Davila, Randy
Publicado: (2024)
por: Davila, Randy
Publicado: (2024)
Automated conjecturing with \emph{TxGraffiti}
por: Davila, Randy
Publicado: (2024)
por: Davila, Randy
Publicado: (2024)
The \emph{Optimist}: Towards Fully Automated Graph Theory Research
por: Davila, Randy
Publicado: (2024)
por: Davila, Randy
Publicado: (2024)
Active Inference and Reinforcement Learning: A unified inference on continuous state and action spaces under partial observability
por: Malekzadeh, Parvin, et al.
Publicado: (2022)
por: Malekzadeh, Parvin, et al.
Publicado: (2022)
Intelligent prospector v2.0: exploration drill planning under epistemic model uncertainty
por: Mern, John, et al.
Publicado: (2024)
por: Mern, John, et al.
Publicado: (2024)
Adaptive mine planning under geological uncertainty: A POMDP framework for sequential decision-making
por: Khalifi, Hamza, et al.
Publicado: (2026)
por: Khalifi, Hamza, et al.
Publicado: (2026)
Forager: a lightweight testbed for continual learning with partial observability in RL
por: Tang, Steven, et al.
Publicado: (2026)
por: Tang, Steven, et al.
Publicado: (2026)
Balancing Sparse RNNs with Hyperparameterization Benefiting Meta-Learning
por: Hershey, Quincy, et al.
Publicado: (2025)
por: Hershey, Quincy, et al.
Publicado: (2025)
Automated planning with ontologies under coherence update semantics (Extended Version)
por: Borgwardt, Stefan, et al.
Publicado: (2025)
por: Borgwardt, Stefan, et al.
Publicado: (2025)
Optimal sensor deception in stochastic environments with partial observability to mislead a robot to a decoy goal
por: Rahmani, Hazhar, et al.
Publicado: (2025)
por: Rahmani, Hazhar, et al.
Publicado: (2025)
Safety Implications of Explainable Artificial Intelligence in End-to-End Autonomous Driving
por: Atakishiyev, Shahin, et al.
Publicado: (2024)
por: Atakishiyev, Shahin, et al.
Publicado: (2024)
Incorporating Explanations into Human-Machine Interfaces for Trust and Situation Awareness in Autonomous Vehicles
por: Atakishiyev, Shahin, et al.
Publicado: (2024)
por: Atakishiyev, Shahin, et al.
Publicado: (2024)
A generative foundation model for an all-in-one seismic processing framework
por: Cheng, Shijun, et al.
Publicado: (2025)
por: Cheng, Shijun, et al.
Publicado: (2025)
Graph Neural Networks vs Convolutional Neural Networks for Graph Domination Number Prediction
por: Davila, Randy, et al.
Publicado: (2025)
por: Davila, Randy, et al.
Publicado: (2025)
Principled Curriculum Learning using Parameter Continuation Methods
por: Pathak, Harsh Nilesh, et al.
Publicado: (2025)
por: Pathak, Harsh Nilesh, et al.
Publicado: (2025)
IMAGINE: An 8-to-1b 22nm FD-SOI Compute-In-Memory CNN Accelerator With an End-to-End Analog Charge-Based 0.15-8POPS/W Macro Featuring Distribution-Aware Data Reshaping
por: Kneip, Adrian, et al.
Publicado: (2024)
por: Kneip, Adrian, et al.
Publicado: (2024)
An unconditional distribution learning advantage with shallow quantum circuits
por: Pirnay, N., et al.
Publicado: (2024)
por: Pirnay, N., et al.
Publicado: (2024)
Experience-driven discovery of planning strategies
por: He, Ruiqi, et al.
Publicado: (2024)
por: He, Ruiqi, et al.
Publicado: (2024)
Synthesizing world models for bilevel planning
por: Ahmed, Zergham, et al.
Publicado: (2025)
por: Ahmed, Zergham, et al.
Publicado: (2025)
Getting SMARTER for Motion Planning in Autonomous Driving Systems
por: Alban, Montgomery, et al.
Publicado: (2025)
por: Alban, Montgomery, et al.
Publicado: (2025)
Ask before you Build: Rethinking AI-for-Good in Human Trafficking Interventions
por: Nair, Pratheeksha, et al.
Publicado: (2025)
por: Nair, Pratheeksha, et al.
Publicado: (2025)
LLM Robustness Leaderboard v1 --Technical report
por: Lefebvre, Pierre Peigné -, et al.
Publicado: (2025)
por: Lefebvre, Pierre Peigné -, et al.
Publicado: (2025)
Robust Fine-Tuning from Non-Robust Pretrained Models: Mitigating Suboptimal Transfer With Epsilon-Scheduling
por: Ngnawé, Jonas, et al.
Publicado: (2025)
por: Ngnawé, Jonas, et al.
Publicado: (2025)
CSM-H-R: A Context Modeling Framework in Supporting Reasoning Automation for Interoperable Intelligent Systems and Privacy Protection
por: Yue, Songhui, et al.
Publicado: (2023)
por: Yue, Songhui, et al.
Publicado: (2023)
Synthesis of timeline-based planning strategies avoiding determinization
por: Della Monica, Dario, et al.
Publicado: (2025)
por: Della Monica, Dario, et al.
Publicado: (2025)
ASKCOS: an open source software suite for synthesis planning
por: Tu, Zhengkai, et al.
Publicado: (2025)
por: Tu, Zhengkai, et al.
Publicado: (2025)
Individual differences in the cognitive mechanisms of planning strategy discovery
por: He, Ruiqi, et al.
Publicado: (2025)
por: He, Ruiqi, et al.
Publicado: (2025)
Estimating cognitive biases with attention-aware inverse planning
por: Banerjee, Sounak, et al.
Publicado: (2025)
por: Banerjee, Sounak, et al.
Publicado: (2025)
Developing Autonomous Robot-Mediated Behavior Coaching Sessions with Haru
por: Jelínek, Matouš, et al.
Publicado: (2024)
por: Jelínek, Matouš, et al.
Publicado: (2024)
In Reverie Together: Ten Years of Mathematical Discovery with a Machine Collaborator
por: Davila, Randy, et al.
Publicado: (2025)
por: Davila, Randy, et al.
Publicado: (2025)
Solo Connection: A Parameter Efficient Fine-Tuning Technique for Transformers
por: Pathak, Harsh Nilesh, et al.
Publicado: (2025)
por: Pathak, Harsh Nilesh, et al.
Publicado: (2025)
Proceedings of the First International Workshop on Next-Generation Language Models for Knowledge Representation and Reasoning (NeLaMKRR 2024)
por: Satoh, Ken, et al.
Publicado: (2024)
por: Satoh, Ken, et al.
Publicado: (2024)
Proceedings of the Second International Workshop on Next-Generation Language Models for Knowledge Representation and Reasoning (NeLaMKRR 2025)
por: Nguyen, Ha-Thanh, et al.
Publicado: (2025)
por: Nguyen, Ha-Thanh, et al.
Publicado: (2025)
Transformers perform adaptive partial pooling
por: Kapatsinski, Vsevolod
Publicado: (2026)
por: Kapatsinski, Vsevolod
Publicado: (2026)
Synthelite: Chemist-aligned and feasibility-aware synthesis planning with LLMs
por: Xuan-Vu, Nguyen, et al.
Publicado: (2025)
por: Xuan-Vu, Nguyen, et al.
Publicado: (2025)
Exploring the hierarchical structure of human plans via program generation
por: Correa, Carlos G., et al.
Publicado: (2023)
por: Correa, Carlos G., et al.
Publicado: (2023)
Ejemplares similares
-
An intelligent tutor for planning in large partially observable environments
por: Heindrich, Lovis, et al.
Publicado: (2023) -
LLM-as-a-Judge: Toward World Models for Slate Recommendation Systems
por: Bonin, Baptiste, et al.
Publicado: (2025) -
Propagating the prior from shallow to deep with a pre-trained velocity-model Generative Transformer network
por: Harsuko, Randy, et al.
Publicado: (2024) -
Nested-ReFT: Efficient Reinforcement Learning for Large Language Model Fine-Tuning via Off-Policy Rollouts
por: Heuillet, Maxime, et al.
Publicado: (2025) -
Artificial intelligence and machine learning generated conjectures with TxGraffiti
por: Davila, Randy
Publicado: (2024)