Saved in:
| Main Authors: | Lázaro-Gredilla, Miguel, Ku, Li Yang, Murphy, Kevin P., George, Dileep |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2406.17863 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Learning Cognitive Maps from Transformer Representations for Efficient Planning in Partially Observed Environments
by: Dedieu, Antoine, et al.
Published: (2024)
by: Dedieu, Antoine, et al.
Published: (2024)
Improving Transformer World Models for Data-Efficient RL
by: Dedieu, Antoine, et al.
Published: (2025)
by: Dedieu, Antoine, et al.
Published: (2025)
Diffusion Model Predictive Control
by: Zhou, Guangyao, et al.
Published: (2024)
by: Zhou, Guangyao, et al.
Published: (2024)
Model Predictive Simulation Using Structured Graphical Models and Transformers
by: Lou, Xinghua, et al.
Published: (2024)
by: Lou, Xinghua, et al.
Published: (2024)
AutoHarness: improving LLM agents by automatically synthesizing a code harness
by: Lou, Xinghua, et al.
Published: (2026)
by: Lou, Xinghua, et al.
Published: (2026)
Dynamic planning in hierarchical active inference
by: Priorelli, Matteo, et al.
Published: (2024)
by: Priorelli, Matteo, et al.
Published: (2024)
Reinforcement Learning: An Overview
by: Murphy, Kevin
Published: (2024)
by: Murphy, Kevin
Published: (2024)
On Predictive planning and counterfactual learning in active inference
by: Paul, Aswin, et al.
Published: (2024)
by: Paul, Aswin, et al.
Published: (2024)
Joint Learning of Hierarchical Neural Options and Abstract World Model
by: Piriyakulkij, Wasu Top, et al.
Published: (2026)
by: Piriyakulkij, Wasu Top, et al.
Published: (2026)
What's the plan? Metrics for implicit planning in LLMs and their application to rhyme generation and question answering
by: Maar, Jim, et al.
Published: (2026)
by: Maar, Jim, et al.
Published: (2026)
Federated Ensemble-Directed Offline Reinforcement Learning
by: Rengarajan, Desik, et al.
Published: (2023)
by: Rengarajan, Desik, et al.
Published: (2023)
Retro-fallback: retrosynthetic planning in an uncertain world
by: Tripp, Austin, et al.
Published: (2023)
by: Tripp, Austin, et al.
Published: (2023)
From Generalist to Specialist Representation
by: Zheng, Yujia, et al.
Published: (2026)
by: Zheng, Yujia, et al.
Published: (2026)
Dual policy as self-model for planning
by: Yoo, Jaesung, et al.
Published: (2023)
by: Yoo, Jaesung, et al.
Published: (2023)
Diffusion model for relational inference
by: Zheng, Shuhan, et al.
Published: (2024)
by: Zheng, Shuhan, et al.
Published: (2024)
Robust LLM Alignment via Distributionally Robust Direct Preference Optimization
by: Xu, Zaiyan, et al.
Published: (2025)
by: Xu, Zaiyan, et al.
Published: (2025)
What makes Reasoning Models Different? Follow the Reasoning Leader for Efficient Decoding
by: Li, Ming, et al.
Published: (2025)
by: Li, Ming, et al.
Published: (2025)
Optimistic World Models: Efficient Exploration in Model-Based Deep Reinforcement Learning
by: Mete, Akshay, et al.
Published: (2026)
by: Mete, Akshay, et al.
Published: (2026)
Time-Varying Constraint-Aware Reinforcement Learning for Energy Storage Control
by: Jeong, Jaeik, et al.
Published: (2024)
by: Jeong, Jaeik, et al.
Published: (2024)
Conformalized Neural Networks for Federated Uncertainty Quantification under Dual Heterogeneity
by: Nguyen, Quang-Huy, et al.
Published: (2026)
by: Nguyen, Quang-Huy, et al.
Published: (2026)
All-in-one simulation-based inference
by: Gloeckler, Manuel, et al.
Published: (2024)
by: Gloeckler, Manuel, et al.
Published: (2024)
Explainability as statistical inference
by: Senetaire, Hugo Henri Joseph, et al.
Published: (2022)
by: Senetaire, Hugo Henri Joseph, et al.
Published: (2022)
DMC-VB: A Benchmark for Representation Learning for Control with Visual Distractors
by: Ortiz, Joseph, et al.
Published: (2024)
by: Ortiz, Joseph, et al.
Published: (2024)
EM Distillation for One-step Diffusion Models
by: Xie, Sirui, et al.
Published: (2024)
by: Xie, Sirui, et al.
Published: (2024)
Does "Do Differentiable Simulators Give Better Policy Gradients?'' Give Better Policy Gradients?
by: Onoda, Ku, et al.
Published: (2026)
by: Onoda, Ku, et al.
Published: (2026)
Delving into Instance-Dependent Label Noise in Graph Data: A Comprehensive Study and Benchmark
by: Kim, Suyeon, et al.
Published: (2025)
by: Kim, Suyeon, et al.
Published: (2025)
Structured Reinforcement Learning for Media Streaming at the Wireless Edge
by: Bura, Archana, et al.
Published: (2024)
by: Bura, Archana, et al.
Published: (2024)
Rotated Runtime Smooth: Training-Free Activation Smoother for accurate INT4 inference
by: Yi, Ke, et al.
Published: (2024)
by: Yi, Ke, et al.
Published: (2024)
Curiosity-driven RL for symbolic equation solving
by: O'Keeffe, Kevin P.
Published: (2025)
by: O'Keeffe, Kevin P.
Published: (2025)
Missing Data Multiple Imputation for Tabular Q-Learning in Online RL
by: Chasalow, Kyla, et al.
Published: (2025)
by: Chasalow, Kyla, et al.
Published: (2025)
Learning to refine domain knowledge for biological network inference
by: Li, Peiwen, et al.
Published: (2024)
by: Li, Peiwen, et al.
Published: (2024)
Towards a Mechanistic Explanation of Diffusion Model Generalization
by: Niedoba, Matthew, et al.
Published: (2024)
by: Niedoba, Matthew, et al.
Published: (2024)
Using large language models for embodied planning introduces systematic safety risks
by: Zhang, Tao, et al.
Published: (2026)
by: Zhang, Tao, et al.
Published: (2026)
Simulating the Unseen: Crash Prediction Must Learn from What Did Not Happen
by: Li, Zihao, et al.
Published: (2025)
by: Li, Zihao, et al.
Published: (2025)
Cross-Entropy Optimization for Hyperparameter Optimization in Stochastic Gradient-based Approaches to Train Deep Neural Networks
by: Li, Kevin, et al.
Published: (2024)
by: Li, Kevin, et al.
Published: (2024)
Reinforcement Learning for Diffusion LLMs with Entropy-Guided Step Selection and Stepwise Advantages
by: Kunde, Vishnu Teja, et al.
Published: (2026)
by: Kunde, Vishnu Teja, et al.
Published: (2026)
Improving planning and MBRL with temporally-extended actions
by: Chatterjee, Palash, et al.
Published: (2025)
by: Chatterjee, Palash, et al.
Published: (2025)
Electrostatics-based particle sampling and approximate inference
by: Huang, Yongchao
Published: (2024)
by: Huang, Yongchao
Published: (2024)
Post-detection inference for sequential changepoint localization
by: Saha, Aytijhya, et al.
Published: (2025)
by: Saha, Aytijhya, et al.
Published: (2025)
LLM enhanced graph inference for long-term disease progression modelling
by: He, Tiantian, et al.
Published: (2025)
by: He, Tiantian, et al.
Published: (2025)
Similar Items
-
Learning Cognitive Maps from Transformer Representations for Efficient Planning in Partially Observed Environments
by: Dedieu, Antoine, et al.
Published: (2024) -
Improving Transformer World Models for Data-Efficient RL
by: Dedieu, Antoine, et al.
Published: (2025) -
Diffusion Model Predictive Control
by: Zhou, Guangyao, et al.
Published: (2024) -
Model Predictive Simulation Using Structured Graphical Models and Transformers
by: Lou, Xinghua, et al.
Published: (2024) -
AutoHarness: improving LLM agents by automatically synthesizing a code harness
by: Lou, Xinghua, et al.
Published: (2026)