What type of inference is planning?
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Lázaro-Gredilla, Miguel, Ku, Li Yang, Murphy, Kevin P., George, Dileep |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Learning Cognitive Maps from Transformer Representations for Efficient Planning in Partially Observed Environments
von: Dedieu, Antoine, et al.
Veröffentlicht: (2024)
von: Dedieu, Antoine, et al.
Veröffentlicht: (2024)
Improving Transformer World Models for Data-Efficient RL
von: Dedieu, Antoine, et al.
Veröffentlicht: (2025)
von: Dedieu, Antoine, et al.
Veröffentlicht: (2025)
Diffusion Model Predictive Control
von: Zhou, Guangyao, et al.
Veröffentlicht: (2024)
von: Zhou, Guangyao, et al.
Veröffentlicht: (2024)
Reinforcement Learning: An Overview
von: Murphy, Kevin
Veröffentlicht: (2024)
von: Murphy, Kevin
Veröffentlicht: (2024)
Dynamic planning in hierarchical active inference
von: Priorelli, Matteo, et al.
Veröffentlicht: (2024)
von: Priorelli, Matteo, et al.
Veröffentlicht: (2024)
On Predictive planning and counterfactual learning in active inference
von: Paul, Aswin, et al.
Veröffentlicht: (2024)
von: Paul, Aswin, et al.
Veröffentlicht: (2024)
Model Predictive Simulation Using Structured Graphical Models and Transformers
von: Lou, Xinghua, et al.
Veröffentlicht: (2024)
von: Lou, Xinghua, et al.
Veröffentlicht: (2024)
AutoHarness: improving LLM agents by automatically synthesizing a code harness
von: Lou, Xinghua, et al.
Veröffentlicht: (2026)
von: Lou, Xinghua, et al.
Veröffentlicht: (2026)
Joint Learning of Hierarchical Neural Options and Abstract World Model
von: Piriyakulkij, Wasu Top, et al.
Veröffentlicht: (2026)
von: Piriyakulkij, Wasu Top, et al.
Veröffentlicht: (2026)
Federated Ensemble-Directed Offline Reinforcement Learning
von: Rengarajan, Desik, et al.
Veröffentlicht: (2023)
von: Rengarajan, Desik, et al.
Veröffentlicht: (2023)
What's the plan? Metrics for implicit planning in LLMs and their application to rhyme generation and question answering
von: Maar, Jim, et al.
Veröffentlicht: (2026)
von: Maar, Jim, et al.
Veröffentlicht: (2026)
Diffusion model for relational inference
von: Zheng, Shuhan, et al.
Veröffentlicht: (2024)
von: Zheng, Shuhan, et al.
Veröffentlicht: (2024)
Retro-fallback: retrosynthetic planning in an uncertain world
von: Tripp, Austin, et al.
Veröffentlicht: (2023)
von: Tripp, Austin, et al.
Veröffentlicht: (2023)
From Generalist to Specialist Representation
von: Zheng, Yujia, et al.
Veröffentlicht: (2026)
von: Zheng, Yujia, et al.
Veröffentlicht: (2026)
Dual policy as self-model for planning
von: Yoo, Jaesung, et al.
Veröffentlicht: (2023)
von: Yoo, Jaesung, et al.
Veröffentlicht: (2023)
Robust LLM Alignment via Distributionally Robust Direct Preference Optimization
von: Xu, Zaiyan, et al.
Veröffentlicht: (2025)
von: Xu, Zaiyan, et al.
Veröffentlicht: (2025)
What makes Reasoning Models Different? Follow the Reasoning Leader for Efficient Decoding
von: Li, Ming, et al.
Veröffentlicht: (2025)
von: Li, Ming, et al.
Veröffentlicht: (2025)
All-in-one simulation-based inference
von: Gloeckler, Manuel, et al.
Veröffentlicht: (2024)
von: Gloeckler, Manuel, et al.
Veröffentlicht: (2024)
Time-Varying Constraint-Aware Reinforcement Learning for Energy Storage Control
von: Jeong, Jaeik, et al.
Veröffentlicht: (2024)
von: Jeong, Jaeik, et al.
Veröffentlicht: (2024)
Conformalized Neural Networks for Federated Uncertainty Quantification under Dual Heterogeneity
von: Nguyen, Quang-Huy, et al.
Veröffentlicht: (2026)
von: Nguyen, Quang-Huy, et al.
Veröffentlicht: (2026)
Explainability as statistical inference
von: Senetaire, Hugo Henri Joseph, et al.
Veröffentlicht: (2022)
von: Senetaire, Hugo Henri Joseph, et al.
Veröffentlicht: (2022)
Optimistic World Models: Efficient Exploration in Model-Based Deep Reinforcement Learning
von: Mete, Akshay, et al.
Veröffentlicht: (2026)
von: Mete, Akshay, et al.
Veröffentlicht: (2026)
Rotated Runtime Smooth: Training-Free Activation Smoother for accurate INT4 inference
von: Yi, Ke, et al.
Veröffentlicht: (2024)
von: Yi, Ke, et al.
Veröffentlicht: (2024)
Curiosity-driven RL for symbolic equation solving
von: O'Keeffe, Kevin P.
Veröffentlicht: (2025)
von: O'Keeffe, Kevin P.
Veröffentlicht: (2025)
Cross-Entropy Optimization for Hyperparameter Optimization in Stochastic Gradient-based Approaches to Train Deep Neural Networks
von: Li, Kevin, et al.
Veröffentlicht: (2024)
von: Li, Kevin, et al.
Veröffentlicht: (2024)
Simulating the Unseen: Crash Prediction Must Learn from What Did Not Happen
von: Li, Zihao, et al.
Veröffentlicht: (2025)
von: Li, Zihao, et al.
Veröffentlicht: (2025)
EM Distillation for One-step Diffusion Models
von: Xie, Sirui, et al.
Veröffentlicht: (2024)
von: Xie, Sirui, et al.
Veröffentlicht: (2024)
Missing Data Multiple Imputation for Tabular Q-Learning in Online RL
von: Chasalow, Kyla, et al.
Veröffentlicht: (2025)
von: Chasalow, Kyla, et al.
Veröffentlicht: (2025)
Delving into Instance-Dependent Label Noise in Graph Data: A Comprehensive Study and Benchmark
von: Kim, Suyeon, et al.
Veröffentlicht: (2025)
von: Kim, Suyeon, et al.
Veröffentlicht: (2025)
LLM enhanced graph inference for long-term disease progression modelling
von: He, Tiantian, et al.
Veröffentlicht: (2025)
von: He, Tiantian, et al.
Veröffentlicht: (2025)
Diffusion Tree Sampling: Scalable inference-time alignment of diffusion models
von: Jain, Vineet, et al.
Veröffentlicht: (2025)
von: Jain, Vineet, et al.
Veröffentlicht: (2025)
Learning to refine domain knowledge for biological network inference
von: Li, Peiwen, et al.
Veröffentlicht: (2024)
von: Li, Peiwen, et al.
Veröffentlicht: (2024)
Does "Do Differentiable Simulators Give Better Policy Gradients?'' Give Better Policy Gradients?
von: Onoda, Ku, et al.
Veröffentlicht: (2026)
von: Onoda, Ku, et al.
Veröffentlicht: (2026)
Remedying uncertainty representations in visual inference through Explaining-Away Variational Autoencoders
von: Catoni, Josefina, et al.
Veröffentlicht: (2024)
von: Catoni, Josefina, et al.
Veröffentlicht: (2024)
Electrostatics-based particle sampling and approximate inference
von: Huang, Yongchao
Veröffentlicht: (2024)
von: Huang, Yongchao
Veröffentlicht: (2024)
Post-detection inference for sequential changepoint localization
von: Saha, Aytijhya, et al.
Veröffentlicht: (2025)
von: Saha, Aytijhya, et al.
Veröffentlicht: (2025)
Beyond What to Select: A Plug-and-play Oscillatory Data-Volume Scheduling for Efficient Model Training
von: Yang, Suorong, et al.
Veröffentlicht: (2026)
von: Yang, Suorong, et al.
Veröffentlicht: (2026)
Using large language models for embodied planning introduces systematic safety risks
von: Zhang, Tao, et al.
Veröffentlicht: (2026)
von: Zhang, Tao, et al.
Veröffentlicht: (2026)
What Makes a Good Diffusion Planner for Decision Making?
von: Lu, Haofei, et al.
Veröffentlicht: (2025)
von: Lu, Haofei, et al.
Veröffentlicht: (2025)
GeoUni: A Unified Model for Generating Geometry Diagrams, Problems and Problem Solutions
von: Cheng, Jo-Ku, et al.
Veröffentlicht: (2025)
von: Cheng, Jo-Ku, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Learning Cognitive Maps from Transformer Representations for Efficient Planning in Partially Observed Environments
von: Dedieu, Antoine, et al.
Veröffentlicht: (2024) -
Improving Transformer World Models for Data-Efficient RL
von: Dedieu, Antoine, et al.
Veröffentlicht: (2025) -
Diffusion Model Predictive Control
von: Zhou, Guangyao, et al.
Veröffentlicht: (2024) -
Reinforcement Learning: An Overview
von: Murphy, Kevin
Veröffentlicht: (2024) -
Dynamic planning in hierarchical active inference
von: Priorelli, Matteo, et al.
Veröffentlicht: (2024)