DITTO: Offline Imitation Learning with World Models
Fuente:
arXiv
Guardado en:
| Autores principales: | DeMoss, Branton, Duckworth, Paul, Foerster, Jakob, Hawes, Nick, Posner, Ingmar |
|---|---|
| Formato: | Preprint |
| Publicado: |
2023
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
The Complexity Dynamics of Grokking
por: DeMoss, Branton, et al.
Publicado: (2024)
por: DeMoss, Branton, et al.
Publicado: (2024)
LUMOS: Language-Conditioned Imitation Learning with World Models
por: Nematollahi, Iman, et al.
Publicado: (2025)
por: Nematollahi, Iman, et al.
Publicado: (2025)
World Models via Policy-Guided Trajectory Diffusion
por: Rigter, Marc, et al.
Publicado: (2023)
por: Rigter, Marc, et al.
Publicado: (2023)
The Edge-of-Reach Problem in Offline Model-Based Reinforcement Learning
por: Sims, Anya, et al.
Publicado: (2024)
por: Sims, Anya, et al.
Publicado: (2024)
A Model-Based Solution to the Offline Multi-Agent Reinforcement Learning Coordination Problem
por: Barde, Paul, et al.
Publicado: (2023)
por: Barde, Paul, et al.
Publicado: (2023)
No Regrets: Investigating and Improving Regret Approximations for Curriculum Discovery
por: Rutherford, Alexander, et al.
Publicado: (2024)
por: Rutherford, Alexander, et al.
Publicado: (2024)
Offline Adaptation of Quadruped Locomotion using Diffusion Models
por: O'Mahoney, Reece, et al.
Publicado: (2024)
por: O'Mahoney, Reece, et al.
Publicado: (2024)
Abstraction for Offline Goal-Conditioned Reinforcement Learning
por: Wibault, Clarisse, et al.
Publicado: (2026)
por: Wibault, Clarisse, et al.
Publicado: (2026)
A Clean Slate for Offline Reinforcement Learning
por: Jackson, Matthew Thomas, et al.
Publicado: (2025)
por: Jackson, Matthew Thomas, et al.
Publicado: (2025)
Offline Imitation Learning with Model-based Reverse Augmentation
por: Shao, Jie-Jing, et al.
Publicado: (2024)
por: Shao, Jie-Jing, et al.
Publicado: (2024)
SOReL and TOReL: Two Methods for Fully Offline Reinforcement Learning
por: Fellows, Mattie, et al.
Publicado: (2025)
por: Fellows, Mattie, et al.
Publicado: (2025)
How to Leverage Diverse Demonstrations in Offline Imitation Learning
por: Yue, Sheng, et al.
Publicado: (2024)
por: Yue, Sheng, et al.
Publicado: (2024)
XQC: Well-conditioned Optimization Accelerates Deep Reinforcement Learning
por: Palenicek, Daniel, et al.
Publicado: (2025)
por: Palenicek, Daniel, et al.
Publicado: (2025)
Select to Perfect: Imitating desired behavior from large multi-agent data
por: Franzmeyer, Tim, et al.
Publicado: (2024)
por: Franzmeyer, Tim, et al.
Publicado: (2024)
OLLIE: Imitation Learning from Offline Pretraining to Online Finetuning
por: Yue, Sheng, et al.
Publicado: (2024)
por: Yue, Sheng, et al.
Publicado: (2024)
Balance Equation-based Distributionally Robust Offline Imitation Learning
por: Agrawal, Rishabh, et al.
Publicado: (2025)
por: Agrawal, Rishabh, et al.
Publicado: (2025)
SPRINQL: Sub-optimal Demonstrations driven Offline Imitation Learning
por: Hoang, Huy, et al.
Publicado: (2024)
por: Hoang, Huy, et al.
Publicado: (2024)
Offline Imitation Learning Through Graph Search and Retrieval
por: Yin, Zhao-Heng, et al.
Publicado: (2024)
por: Yin, Zhao-Heng, et al.
Publicado: (2024)
SEABO: A Simple Search-Based Method for Offline Imitation Learning
por: Lyu, Jiafei, et al.
Publicado: (2024)
por: Lyu, Jiafei, et al.
Publicado: (2024)
Align Your Intents: Offline Imitation Learning via Optimal Transport
por: Bobrin, Maksim, et al.
Publicado: (2024)
por: Bobrin, Maksim, et al.
Publicado: (2024)
JaxWildfire: A GPU-Accelerated Wildfire Simulator for Reinforcement Learning
por: Çakır, Ufuk, et al.
Publicado: (2025)
por: Çakır, Ufuk, et al.
Publicado: (2025)
Vision-Language-Action Models for Robotics: A Review Towards Real-World Applications
por: Kawaharazuka, Kento, et al.
Publicado: (2025)
por: Kawaharazuka, Kento, et al.
Publicado: (2025)
DITTO-2: Distilled Diffusion Inference-Time T-Optimization for Music Generation
por: Novack, Zachary, et al.
Publicado: (2024)
por: Novack, Zachary, et al.
Publicado: (2024)
Learning Multi-Agent Communication with Contrastive Learning
por: Lo, Yat Long, et al.
Publicado: (2023)
por: Lo, Yat Long, et al.
Publicado: (2023)
Markov Balance Satisfaction Improves Performance in Strictly Batch Offline Imitation Learning
por: Agrawal, Rishabh, et al.
Publicado: (2024)
por: Agrawal, Rishabh, et al.
Publicado: (2024)
JaxUED: A simple and useable UED library in Jax
por: Coward, Samuel, et al.
Publicado: (2024)
por: Coward, Samuel, et al.
Publicado: (2024)
Offline Diversity Maximization Under Imitation Constraints
por: Vlastelica, Marin, et al.
Publicado: (2023)
por: Vlastelica, Marin, et al.
Publicado: (2023)
A Dual Approach to Imitation Learning from Observations with Offline Datasets
por: Sikchi, Harshit, et al.
Publicado: (2024)
por: Sikchi, Harshit, et al.
Publicado: (2024)
Mirror Learning: A Unifying Framework of Policy Optimisation
por: Kuba, Jakub Grudzien, et al.
Publicado: (2022)
por: Kuba, Jakub Grudzien, et al.
Publicado: (2022)
Return Capping: Sample-Efficient CVaR Policy Gradient Optimisation
por: Mead, Harry, et al.
Publicado: (2025)
por: Mead, Harry, et al.
Publicado: (2025)
Monte Carlo Tree Search with Boltzmann Exploration
por: Painter, Michael, et al.
Publicado: (2024)
por: Painter, Michael, et al.
Publicado: (2024)
PROF: An LLM-based Reward Code Preference Optimization Framework for Offline Imitation Learning
por: Sun, Shengjie, et al.
Publicado: (2025)
por: Sun, Shengjie, et al.
Publicado: (2025)
SafeMIL: Learning Offline Safe Imitation Policy from Non-Preferred Trajectories
por: Burnwal, Returaj, et al.
Publicado: (2025)
por: Burnwal, Returaj, et al.
Publicado: (2025)
Learning What to Do and What Not To Do: Offline Imitation from Expert and Undesirable Demonstrations
por: Hoang, Huy, et al.
Publicado: (2025)
por: Hoang, Huy, et al.
Publicado: (2025)
When Dynamics Shift, Robust Task Inference Wins: Offline Imitation Learning with Behavior Foundation Models Revisited
por: Agrawal, Rishabh, et al.
Publicado: (2026)
por: Agrawal, Rishabh, et al.
Publicado: (2026)
Intrinsically Interpretable Attention via Sparse Post-Training
por: Draye, Florent, et al.
Publicado: (2025)
por: Draye, Florent, et al.
Publicado: (2025)
Tackling GNARLy Problems: Graph Neural Algorithmic Reasoning Reimagined through Reinforcement Learning
por: Schutz, Alex, et al.
Publicado: (2025)
por: Schutz, Alex, et al.
Publicado: (2025)
Using Non-Expert Data to Robustify Imitation Learning via Offline Reinforcement Learning
por: Huang, Kevin, et al.
Publicado: (2025)
por: Huang, Kevin, et al.
Publicado: (2025)
DITTO: Diffusion Inference-Time T-Optimization for Music Generation
por: Novack, Zachary, et al.
Publicado: (2024)
por: Novack, Zachary, et al.
Publicado: (2024)
Grasp-MPC: Closed-Loop Visual Grasping via Value-Guided Model Predictive Control
por: Yamada, Jun, et al.
Publicado: (2025)
por: Yamada, Jun, et al.
Publicado: (2025)
Ejemplares similares
-
The Complexity Dynamics of Grokking
por: DeMoss, Branton, et al.
Publicado: (2024) -
LUMOS: Language-Conditioned Imitation Learning with World Models
por: Nematollahi, Iman, et al.
Publicado: (2025) -
World Models via Policy-Guided Trajectory Diffusion
por: Rigter, Marc, et al.
Publicado: (2023) -
The Edge-of-Reach Problem in Offline Model-Based Reinforcement Learning
por: Sims, Anya, et al.
Publicado: (2024) -
A Model-Based Solution to the Offline Multi-Agent Reinforcement Learning Coordination Problem
por: Barde, Paul, et al.
Publicado: (2023)