Coupled Distributional Random Expert Distillation for World Model Online Imitation Learning
Fuente:
arXiv
Saved in:
| Main Authors: | Li, Shangzhe, Huang, Zhiao, Su, Hao |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Reward-free World Models for Online Imitation Learning
by: Li, Shangzhe, et al.
Published: (2024)
by: Li, Shangzhe, et al.
Published: (2024)
Uncovering Capabilities of Model Pruning in Graph Contrastive Learning
by: Wu, Junran, et al.
Published: (2024)
by: Wu, Junran, et al.
Published: (2024)
DITTO: Offline Imitation Learning with World Models
by: DeMoss, Branton, et al.
Published: (2023)
by: DeMoss, Branton, et al.
Published: (2023)
Causal Imitation Learning under Expert-Observable and Expert-Unobservable Confounding
by: Shao, Daqian, et al.
Published: (2025)
by: Shao, Daqian, et al.
Published: (2025)
Language Model Distillation: A Temporal Difference Imitation Learning Perspective
by: Yu, Zishun, et al.
Published: (2025)
by: Yu, Zishun, et al.
Published: (2025)
Chain-of-Thought Predictive Control
by: Jia, Zhiwei, et al.
Published: (2023)
by: Jia, Zhiwei, et al.
Published: (2023)
Online Adaptation for Enhancing Imitation Learning Policies
by: Malato, Federico, et al.
Published: (2024)
by: Malato, Federico, et al.
Published: (2024)
Imitation Learning as Return Distribution Matching
by: Lazzati, Filippo, et al.
Published: (2025)
by: Lazzati, Filippo, et al.
Published: (2025)
Offline Imitation Learning with Model-based Reverse Augmentation
by: Shao, Jie-Jing, et al.
Published: (2024)
by: Shao, Jie-Jing, et al.
Published: (2024)
IDIL: Imitation Learning of Intent-Driven Expert Behavior
by: Seo, Sangwon, et al.
Published: (2024)
by: Seo, Sangwon, et al.
Published: (2024)
Using Non-Expert Data to Robustify Imitation Learning via Offline Reinforcement Learning
by: Huang, Kevin, et al.
Published: (2025)
by: Huang, Kevin, et al.
Published: (2025)
PoE-World: Compositional World Modeling with Products of Programmatic Experts
by: Piriyakulkij, Wasu Top, et al.
Published: (2025)
by: Piriyakulkij, Wasu Top, et al.
Published: (2025)
OLLIE: Imitation Learning from Offline Pretraining to Online Finetuning
by: Yue, Sheng, et al.
Published: (2024)
by: Yue, Sheng, et al.
Published: (2024)
Beyond-Expert Performance with Limited Demonstrations: Efficient Imitation Learning with Double Exploration
by: Zhao, Heyang, et al.
Published: (2025)
by: Zhao, Heyang, et al.
Published: (2025)
Imitation Learning for Multi-turn LM Agents via On-policy Expert Corrections
by: Lauffer, Niklas, et al.
Published: (2025)
by: Lauffer, Niklas, et al.
Published: (2025)
Distilled Thompson Sampling: Practical and Efficient Thompson Sampling via Imitation Learning
by: Namkoong, Hongseok, et al.
Published: (2020)
by: Namkoong, Hongseok, et al.
Published: (2020)
Self-evolved Imitation Learning in Simulated World
by: Ye, Yifan, et al.
Published: (2025)
by: Ye, Yifan, et al.
Published: (2025)
Learning What to Do and What Not To Do: Offline Imitation from Expert and Undesirable Demonstrations
by: Hoang, Huy, et al.
Published: (2025)
by: Hoang, Huy, et al.
Published: (2025)
MoDE: A Mixture-of-Experts Model with Mutual Distillation among the Experts
by: Xie, Zhitian, et al.
Published: (2024)
by: Xie, Zhitian, et al.
Published: (2024)
Stackelberg Coupling of Online Representation Learning and Reinforcement Learning
by: Martinez, Fernando, et al.
Published: (2025)
by: Martinez, Fernando, et al.
Published: (2025)
Learning for Long-Horizon Planning via Neuro-Symbolic Abductive Imitation
by: Shao, Jie-Jing, et al.
Published: (2024)
by: Shao, Jie-Jing, et al.
Published: (2024)
Balance Equation-based Distributionally Robust Offline Imitation Learning
by: Agrawal, Rishabh, et al.
Published: (2025)
by: Agrawal, Rishabh, et al.
Published: (2025)
Denoising-based Contractive Imitation Learning
by: Shen, Macheng, et al.
Published: (2025)
by: Shen, Macheng, et al.
Published: (2025)
Asymmetric On-Policy Distillation: Bridging Exploitation and Imitation at the Token Level
by: Jia, Nan, et al.
Published: (2026)
by: Jia, Nan, et al.
Published: (2026)
SENSOR: Imitate Third-Person Expert's Behaviors via Active Sensoring
by: Huang, Kaichen, et al.
Published: (2024)
by: Huang, Kaichen, et al.
Published: (2024)
Online Policy Distillation with Decision-Attention
by: Yu, Xinqiang, et al.
Published: (2024)
by: Yu, Xinqiang, et al.
Published: (2024)
Unveiling the Role of Expert Guidance: A Comparative Analysis of User-centered Imitation Learning and Traditional Reinforcement Learning
by: Gomaa, Amr, et al.
Published: (2024)
by: Gomaa, Amr, et al.
Published: (2024)
A Note on Hybrid Online Reinforcement and Imitation Learning for LLMs: Formulations and Algorithms
by: Li, Yingru, et al.
Published: (2025)
by: Li, Yingru, et al.
Published: (2025)
Learning Soft Driving Constraints from Vectorized Scene Embeddings while Imitating Expert Trajectories
by: Mobarakeh, Niloufar Saeidi, et al.
Published: (2024)
by: Mobarakeh, Niloufar Saeidi, et al.
Published: (2024)
A Recipe for Efficient Sim-to-Real Transfer in Manipulation with Online Imitation-Pretrained World Models
by: Wang, Yilin, et al.
Published: (2025)
by: Wang, Yilin, et al.
Published: (2025)
Continual Reinforcement Learning by Planning with Online World Models
by: Liu, Zichen, et al.
Published: (2025)
by: Liu, Zichen, et al.
Published: (2025)
Adversarial Imitation Learning via Boosting
by: Chang, Jonathan D., et al.
Published: (2024)
by: Chang, Jonathan D., et al.
Published: (2024)
Towards Generalisable Imitation Learning Through Conditioned Transition Estimation and Online Behaviour Alignment
by: Gavenski, Nathan, et al.
Published: (2026)
by: Gavenski, Nathan, et al.
Published: (2026)
Hierarchically Gated Experts for Efficient Online Continual Learning
by: Luong, Kevin, et al.
Published: (2024)
by: Luong, Kevin, et al.
Published: (2024)
Deep Clustering Survival Machines with Interpretable Expert Distributions
by: Hou, Bojian, et al.
Published: (2023)
by: Hou, Bojian, et al.
Published: (2023)
Sample-Efficient Expert Query Control in Active Imitation Learning via Conformal Prediction
by: Firouzkouhi, Arad, et al.
Published: (2025)
by: Firouzkouhi, Arad, et al.
Published: (2025)
Graph Knowledge Distillation to Mixture of Experts
by: Rumiantsev, Pavel, et al.
Published: (2024)
by: Rumiantsev, Pavel, et al.
Published: (2024)
Random Network Distillation Based Deep Reinforcement Learning for AGV Path Planning
by: Yin, Huilin, et al.
Published: (2024)
by: Yin, Huilin, et al.
Published: (2024)
Rethinking Momentum Knowledge Distillation in Online Continual Learning
by: Michel, Nicolas, et al.
Published: (2023)
by: Michel, Nicolas, et al.
Published: (2023)
Composition of Memory Experts for Diffusion World Models
by: Stapf, Sebastian, et al.
Published: (2026)
by: Stapf, Sebastian, et al.
Published: (2026)
Similar Items
-
Reward-free World Models for Online Imitation Learning
by: Li, Shangzhe, et al.
Published: (2024) -
Uncovering Capabilities of Model Pruning in Graph Contrastive Learning
by: Wu, Junran, et al.
Published: (2024) -
DITTO: Offline Imitation Learning with World Models
by: DeMoss, Branton, et al.
Published: (2023) -
Causal Imitation Learning under Expert-Observable and Expert-Unobservable Confounding
by: Shao, Daqian, et al.
Published: (2025) -
Language Model Distillation: A Temporal Difference Imitation Learning Perspective
by: Yu, Zishun, et al.
Published: (2025)