Bootstrap Off-policy with World Model
Fuente:
arXiv
Saved in:
| Main Authors: | Zhan, Guojian, Wang, Likun, Zhang, Xiangteng, Gao, Jiaxin, Tomizuka, Masayoshi, Li, Shengbo Eben |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Mind Your Entropy: From Maximum Entropy to Trajectory Entropy-Constrained RL
by: Zhan, Guojian, et al.
Published: (2025)
by: Zhan, Guojian, et al.
Published: (2025)
Off-policy Reinforcement Learning with Model-based Exploration Augmentation
by: Wang, Likun, et al.
Published: (2025)
by: Wang, Likun, et al.
Published: (2025)
Distributional Soft Actor-Critic with Harmonic Gradient for Safe and Efficient Autonomous Driving in Multi-lane Scenarios
by: Zhang, Feihong, et al.
Published: (2025)
by: Zhang, Feihong, et al.
Published: (2025)
Mean Flow Policy with Instantaneous Velocity Constraint for One-step Action Generation
by: Zhan, Guojian, et al.
Published: (2026)
by: Zhan, Guojian, et al.
Published: (2026)
Safe Offline Reinforcement Learning with Feasibility-Guided Diffusion Model
by: Zheng, Yinan, et al.
Published: (2024)
by: Zheng, Yinan, et al.
Published: (2024)
LanguageMPC: Large Language Models as Decision Makers for Autonomous Driving
by: Sha, Hao, et al.
Published: (2023)
by: Sha, Hao, et al.
Published: (2023)
Grounded Relational Inference: Domain Knowledge Driven Explainable Autonomous Driving
by: Tang, Chen, et al.
Published: (2021)
by: Tang, Chen, et al.
Published: (2021)
Jump-Start Reinforcement Learning with Self-Evolving Priors for Extreme Monopedal Locomotion
by: Zheng, Ziang, et al.
Published: (2025)
by: Zheng, Ziang, et al.
Published: (2025)
Transferable Latent-to-Latent Locomotion Policy for Efficient and Versatile Motion Control of Diverse Legged Robots
by: Zheng, Ziang, et al.
Published: (2025)
by: Zheng, Ziang, et al.
Published: (2025)
DADP: Domain Adaptive Diffusion Policy
by: Wang, Pengcheng, et al.
Published: (2026)
by: Wang, Pengcheng, et al.
Published: (2026)
Diffusion-Based Planning for Autonomous Driving with Flexible Guidance
by: Zheng, Yinan, et al.
Published: (2025)
by: Zheng, Yinan, et al.
Published: (2025)
Bootstrapped Model Predictive Control
by: Wang, Yuhang, et al.
Published: (2025)
by: Wang, Yuhang, et al.
Published: (2025)
Conformal Symplectic Optimization for Stable Reinforcement Learning
by: Lyu, Yao, et al.
Published: (2024)
by: Lyu, Yao, et al.
Published: (2024)
Spatio-Temporal Graph Dual-Attention Network for Multi-Agent Prediction and Tracking
by: Li, Jiachen, et al.
Published: (2021)
by: Li, Jiachen, et al.
Published: (2021)
Adapting Critic Match Loss Landscape Visualization to Off-policy Reinforcement Learning
by: Liu, Jingyi, et al.
Published: (2026)
by: Liu, Jingyi, et al.
Published: (2026)
VER: Vision Expert Transformer for Robot Learning via Foundation Distillation and Dynamic Routing
by: Wang, Yixiao, et al.
Published: (2025)
by: Wang, Yixiao, et al.
Published: (2025)
FOUNDER: Grounding Foundation Models in World Models for Open-Ended Embodied Decision Making
by: Wang, Yucen, et al.
Published: (2025)
by: Wang, Yucen, et al.
Published: (2025)
Towards Generalizable and Interpretable Motion Prediction: A Deep Variational Bayes Approach
by: Lu, Juanwu, et al.
Published: (2024)
by: Lu, Juanwu, et al.
Published: (2024)
SPACeR: Self-Play Anchoring with Centralized Reference Models
by: Chang, Wei-Jer, et al.
Published: (2025)
by: Chang, Wei-Jer, et al.
Published: (2025)
MATRIX: Multi-Agent Trajectory Generation with Diverse Contexts
by: Xu, Zhuo, et al.
Published: (2024)
by: Xu, Zhuo, et al.
Published: (2024)
Flattening Hierarchies with Policy Bootstrapping
by: Zhou, John L., et al.
Published: (2025)
by: Zhou, John L., et al.
Published: (2025)
Zeroth-Order Actor-Critic: An Evolutionary Framework for Sequential Decision Problems
by: Lei, Yuheng, et al.
Published: (2022)
by: Lei, Yuheng, et al.
Published: (2022)
Enhanced DACER Algorithm with High Diffusion Efficiency
by: Wang, Yinuo, et al.
Published: (2025)
by: Wang, Yinuo, et al.
Published: (2025)
Maximizing Alignment with Minimal Feedback: Efficiently Learning Rewards for Visuomotor Robot Policy Alignment
by: Tian, Ran, et al.
Published: (2024)
by: Tian, Ran, et al.
Published: (2024)
Canonical Form of Datatic Description in Control Systems
by: Zhan, Guojian, et al.
Published: (2024)
by: Zhan, Guojian, et al.
Published: (2024)
FLIP: Flow-Centric Generative Planning as General-Purpose Manipulation World Model
by: Gao, Chongkai, et al.
Published: (2024)
by: Gao, Chongkai, et al.
Published: (2024)
MEReQ: Max-Ent Residual-Q Inverse RL for Sample-Efficient Alignment from Intervention
by: Chen, Yuxin, et al.
Published: (2024)
by: Chen, Yuxin, et al.
Published: (2024)
World Models for Autonomous Driving: An Initial Survey
by: Guan, Yanchen, et al.
Published: (2024)
by: Guan, Yanchen, et al.
Published: (2024)
MOBODY: Model Based Off-Dynamics Offline Reinforcement Learning
by: Guo, Yihong, et al.
Published: (2025)
by: Guo, Yihong, et al.
Published: (2025)
From Imitation to Exploration: End-to-end Autonomous Driving based on World Model
by: Li, Yueyuan, et al.
Published: (2024)
by: Li, Yueyuan, et al.
Published: (2024)
SAFE-SIM: Safety-Critical Closed-Loop Traffic Simulation with Diffusion-Controllable Adversaries
by: Chang, Wei-Jer, et al.
Published: (2023)
by: Chang, Wei-Jer, et al.
Published: (2023)
Real-Time Generative Policy via Langevin-Guided Flow Matching for Autonomous Driving
by: Zhu, Tianze, et al.
Published: (2026)
by: Zhu, Tianze, et al.
Published: (2026)
BLAZER: Bootstrapping LLM-based Manipulation Agents with Zero-Shot Data Generation
by: Das, Rocktim Jyoti, et al.
Published: (2025)
by: Das, Rocktim Jyoti, et al.
Published: (2025)
Return Augmented Decision Transformer for Off-Dynamics Reinforcement Learning
by: Wang, Ruhan, et al.
Published: (2024)
by: Wang, Ruhan, et al.
Published: (2024)
Semantic World Models
by: Berg, Jacob, et al.
Published: (2025)
by: Berg, Jacob, et al.
Published: (2025)
World Action Verifier: Self-Improving World Models via Forward-Inverse Asymmetry
by: Liu, Yuejiang, et al.
Published: (2026)
by: Liu, Yuejiang, et al.
Published: (2026)
Sample Complexity of Distributionally Robust Off-Dynamics Reinforcement Learning with Online Interaction
by: He, Yiting, et al.
Published: (2025)
by: He, Yiting, et al.
Published: (2025)
TADPO: Reinforcement Learning Goes Off-road
by: Wu, Zhouchonghao, et al.
Published: (2026)
by: Wu, Zhouchonghao, et al.
Published: (2026)
Off-Dynamics Reinforcement Learning via Domain Adaptation and Reward Augmented Imitation
by: Guo, Yihong, et al.
Published: (2024)
by: Guo, Yihong, et al.
Published: (2024)
Generative World Modelling for Humanoids: 1X World Model Challenge Technical Report
by: Mereu, Riccardo, et al.
Published: (2025)
by: Mereu, Riccardo, et al.
Published: (2025)
Similar Items
-
Mind Your Entropy: From Maximum Entropy to Trajectory Entropy-Constrained RL
by: Zhan, Guojian, et al.
Published: (2025) -
Off-policy Reinforcement Learning with Model-based Exploration Augmentation
by: Wang, Likun, et al.
Published: (2025) -
Distributional Soft Actor-Critic with Harmonic Gradient for Safe and Efficient Autonomous Driving in Multi-lane Scenarios
by: Zhang, Feihong, et al.
Published: (2025) -
Mean Flow Policy with Instantaneous Velocity Constraint for One-step Action Generation
by: Zhan, Guojian, et al.
Published: (2026) -
Safe Offline Reinforcement Learning with Feasibility-Guided Diffusion Model
by: Zheng, Yinan, et al.
Published: (2024)