Performance Asymmetry in Model-Based Reinforcement Learning
Fuente:
arXiv
Saved in:
| Main Authors: | Lim, Jing Yu, Shah, Rushi, Ikram, Zarif, Yu, Samson, Ma, Haozhe, Leong, Tze-Yun, Liu, Dianbo |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
JEDI: Joint Embedding Diffusion World Model for Online Model-Based Reinforcement Learning
by: Lim, Jing Yu, et al.
Published: (2026)
by: Lim, Jing Yu, et al.
Published: (2026)
Evolution Guided Generative Flow Networks
by: Ikram, Zarif, et al.
Published: (2024)
by: Ikram, Zarif, et al.
Published: (2024)
Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic
by: Vo, Thanh Vinh, et al.
Published: (2025)
by: Vo, Thanh Vinh, et al.
Published: (2025)
Highly Efficient Self-Adaptive Reward Shaping for Reinforcement Learning
by: Ma, Haozhe, et al.
Published: (2024)
by: Ma, Haozhe, et al.
Published: (2024)
Centralized Reward Agent for Knowledge Sharing and Transfer in Multi-Task Reinforcement Learning
by: Ma, Haozhe, et al.
Published: (2024)
by: Ma, Haozhe, et al.
Published: (2024)
MOBODY: Model Based Off-Dynamics Offline Reinforcement Learning
by: Guo, Yihong, et al.
Published: (2025)
by: Guo, Yihong, et al.
Published: (2025)
Representation Collapsing Problems in Vector Quantization
by: Zhao, Wenhao, et al.
Published: (2024)
by: Zhao, Wenhao, et al.
Published: (2024)
Hierarchical Molecular Representation Learning via Fragment-Based Self-Supervised Embedding Prediction
by: Wu, Jiele, et al.
Published: (2026)
by: Wu, Jiele, et al.
Published: (2026)
Improving Discrete Optimisation Via Decoupled Straight-Through Estimator
by: Shah, Rushi, et al.
Published: (2024)
by: Shah, Rushi, et al.
Published: (2024)
Exploration by Random Reward Perturbation
by: Ma, Haozhe, et al.
Published: (2025)
by: Ma, Haozhe, et al.
Published: (2025)
Safe Offline Reinforcement Learning with Feasibility-Guided Diffusion Model
by: Zheng, Yinan, et al.
Published: (2024)
by: Zheng, Yinan, et al.
Published: (2024)
RL-100: Performant Robotic Manipulation with Real-World Reinforcement Learning
by: Lei, Kun, et al.
Published: (2025)
by: Lei, Kun, et al.
Published: (2025)
Continual Model-Based Reinforcement Learning with Hypernetworks
by: Huang, Yizhou, et al.
Published: (2020)
by: Huang, Yizhou, et al.
Published: (2020)
Masked Generative Priors Improve World Models Sequence Modelling Capabilities
by: Meo, Cristian, et al.
Published: (2024)
by: Meo, Cristian, et al.
Published: (2024)
Return Augmented Decision Transformer for Off-Dynamics Reinforcement Learning
by: Wang, Ruhan, et al.
Published: (2024)
by: Wang, Ruhan, et al.
Published: (2024)
Boundary-to-Region Supervision for Offline Safe Reinforcement Learning
by: Su, Huikang, et al.
Published: (2025)
by: Su, Huikang, et al.
Published: (2025)
Learning to Play Air Hockey with Model-Based Deep Reinforcement Learning
by: Orsula, Andrej
Published: (2024)
by: Orsula, Andrej
Published: (2024)
Learning global control of underactuated systems with Model-Based Reinforcement Learning
by: Turcato, Niccolò, et al.
Published: (2025)
by: Turcato, Niccolò, et al.
Published: (2025)
World Action Verifier: Self-Improving World Models via Forward-Inverse Asymmetry
by: Liu, Yuejiang, et al.
Published: (2026)
by: Liu, Yuejiang, et al.
Published: (2026)
Model-Based Reinforcement Learning with Multi-Task Offline Pretraining
by: Pan, Minting, et al.
Published: (2023)
by: Pan, Minting, et al.
Published: (2023)
OmniDrones: An Efficient and Flexible Platform for Reinforcement Learning in Drone Control
by: Xu, Botian, et al.
Published: (2023)
by: Xu, Botian, et al.
Published: (2023)
Self-Improving Safety Performance of Reinforcement Learning Based Driving with Black-Box Verification Algorithms
by: Dagdanov, Resul, et al.
Published: (2022)
by: Dagdanov, Resul, et al.
Published: (2022)
Control Synthesis from Linear Temporal Logic Specifications using Model-Free Reinforcement Learning
by: Bozkurt, Alper Kamil, et al.
Published: (2019)
by: Bozkurt, Alper Kamil, et al.
Published: (2019)
Towards Robust Zero-Shot Reinforcement Learning
by: Zheng, Kexin, et al.
Published: (2025)
by: Zheng, Kexin, et al.
Published: (2025)
Early Quantization Shrinks Codebook: A Simple Fix for Diversity-Preserving Tokenization
by: Zhao, Wenhao, et al.
Published: (2026)
by: Zhao, Wenhao, et al.
Published: (2026)
Video-Enhanced Offline Reinforcement Learning: A Model-Based Approach
by: Pan, Minting, et al.
Published: (2025)
by: Pan, Minting, et al.
Published: (2025)
STORI: A Benchmark and Taxonomy for Stochastic Environments
by: Barsainyan, Aryan Amit, et al.
Published: (2025)
by: Barsainyan, Aryan Amit, et al.
Published: (2025)
Localized Dynamics-Aware Domain Adaption for Off-Dynamics Offline Reinforcement Learning
by: Xia, Zhangjie, et al.
Published: (2026)
by: Xia, Zhangjie, et al.
Published: (2026)
Sampling-Based Safe Reinforcement Learning
by: Vignola, Luca, et al.
Published: (2026)
by: Vignola, Luca, et al.
Published: (2026)
D-SPEAR: Dual-Stream Prioritized Experience Adaptive Replay for Stable Reinforcement Learning in Robotic Manipulation
by: Zhang, Yu, et al.
Published: (2026)
by: Zhang, Yu, et al.
Published: (2026)
Replication of Impedance Identification Experiments on a Reinforcement-Learning-Controlled Digital Twin of Human Elbows
by: Yu, Hao, et al.
Published: (2024)
by: Yu, Hao, et al.
Published: (2024)
Skill-aware Mutual Information Optimisation for Generalisation in Reinforcement Learning
by: Yu, Xuehui, et al.
Published: (2024)
by: Yu, Xuehui, et al.
Published: (2024)
Drama: Mamba-Enabled Model-Based Reinforcement Learning Is Sample and Parameter Efficient
by: Wang, Wenlong, et al.
Published: (2024)
by: Wang, Wenlong, et al.
Published: (2024)
Demonstrating the Octopi-1.5 Visual-Tactile-Language Model
by: Yu, Samson, et al.
Published: (2025)
by: Yu, Samson, et al.
Published: (2025)
Cross-Domain Energy-Guided Diffusion Generation for Off-Dynamics Reinforcement Learning
by: Yang, Yu, et al.
Published: (2026)
by: Yang, Yu, et al.
Published: (2026)
Constraints as Rewards: Reinforcement Learning for Robots without Reward Functions
by: Ishihara, Yu, et al.
Published: (2025)
by: Ishihara, Yu, et al.
Published: (2025)
ENOTO: Improving Offline-to-Online Reinforcement Learning with Q-Ensembles
by: Zhao, Kai, et al.
Published: (2023)
by: Zhao, Kai, et al.
Published: (2023)
Improving Offline Reinforcement Learning with Inaccurate Simulators
by: Hou, Yiwen, et al.
Published: (2024)
by: Hou, Yiwen, et al.
Published: (2024)
SCIZOR: A Self-Supervised Approach to Data Curation for Large-Scale Imitation Learning
by: Zhang, Yu, et al.
Published: (2025)
by: Zhang, Yu, et al.
Published: (2025)
Diffusion Policies with Value-Conditional Optimization for Offline Reinforcement Learning
by: Ma, Yunchang, et al.
Published: (2025)
by: Ma, Yunchang, et al.
Published: (2025)
Similar Items
-
JEDI: Joint Embedding Diffusion World Model for Online Model-Based Reinforcement Learning
by: Lim, Jing Yu, et al.
Published: (2026) -
Evolution Guided Generative Flow Networks
by: Ikram, Zarif, et al.
Published: (2024) -
Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic
by: Vo, Thanh Vinh, et al.
Published: (2025) -
Highly Efficient Self-Adaptive Reward Shaping for Reinforcement Learning
by: Ma, Haozhe, et al.
Published: (2024) -
Centralized Reward Agent for Knowledge Sharing and Transfer in Multi-Task Reinforcement Learning
by: Ma, Haozhe, et al.
Published: (2024)