Meta-DT: Offline Meta-RL as Conditional Sequence Modeling with World Model Disentanglement
Fuente:
arXiv
Saved in:
| Main Authors: | Wang, Zhi, Zhang, Li, Wu, Wenhao, Zhu, Yuanheng, Zhao, Dongbin, Chen, Chunlin |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Discretizing Continuous Action Space with Unimodal Probability Distributions for On-Policy Reinforcement Learning
by: Zhu, Yuanyang, et al.
Published: (2024)
by: Zhu, Yuanyang, et al.
Published: (2024)
Reinformer: Max-Return Sequence Modeling for Offline RL
by: Zhuang, Zifeng, et al.
Published: (2024)
by: Zhuang, Zifeng, et al.
Published: (2024)
ARAC: Adaptive Regularized Multi-Agent Soft Actor-Critic in Graph-Structured Adversarial Games
by: Shi, Ruochuan, et al.
Published: (2025)
by: Shi, Ruochuan, et al.
Published: (2025)
R2PS: Worst-Case Robust Real-Time Pursuit Strategies under Partial Observability
by: Lu, Runyu, et al.
Published: (2025)
by: Lu, Runyu, et al.
Published: (2025)
Decision MetaMamba: Enhancing Selective SSM in Offline RL with Heterogeneous Sequence Mixing
by: Kim, Wall, et al.
Published: (2024)
by: Kim, Wall, et al.
Published: (2024)
Decision MetaMamba: Enhancing Selective SSM in Offline RL with Heterogeneous Sequence Mixing
by: Kim, Wall, et al.
Published: (2026)
by: Kim, Wall, et al.
Published: (2026)
Contextual Latent World Models for Offline Meta Reinforcement Learning
by: Nakheai, Mohammadreza, et al.
Published: (2026)
by: Nakheai, Mohammadreza, et al.
Published: (2026)
Making Offline RL Online: Collaborative World Models for Offline Visual Reinforcement Learning
by: Wang, Qi, et al.
Published: (2023)
by: Wang, Qi, et al.
Published: (2023)
Behavior-Invariant Task Representation Learning with Transformer-based World Models for Offline Meta-Reinforcement Learning
by: Qian, Fuyuan, et al.
Published: (2026)
by: Qian, Fuyuan, et al.
Published: (2026)
Meta-World+: An Improved, Standardized, RL Benchmark
by: McLean, Reginald, et al.
Published: (2025)
by: McLean, Reginald, et al.
Published: (2025)
Towards an Information Theoretic Framework of Context-Based Offline Meta-Reinforcement Learning
by: Li, Lanqing, et al.
Published: (2024)
by: Li, Lanqing, et al.
Published: (2024)
Inference Time Policy Optimization for Offline RL with Differentiable World Models
by: Deb, Rohan, et al.
Published: (2026)
by: Deb, Rohan, et al.
Published: (2026)
DipLLM: Fine-Tuning LLM for Strategic Decision-making in Diplomacy
by: Xu, Kaixuan, et al.
Published: (2025)
by: Xu, Kaixuan, et al.
Published: (2025)
Are Expressive Models Truly Necessary for Offline RL?
by: Wang, Guan, et al.
Published: (2024)
by: Wang, Guan, et al.
Published: (2024)
Conditional Diffusion Modeling with Attention for Probabilistic Battery Capacity Prediction under Real-World Condition
by: Jiang, Chunlin, et al.
Published: (2025)
by: Jiang, Chunlin, et al.
Published: (2025)
Meta-DiffuB: A Contextualized Sequence-to-Sequence Text Diffusion Model with Meta-Exploration
by: Chuang, Yun-Yen, et al.
Published: (2024)
by: Chuang, Yun-Yen, et al.
Published: (2024)
HarmoDT: Harmony Multi-Task Decision Transformer for Offline Reinforcement Learning
by: Hu, Shengchao, et al.
Published: (2024)
by: Hu, Shengchao, et al.
Published: (2024)
OASIS: Conditional Distribution Shaping for Offline Safe Reinforcement Learning
by: Yao, Yihang, et al.
Published: (2024)
by: Yao, Yihang, et al.
Published: (2024)
Scaling Offline Model-Based RL via Jointly-Optimized World-Action Model Pretraining
by: Cheng, Jie, et al.
Published: (2024)
by: Cheng, Jie, et al.
Published: (2024)
From Path Signatures to Sequential Modeling: Incremental Signature Contributions for Offline RL
by: Zhao, Ziyi, et al.
Published: (2026)
by: Zhao, Ziyi, et al.
Published: (2026)
Continual Offline Reinforcement Learning via Diffusion-based Dual Generative Replay
by: Liu, Jinmei, et al.
Published: (2024)
by: Liu, Jinmei, et al.
Published: (2024)
Equilibrium Policy Generalization: A Reinforcement Learning Framework for Cross-Graph Zero-Shot Generalization in Pursuit-Evasion Games
by: Lu, Runyu, et al.
Published: (2025)
by: Lu, Runyu, et al.
Published: (2025)
Meta-probabilistic Modeling
by: Zhang, Kevin, et al.
Published: (2026)
by: Zhang, Kevin, et al.
Published: (2026)
Dual Alignment Maximin Optimization for Offline Model-based RL
by: Zhou, Chi, et al.
Published: (2025)
by: Zhou, Chi, et al.
Published: (2025)
When Are RL Hyperparameters Benign? A Study in Offline Goal-Conditioned RL
by: Töpperwien, Jan Malte, et al.
Published: (2026)
by: Töpperwien, Jan Malte, et al.
Published: (2026)
DEAS: DEtached value learning with Action Sequence for Scalable Offline RL
by: Kim, Changyeon, et al.
Published: (2025)
by: Kim, Changyeon, et al.
Published: (2025)
RLAE: Reinforcement Learning-Assisted Ensemble for LLMs
by: Fu, Yuqian, et al.
Published: (2025)
by: Fu, Yuqian, et al.
Published: (2025)
OGBench: Benchmarking Offline Goal-Conditioned RL
by: Park, Seohong, et al.
Published: (2024)
by: Park, Seohong, et al.
Published: (2024)
Exploiting Structure in Offline Multi-Agent RL: The Benefits of Low Interaction Rank
by: Zhan, Wenhao, et al.
Published: (2024)
by: Zhan, Wenhao, et al.
Published: (2024)
Multi-fidelity Parameter Estimation Using Conditional Diffusion Models
by: Tatsuoka, Caroline, et al.
Published: (2025)
by: Tatsuoka, Caroline, et al.
Published: (2025)
MetaDiff: Meta-Learning with Conditional Diffusion for Few-Shot Learning
by: Zhang, Baoquan, et al.
Published: (2023)
by: Zhang, Baoquan, et al.
Published: (2023)
A Training-Free Conditional Diffusion Model for Learning Stochastic Dynamical Systems
by: Liu, Yanfang, et al.
Published: (2024)
by: Liu, Yanfang, et al.
Published: (2024)
GAS: Enhancing Reward-Cost Balance of Generative Model-assisted Offline Safe RL
by: Liu, Zifan, et al.
Published: (2026)
by: Liu, Zifan, et al.
Published: (2026)
BECAUSE: Bilinear Causal Representation for Generalizable Offline Model-based Reinforcement Learning
by: Lin, Haohong, et al.
Published: (2024)
by: Lin, Haohong, et al.
Published: (2024)
Dream to Drive with Predictive Individual World Model
by: Gao, Yinfeng, et al.
Published: (2025)
by: Gao, Yinfeng, et al.
Published: (2025)
Augmenting Offline RL with Unlabeled Data
by: Wang, Zhao, et al.
Published: (2024)
by: Wang, Zhao, et al.
Published: (2024)
Fat-to-Thin Policy Optimization: Offline RL with Sparse Policies
by: Zhu, Lingwei, et al.
Published: (2025)
by: Zhu, Lingwei, et al.
Published: (2025)
Generalizing Consistency Policy to Visual RL with Prioritized Proximal Experience Regularization
by: Li, Haoran, et al.
Published: (2024)
by: Li, Haoran, et al.
Published: (2024)
Mitigating Distribution Shift in Model-based Offline RL via Shifts-aware Reward Learning
by: Luo, Wang, et al.
Published: (2024)
by: Luo, Wang, et al.
Published: (2024)
Dual-Granularity Contrastive Reward via Generated Episodic Guidance for Efficient Embodied RL
by: Liu, Xin, et al.
Published: (2026)
by: Liu, Xin, et al.
Published: (2026)
Similar Items
-
Discretizing Continuous Action Space with Unimodal Probability Distributions for On-Policy Reinforcement Learning
by: Zhu, Yuanyang, et al.
Published: (2024) -
Reinformer: Max-Return Sequence Modeling for Offline RL
by: Zhuang, Zifeng, et al.
Published: (2024) -
ARAC: Adaptive Regularized Multi-Agent Soft Actor-Critic in Graph-Structured Adversarial Games
by: Shi, Ruochuan, et al.
Published: (2025) -
R2PS: Worst-Case Robust Real-Time Pursuit Strategies under Partial Observability
by: Lu, Runyu, et al.
Published: (2025) -
Decision MetaMamba: Enhancing Selective SSM in Offline RL with Heterogeneous Sequence Mixing
by: Kim, Wall, et al.
Published: (2024)