TOP-ERL: Transformer-based Off-Policy Episodic Reinforcement Learning
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Li, Ge, Tian, Dong, Zhou, Hongyi, Jiang, Xinkai, Lioutikov, Rudolf, Neumann, Gerhard |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
MoRe-ERL: Learning Motion Residuals using Episodic Reinforcement Learning
von: Huang, Xi, et al.
Veröffentlicht: (2025)
von: Huang, Xi, et al.
Veröffentlicht: (2025)
Open the Black Box: Step-based Policy Updates for Temporally-Correlated Episodic Reinforcement Learning
von: Li, Ge, et al.
Veröffentlicht: (2024)
von: Li, Ge, et al.
Veröffentlicht: (2024)
Variational Distillation of Diffusion Policies into Mixture of Experts
von: Zhou, Hongyi, et al.
Veröffentlicht: (2024)
von: Zhou, Hongyi, et al.
Veröffentlicht: (2024)
PointMapPolicy: Structured Point Cloud Processing for Multi-Modal Imitation Learning
von: Jia, Xiaogang, et al.
Veröffentlicht: (2025)
von: Jia, Xiaogang, et al.
Veröffentlicht: (2025)
Efficient Diffusion Transformer Policies with Mixture of Expert Denoisers for Multitask Learning
von: Reuss, Moritz, et al.
Veröffentlicht: (2024)
von: Reuss, Moritz, et al.
Veröffentlicht: (2024)
MaIL: Improving Imitation Learning with Mamba
von: Jia, Xiaogang, et al.
Veröffentlicht: (2024)
von: Jia, Xiaogang, et al.
Veröffentlicht: (2024)
X-IL: Exploring the Design Space of Imitation Learning Policies
von: Jia, Xiaogang, et al.
Veröffentlicht: (2025)
von: Jia, Xiaogang, et al.
Veröffentlicht: (2025)
IRIS: An Immersive Robot Interaction System
von: Jiang, Xinkai, et al.
Veröffentlicht: (2025)
von: Jiang, Xinkai, et al.
Veröffentlicht: (2025)
BMP: Bridging the Gap between B-Spline and Movement Primitives
von: Liao, Weiran, et al.
Veröffentlicht: (2024)
von: Liao, Weiran, et al.
Veröffentlicht: (2024)
BEAST: Efficient Tokenization of B-Splines Encoded Action Sequences for Imitation Learning
von: Zhou, Hongyi, et al.
Veröffentlicht: (2025)
von: Zhou, Hongyi, et al.
Veröffentlicht: (2025)
Beyond Visuals: Investigating Force Feedback in Extended Reality for Robot Data Collection
von: Li, Xueyin, et al.
Veröffentlicht: (2025)
von: Li, Xueyin, et al.
Veröffentlicht: (2025)
Use the Force, Bot! -- Force-Aware ProDMP with Event-Based Replanning
von: Lödige, Paul Werner, et al.
Veröffentlicht: (2024)
von: Lödige, Paul Werner, et al.
Veröffentlicht: (2024)
Movement Primitive Diffusion: Learning Gentle Robotic Manipulation of Deformable Objects
von: Scheikl, Paul Maria, et al.
Veröffentlicht: (2023)
von: Scheikl, Paul Maria, et al.
Veröffentlicht: (2023)
Towards Diverse Behaviors: A Benchmark for Imitation Learning with Human Demonstrations
von: Jia, Xiaogang, et al.
Veröffentlicht: (2024)
von: Jia, Xiaogang, et al.
Veröffentlicht: (2024)
Off-Policy Safe Reinforcement Learning with Constrained Optimistic Exploration
von: Li, Guopeng, et al.
Veröffentlicht: (2026)
von: Li, Guopeng, et al.
Veröffentlicht: (2026)
Acquiring Diverse Skills using Curriculum Reinforcement Learning with Mixture of Experts
von: Celik, Onur, et al.
Veröffentlicht: (2024)
von: Celik, Onur, et al.
Veröffentlicht: (2024)
Off Policy Lyapunov Stability in Reinforcement Learning
von: Gill, Sarvan, et al.
Veröffentlicht: (2025)
von: Gill, Sarvan, et al.
Veröffentlicht: (2025)
Return Augmented Decision Transformer for Off-Dynamics Reinforcement Learning
von: Wang, Ruhan, et al.
Veröffentlicht: (2024)
von: Wang, Ruhan, et al.
Veröffentlicht: (2024)
PointPatchRL -- Masked Reconstruction Improves Reinforcement Learning on Point Clouds
von: Gyenes, Balázs, et al.
Veröffentlicht: (2024)
von: Gyenes, Balázs, et al.
Veröffentlicht: (2024)
Formulating Reinforcement Learning for Human-Robot Collaboration through Off-Policy Evaluation
von: Singh, Saurav, et al.
Veröffentlicht: (2026)
von: Singh, Saurav, et al.
Veröffentlicht: (2026)
FlashSAC: Fast and Stable Off-Policy Reinforcement Learning for High-Dimensional Robot Control
von: Kim, Donghu, et al.
Veröffentlicht: (2026)
von: Kim, Donghu, et al.
Veröffentlicht: (2026)
Scaling Robot Policy Learning via Zero-Shot Labeling with Foundation Models
von: Blank, Nils, et al.
Veröffentlicht: (2024)
von: Blank, Nils, et al.
Veröffentlicht: (2024)
Residual Off-Policy RL for Finetuning Behavior Cloning Policies
von: Ankile, Lars, et al.
Veröffentlicht: (2025)
von: Ankile, Lars, et al.
Veröffentlicht: (2025)
FLOWER: Democratizing Generalist Robot Policies with Efficient Vision-Language-Action Flow Policies
von: Reuss, Moritz, et al.
Veröffentlicht: (2025)
von: Reuss, Moritz, et al.
Veröffentlicht: (2025)
HOPE: A Reinforcement Learning-based Hybrid Policy Path Planner for Diverse Parking Scenarios
von: Jiang, Mingyang, et al.
Veröffentlicht: (2024)
von: Jiang, Mingyang, et al.
Veröffentlicht: (2024)
Quantum-Inspired Episode Selection for Monte Carlo Reinforcement Learning via QUBO Optimization
von: Salloum, Hadi, et al.
Veröffentlicht: (2026)
von: Salloum, Hadi, et al.
Veröffentlicht: (2026)
A Retrospective on the Robot Air Hockey Challenge: Benchmarking Robust, Reliable, and Safe Learning Techniques for Real-world Robotics
von: Liu, Puze, et al.
Veröffentlicht: (2024)
von: Liu, Puze, et al.
Veröffentlicht: (2024)
SATA: Safe and Adaptive Torque-Based Locomotion Policies Inspired by Animal Learning
von: Li, Peizhuo, et al.
Veröffentlicht: (2025)
von: Li, Peizhuo, et al.
Veröffentlicht: (2025)
Towards Fusing Point Cloud and Visual Representations for Imitation Learning
von: Donat, Atalay, et al.
Veröffentlicht: (2025)
von: Donat, Atalay, et al.
Veröffentlicht: (2025)
MagBotSim: Physics-Based Simulation and Reinforcement Learning Environments for Magnetic Robotics
von: Bergmann, Lara, et al.
Veröffentlicht: (2025)
von: Bergmann, Lara, et al.
Veröffentlicht: (2025)
Context-aware Learned Mesh-based Simulation via Trajectory-Level Meta-Learning
von: Dahlinger, Philipp, et al.
Veröffentlicht: (2025)
von: Dahlinger, Philipp, et al.
Veröffentlicht: (2025)
TADPO: Reinforcement Learning Goes Off-road
von: Wu, Zhouchonghao, et al.
Veröffentlicht: (2026)
von: Wu, Zhouchonghao, et al.
Veröffentlicht: (2026)
VPWEM: Non-Markovian Visuomotor Policy with Working and Episodic Memory
von: Lei, Yuheng, et al.
Veröffentlicht: (2026)
von: Lei, Yuheng, et al.
Veröffentlicht: (2026)
DRARL: Disengagement-Reason-Augmented Reinforcement Learning for Efficient Improvement of Autonomous Driving Policy
von: Zhou, Weitao, et al.
Veröffentlicht: (2025)
von: Zhou, Weitao, et al.
Veröffentlicht: (2025)
MuTT: A Multimodal Trajectory Transformer for Robot Skills
von: Kienle, Claudius, et al.
Veröffentlicht: (2024)
von: Kienle, Claudius, et al.
Veröffentlicht: (2024)
PolicyFlow: Policy Optimization with Continuous Normalizing Flow in Reinforcement Learning
von: Yang, Shunpeng, et al.
Veröffentlicht: (2026)
von: Yang, Shunpeng, et al.
Veröffentlicht: (2026)
ReinforceGen: Hybrid Skill Policies with Automated Data Generation and Reinforcement Learning
von: Zhou, Zihan, et al.
Veröffentlicht: (2025)
von: Zhou, Zihan, et al.
Veröffentlicht: (2025)
MOBODY: Model Based Off-Dynamics Offline Reinforcement Learning
von: Guo, Yihong, et al.
Veröffentlicht: (2025)
von: Guo, Yihong, et al.
Veröffentlicht: (2025)
STITCH-OPE: Trajectory Stitching with Guided Diffusion for Off-Policy Evaluation
von: Goli, Hossein, et al.
Veröffentlicht: (2025)
von: Goli, Hossein, et al.
Veröffentlicht: (2025)
ConceptACT: Episode-Level Concepts for Sample-Efficient Robotic Imitation Learning
von: Karalus, Jakob, et al.
Veröffentlicht: (2026)
von: Karalus, Jakob, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
MoRe-ERL: Learning Motion Residuals using Episodic Reinforcement Learning
von: Huang, Xi, et al.
Veröffentlicht: (2025) -
Open the Black Box: Step-based Policy Updates for Temporally-Correlated Episodic Reinforcement Learning
von: Li, Ge, et al.
Veröffentlicht: (2024) -
Variational Distillation of Diffusion Policies into Mixture of Experts
von: Zhou, Hongyi, et al.
Veröffentlicht: (2024) -
PointMapPolicy: Structured Point Cloud Processing for Multi-Modal Imitation Learning
von: Jia, Xiaogang, et al.
Veröffentlicht: (2025) -
Efficient Diffusion Transformer Policies with Mixture of Expert Denoisers for Multitask Learning
von: Reuss, Moritz, et al.
Veröffentlicht: (2024)