Beyond Non-Expert Demonstrations: Outcome-Driven Action Constraint for Offline Reinforcement Learning
Fuente:
arXiv
Guardado en:
| Autores principales: | Jiang, Ke, Jiang, Wen, Li, Yao, Tan, Xiaoyang |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Variational OOD State Correction for Offline Reinforcement Learning
por: Jiang, Ke, et al.
Publicado: (2025)
por: Jiang, Ke, et al.
Publicado: (2025)
Beyond Hard Constraints: Budget-Conditioned Reachability For Safe Offline Reinforcement Learning
por: Brahmanage, Janaka Chathuranga, et al.
Publicado: (2026)
por: Brahmanage, Janaka Chathuranga, et al.
Publicado: (2026)
Using Non-Expert Data to Robustify Imitation Learning via Offline Reinforcement Learning
por: Huang, Kevin, et al.
Publicado: (2025)
por: Huang, Kevin, et al.
Publicado: (2025)
Learning Gentle Grasping from Human-Free Force Control Demonstration
por: Li, Mingxuan, et al.
Publicado: (2024)
por: Li, Mingxuan, et al.
Publicado: (2024)
Offline-to-online Reinforcement Learning for Image-based Grasping with Scarce Demonstrations
por: Chan, Bryan, et al.
Publicado: (2024)
por: Chan, Bryan, et al.
Publicado: (2024)
Robustness Evaluation of Offline Reinforcement Learning for Robot Control Against Action Perturbations
por: Ayabe, Shingo, et al.
Publicado: (2024)
por: Ayabe, Shingo, et al.
Publicado: (2024)
Positive-Unlabeled Constraint Learning for Inferring Nonlinear Continuous Constraints Functions from Expert Demonstrations
por: Peng, Baiyu, et al.
Publicado: (2024)
por: Peng, Baiyu, et al.
Publicado: (2024)
Safe Offline Reinforcement Learning with Real-Time Budget Constraints
por: Lin, Qian, et al.
Publicado: (2023)
por: Lin, Qian, et al.
Publicado: (2023)
Offline Imitation Learning upon Arbitrary Demonstrations by Pre-Training Dynamics Representations
por: Ma, Haitong, et al.
Publicado: (2025)
por: Ma, Haitong, et al.
Publicado: (2025)
Equivariant Offline Reinforcement Learning
por: Tangri, Arsh, et al.
Publicado: (2024)
por: Tangri, Arsh, et al.
Publicado: (2024)
CO-RFT: Efficient Fine-Tuning of Vision-Language-Action Models through Chunked Offline Reinforcement Learning
por: Huang, Dongchi, et al.
Publicado: (2025)
por: Huang, Dongchi, et al.
Publicado: (2025)
Diffusion Models for Offline Multi-agent Reinforcement Learning with Safety Constraints
por: Huang, Jianuo
Publicado: (2024)
por: Huang, Jianuo
Publicado: (2024)
Forecasting in Offline Reinforcement Learning for Non-stationary Environments
por: Ada, Suzan Ece, et al.
Publicado: (2025)
por: Ada, Suzan Ece, et al.
Publicado: (2025)
Long and Short-Term Constraints Driven Safe Reinforcement Learning for Autonomous Driving
por: Hu, Xuemin, et al.
Publicado: (2024)
por: Hu, Xuemin, et al.
Publicado: (2024)
Conditional Neural Expert Processes for Learning Movement Primitives from Demonstration
por: Yildirim, Yigit, et al.
Publicado: (2024)
por: Yildirim, Yigit, et al.
Publicado: (2024)
Guided Data Augmentation for Offline Reinforcement Learning and Imitation Learning
por: Corrado, Nicholas E., et al.
Publicado: (2023)
por: Corrado, Nicholas E., et al.
Publicado: (2023)
Constraint-Aware Reinforcement Learning via Adaptive Action Scaling
por: Dawood, Murad, et al.
Publicado: (2025)
por: Dawood, Murad, et al.
Publicado: (2025)
Trajectory-Level Data Augmentation for Offline Reinforcement Learning
por: Schmähling, Tobias, et al.
Publicado: (2026)
por: Schmähling, Tobias, et al.
Publicado: (2026)
Adaptive Q-Chunking for Offline-to-Online Reinforcement Learning
por: Gireesh, Nandiraju, et al.
Publicado: (2026)
por: Gireesh, Nandiraju, et al.
Publicado: (2026)
Q-value Regularized Decision ConvFormer for Offline Reinforcement Learning
por: Yan, Teng, et al.
Publicado: (2024)
por: Yan, Teng, et al.
Publicado: (2024)
Residual Learning and Context Encoding for Adaptive Offline-to-Online Reinforcement Learning
por: Nakhaei, Mohammadreza, et al.
Publicado: (2024)
por: Nakhaei, Mohammadreza, et al.
Publicado: (2024)
Causal Flow Q-Learning for Robust Offline Reinforcement Learning
por: Li, Mingxuan, et al.
Publicado: (2026)
por: Li, Mingxuan, et al.
Publicado: (2026)
Fast Lifelong Adaptive Inverse Reinforcement Learning from Demonstrations
por: Chen, Letian, et al.
Publicado: (2022)
por: Chen, Letian, et al.
Publicado: (2022)
A Real-World Quadrupedal Locomotion Benchmark for Offline Reinforcement Learning
por: Zhang, Hongyin, et al.
Publicado: (2023)
por: Zhang, Hongyin, et al.
Publicado: (2023)
Demonstration Sidetracks: Categorizing Systematic Non-Optimality in Human Demonstrations
por: Fang, Shijie, et al.
Publicado: (2025)
por: Fang, Shijie, et al.
Publicado: (2025)
Safe Offline Reinforcement Learning with Feasibility-Guided Diffusion Model
por: Zheng, Yinan, et al.
Publicado: (2024)
por: Zheng, Yinan, et al.
Publicado: (2024)
Rainbow-DemoRL: Combining Improvements in Demonstration-Augmented Reinforcement Learning
por: Bhatt, Dwait, et al.
Publicado: (2026)
por: Bhatt, Dwait, et al.
Publicado: (2026)
Offline Reinforcement Learning with Discrete Diffusion Skills
por: Qiao, RuiXi, et al.
Publicado: (2025)
por: Qiao, RuiXi, et al.
Publicado: (2025)
A Clean Slate for Offline Reinforcement Learning
por: Jackson, Matthew Thomas, et al.
Publicado: (2025)
por: Jackson, Matthew Thomas, et al.
Publicado: (2025)
Improving Offline Reinforcement Learning with Inaccurate Simulators
por: Hou, Yiwen, et al.
Publicado: (2024)
por: Hou, Yiwen, et al.
Publicado: (2024)
Learning Parameterized Skills from Demonstrations
por: Gupta, Vedant, et al.
Publicado: (2025)
por: Gupta, Vedant, et al.
Publicado: (2025)
Reinforcement Learning with Action Chunking
por: Li, Qiyang, et al.
Publicado: (2025)
por: Li, Qiyang, et al.
Publicado: (2025)
Adaptive Advantage-Guided Policy Regularization for Offline Reinforcement Learning
por: Liu, Tenglong, et al.
Publicado: (2024)
por: Liu, Tenglong, et al.
Publicado: (2024)
Dual Action Policy for Robust Sim-to-Real Reinforcement Learning
por: Terence, Ng Wen Zheng, et al.
Publicado: (2024)
por: Terence, Ng Wen Zheng, et al.
Publicado: (2024)
Multiagent Reinforcement Learning with Neighbor Action Estimation
por: Luo, Zhenglong, et al.
Publicado: (2026)
por: Luo, Zhenglong, et al.
Publicado: (2026)
Occupancy Reward Shaping: Improving Credit Assignment for Offline Goal-Conditioned Reinforcement Learning
por: Venugopal, Aravind, et al.
Publicado: (2026)
por: Venugopal, Aravind, et al.
Publicado: (2026)
Boundary-to-Region Supervision for Offline Safe Reinforcement Learning
por: Su, Huikang, et al.
Publicado: (2025)
por: Su, Huikang, et al.
Publicado: (2025)
Graph-Assisted Stitching for Offline Hierarchical Reinforcement Learning
por: Baek, Seungho, et al.
Publicado: (2025)
por: Baek, Seungho, et al.
Publicado: (2025)
Goal-Conditioned Data Augmentation for Offline Reinforcement Learning
por: Huang, Xingshuai, et al.
Publicado: (2024)
por: Huang, Xingshuai, et al.
Publicado: (2024)
Robotic Manipulation Datasets for Offline Compositional Reinforcement Learning
por: Hussing, Marcel, et al.
Publicado: (2023)
por: Hussing, Marcel, et al.
Publicado: (2023)
Ejemplares similares
-
Variational OOD State Correction for Offline Reinforcement Learning
por: Jiang, Ke, et al.
Publicado: (2025) -
Beyond Hard Constraints: Budget-Conditioned Reachability For Safe Offline Reinforcement Learning
por: Brahmanage, Janaka Chathuranga, et al.
Publicado: (2026) -
Using Non-Expert Data to Robustify Imitation Learning via Offline Reinforcement Learning
por: Huang, Kevin, et al.
Publicado: (2025) -
Learning Gentle Grasping from Human-Free Force Control Demonstration
por: Li, Mingxuan, et al.
Publicado: (2024) -
Offline-to-online Reinforcement Learning for Image-based Grasping with Scarce Demonstrations
por: Chan, Bryan, et al.
Publicado: (2024)