SPRINQL: Sub-optimal Demonstrations driven Offline Imitation Learning
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Hoang, Huy, Mai, Tien, Varakantham, Pradeep |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Learning What to Do and What Not To Do: Offline Imitation from Expert and Undesirable Demonstrations
von: Hoang, Huy, et al.
Veröffentlicht: (2025)
von: Hoang, Huy, et al.
Veröffentlicht: (2025)
UNIQ: Offline Inverse Q-learning for Avoiding Undesirable Demonstrations
von: Hoang, Huy, et al.
Veröffentlicht: (2024)
von: Hoang, Huy, et al.
Veröffentlicht: (2024)
Imitate the Good and Avoid the Bad: An Incremental Approach to Safe Reinforcement Learning
von: Hoang, Huy, et al.
Veröffentlicht: (2023)
von: Hoang, Huy, et al.
Veröffentlicht: (2023)
Offline Safe Reinforcement Learning Using Trajectory Classification
von: Gong, Ze, et al.
Veröffentlicht: (2024)
von: Gong, Ze, et al.
Veröffentlicht: (2024)
Imitating Cost-Constrained Behaviors in Reinforcement Learning
von: Shao, Qian, et al.
Veröffentlicht: (2024)
von: Shao, Qian, et al.
Veröffentlicht: (2024)
On Learning Informative Trajectory Embeddings for Imitation, Classification and Regression
von: Ge, Zichang, et al.
Veröffentlicht: (2025)
von: Ge, Zichang, et al.
Veröffentlicht: (2025)
Solving Richly Constrained Reinforcement Learning through State Augmentation and Reward Penalties
von: Jiang, Hao, et al.
Veröffentlicht: (2023)
von: Jiang, Hao, et al.
Veröffentlicht: (2023)
How to Leverage Diverse Demonstrations in Offline Imitation Learning
von: Yue, Sheng, et al.
Veröffentlicht: (2024)
von: Yue, Sheng, et al.
Veröffentlicht: (2024)
Enhancing the Hierarchical Environment Design via Generative Trajectory Modeling
von: Li, Dexun, et al.
Veröffentlicht: (2023)
von: Li, Dexun, et al.
Veröffentlicht: (2023)
MisoDICE: Multi-Agent Imitation from Unlabeled Mixed-Quality Demonstrations
von: Bui, The Viet, et al.
Veröffentlicht: (2025)
von: Bui, The Viet, et al.
Veröffentlicht: (2025)
Handling Long and Richly Constrained Tasks through Constrained Hierarchical Reinforcement Learning
von: Lu, Yuxiao, et al.
Veröffentlicht: (2023)
von: Lu, Yuxiao, et al.
Veröffentlicht: (2023)
On Minimizing Adversarial Counterfactual Error in Adversarial RL
von: Belaire, Roman, et al.
Veröffentlicht: (2024)
von: Belaire, Roman, et al.
Veröffentlicht: (2024)
Automatic LLM Red Teaming
von: Belaire, Roman, et al.
Veröffentlicht: (2025)
von: Belaire, Roman, et al.
Veröffentlicht: (2025)
Regret-Based Defense in Adversarial Reinforcement Learning
von: Belaire, Roman, et al.
Veröffentlicht: (2023)
von: Belaire, Roman, et al.
Veröffentlicht: (2023)
DITTO: Offline Imitation Learning with World Models
von: DeMoss, Branton, et al.
Veröffentlicht: (2023)
von: DeMoss, Branton, et al.
Veröffentlicht: (2023)
Offline Imitation Learning with Model-based Reverse Augmentation
von: Shao, Jie-Jing, et al.
Veröffentlicht: (2024)
von: Shao, Jie-Jing, et al.
Veröffentlicht: (2024)
OLLIE: Imitation Learning from Offline Pretraining to Online Finetuning
von: Yue, Sheng, et al.
Veröffentlicht: (2024)
von: Yue, Sheng, et al.
Veröffentlicht: (2024)
Balance Equation-based Distributionally Robust Offline Imitation Learning
von: Agrawal, Rishabh, et al.
Veröffentlicht: (2025)
von: Agrawal, Rishabh, et al.
Veröffentlicht: (2025)
Offline Imitation Learning Through Graph Search and Retrieval
von: Yin, Zhao-Heng, et al.
Veröffentlicht: (2024)
von: Yin, Zhao-Heng, et al.
Veröffentlicht: (2024)
SEABO: A Simple Search-Based Method for Offline Imitation Learning
von: Lyu, Jiafei, et al.
Veröffentlicht: (2024)
von: Lyu, Jiafei, et al.
Veröffentlicht: (2024)
Align Your Intents: Offline Imitation Learning via Optimal Transport
von: Bobrin, Maksim, et al.
Veröffentlicht: (2024)
von: Bobrin, Maksim, et al.
Veröffentlicht: (2024)
On Discovering Algorithms for Adversarial Imitation Learning
von: Chirra, Shashank Reddy, et al.
Veröffentlicht: (2025)
von: Chirra, Shashank Reddy, et al.
Veröffentlicht: (2025)
Imitation Learning from Suboptimal Demonstrations via Meta-Learning An Action Ranker
von: Fan, Jiangdong, et al.
Veröffentlicht: (2024)
von: Fan, Jiangdong, et al.
Veröffentlicht: (2024)
LMAct: A Benchmark for In-Context Imitation Learning with Long Multimodal Demonstrations
von: Ruoss, Anian, et al.
Veröffentlicht: (2024)
von: Ruoss, Anian, et al.
Veröffentlicht: (2024)
Offline Safe Policy Optimization From Heterogeneous Feedback
von: Gong, Ze, et al.
Veröffentlicht: (2025)
von: Gong, Ze, et al.
Veröffentlicht: (2025)
IRL for Restless Multi-Armed Bandits with Applications in Maternal and Child Health
von: Jain, Gauri, et al.
Veröffentlicht: (2024)
von: Jain, Gauri, et al.
Veröffentlicht: (2024)
Markov Balance Satisfaction Improves Performance in Strictly Batch Offline Imitation Learning
von: Agrawal, Rishabh, et al.
Veröffentlicht: (2024)
von: Agrawal, Rishabh, et al.
Veröffentlicht: (2024)
Beyond-Expert Performance with Limited Demonstrations: Efficient Imitation Learning with Double Exploration
von: Zhao, Heyang, et al.
Veröffentlicht: (2025)
von: Zhao, Heyang, et al.
Veröffentlicht: (2025)
A Dual Approach to Imitation Learning from Observations with Offline Datasets
von: Sikchi, Harshit, et al.
Veröffentlicht: (2024)
von: Sikchi, Harshit, et al.
Veröffentlicht: (2024)
Offline Diversity Maximization Under Imitation Constraints
von: Vlastelica, Marin, et al.
Veröffentlicht: (2023)
von: Vlastelica, Marin, et al.
Veröffentlicht: (2023)
PROF: An LLM-based Reward Code Preference Optimization Framework for Offline Imitation Learning
von: Sun, Shengjie, et al.
Veröffentlicht: (2025)
von: Sun, Shengjie, et al.
Veröffentlicht: (2025)
SafeMIL: Learning Offline Safe Imitation Policy from Non-Preferred Trajectories
von: Burnwal, Returaj, et al.
Veröffentlicht: (2025)
von: Burnwal, Returaj, et al.
Veröffentlicht: (2025)
Learning Surrogates for Offline Black-Box Optimization via Gradient Matching
von: Hoang, Minh, et al.
Veröffentlicht: (2025)
von: Hoang, Minh, et al.
Veröffentlicht: (2025)
Using Non-Expert Data to Robustify Imitation Learning via Offline Reinforcement Learning
von: Huang, Kevin, et al.
Veröffentlicht: (2025)
von: Huang, Kevin, et al.
Veröffentlicht: (2025)
OSIL: Learning Offline Safe Imitation Policies with Safety Inferred from Non-preferred Trajectories
von: Burnwal, Returaj, et al.
Veröffentlicht: (2026)
von: Burnwal, Returaj, et al.
Veröffentlicht: (2026)
Hierarchical Imitation Learning of Team Behavior from Heterogeneous Demonstrations
von: Seo, Sangwon, et al.
Veröffentlicht: (2025)
von: Seo, Sangwon, et al.
Veröffentlicht: (2025)
Offline-to-online Reinforcement Learning for Image-based Grasping with Scarce Demonstrations
von: Chan, Bryan, et al.
Veröffentlicht: (2024)
von: Chan, Bryan, et al.
Veröffentlicht: (2024)
A Unified Linear Programming Framework for Offline Reward Learning from Human Demonstrations and Feedback
von: Kim, Kihyun, et al.
Veröffentlicht: (2024)
von: Kim, Kihyun, et al.
Veröffentlicht: (2024)
Momentum Contrastive Learning with Enhanced Negative Sampling and Hard Negative Filtering
von: Hoang, Duy, et al.
Veröffentlicht: (2025)
von: Hoang, Duy, et al.
Veröffentlicht: (2025)
Offline Imitation of Badminton Player Behavior via Experiential Contexts and Brownian Motion
von: Wang, Kuang-Da, et al.
Veröffentlicht: (2024)
von: Wang, Kuang-Da, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Learning What to Do and What Not To Do: Offline Imitation from Expert and Undesirable Demonstrations
von: Hoang, Huy, et al.
Veröffentlicht: (2025) -
UNIQ: Offline Inverse Q-learning for Avoiding Undesirable Demonstrations
von: Hoang, Huy, et al.
Veröffentlicht: (2024) -
Imitate the Good and Avoid the Bad: An Incremental Approach to Safe Reinforcement Learning
von: Hoang, Huy, et al.
Veröffentlicht: (2023) -
Offline Safe Reinforcement Learning Using Trajectory Classification
von: Gong, Ze, et al.
Veröffentlicht: (2024) -
Imitating Cost-Constrained Behaviors in Reinforcement Learning
von: Shao, Qian, et al.
Veröffentlicht: (2024)