Identifying Latent Actions and Dynamics from Offline Data via Demonstrator Diversity
Fuente:
arXiv
Gespeichert in:
| 1. Verfasser: | Schur, Felix |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
How to Leverage Diverse Demonstrations in Offline Imitation Learning
von: Yue, Sheng, et al.
Veröffentlicht: (2024)
von: Yue, Sheng, et al.
Veröffentlicht: (2024)
Prediction-Intervention Games and Invariant Sets
von: Kühne, Linus, et al.
Veröffentlicht: (2026)
von: Kühne, Linus, et al.
Veröffentlicht: (2026)
Many Experiments, Few Repetitions, Unpaired Data, and Sparse Effects: Is Causal Inference Possible?
von: Schur, Felix, et al.
Veröffentlicht: (2026)
von: Schur, Felix, et al.
Veröffentlicht: (2026)
Constrained Latent Action Policies for Model-Based Offline Reinforcement Learning
von: Alles, Marvin, et al.
Veröffentlicht: (2024)
von: Alles, Marvin, et al.
Veröffentlicht: (2024)
HIQL: Offline Goal-Conditioned RL with Latent States as Actions
von: Park, Seohong, et al.
Veröffentlicht: (2023)
von: Park, Seohong, et al.
Veröffentlicht: (2023)
Leveraging Offline Data in Linear Latent Contextual Bandits
von: Kausik, Chinmaya, et al.
Veröffentlicht: (2024)
von: Kausik, Chinmaya, et al.
Veröffentlicht: (2024)
CLAM: Continuous Latent Action Models for Robot Learning from Unlabeled Demonstrations
von: Liang, Anthony, et al.
Veröffentlicht: (2025)
von: Liang, Anthony, et al.
Veröffentlicht: (2025)
SAMG: Offline-to-Online Reinforcement Learning via State-Action-Conditional Offline Model Guidance
von: Zhang, Liyu, et al.
Veröffentlicht: (2024)
von: Zhang, Liyu, et al.
Veröffentlicht: (2024)
Penalizing Infeasible Actions and Reward Scaling in Reinforcement Learning with Offline Data
von: Kim, Jeonghye, et al.
Veröffentlicht: (2025)
von: Kim, Jeonghye, et al.
Veröffentlicht: (2025)
UNIQ: Offline Inverse Q-learning for Avoiding Undesirable Demonstrations
von: Hoang, Huy, et al.
Veröffentlicht: (2024)
von: Hoang, Huy, et al.
Veröffentlicht: (2024)
SPRINQL: Sub-optimal Demonstrations driven Offline Imitation Learning
von: Hoang, Huy, et al.
Veröffentlicht: (2024)
von: Hoang, Huy, et al.
Veröffentlicht: (2024)
Towards Robust Offline Reinforcement Learning under Diverse Data Corruption
von: Yang, Rui, et al.
Veröffentlicht: (2023)
von: Yang, Rui, et al.
Veröffentlicht: (2023)
DAWM: Diffusion Action World Models for Offline Reinforcement Learning via Action-Inferred Transitions
von: Li, Zongyue, et al.
Veröffentlicht: (2025)
von: Li, Zongyue, et al.
Veröffentlicht: (2025)
Transferring Causal Effects using Proxies
von: Iglesias-Alonso, Manuel, et al.
Veröffentlicht: (2025)
von: Iglesias-Alonso, Manuel, et al.
Veröffentlicht: (2025)
Action-Free Offline-to-Online RL via Discretised State Policies
von: Neggatu, Natinael Solomon, et al.
Veröffentlicht: (2026)
von: Neggatu, Natinael Solomon, et al.
Veröffentlicht: (2026)
Imitation Learning from Suboptimal Demonstrations via Meta-Learning An Action Ranker
von: Fan, Jiangdong, et al.
Veröffentlicht: (2024)
von: Fan, Jiangdong, et al.
Veröffentlicht: (2024)
Latent Adversarial Regularization for Offline Preference Optimization
von: Jiang, Enyi, et al.
Veröffentlicht: (2026)
von: Jiang, Enyi, et al.
Veröffentlicht: (2026)
On Sample-Efficient Offline Reinforcement Learning: Data Diversity, Posterior Sampling, and Beyond
von: Nguyen-Tang, Thanh, et al.
Veröffentlicht: (2024)
von: Nguyen-Tang, Thanh, et al.
Veröffentlicht: (2024)
Robust Policy Expansion for Offline-to-Online RL under Diverse Data Corruption
von: He, Longxiang, et al.
Veröffentlicht: (2025)
von: He, Longxiang, et al.
Veröffentlicht: (2025)
Learning What to Do and What Not To Do: Offline Imitation from Expert and Undesirable Demonstrations
von: Hoang, Huy, et al.
Veröffentlicht: (2025)
von: Hoang, Huy, et al.
Veröffentlicht: (2025)
A Unified Linear Programming Framework for Offline Reward Learning from Human Demonstrations and Feedback
von: Kim, Kihyun, et al.
Veröffentlicht: (2024)
von: Kim, Kihyun, et al.
Veröffentlicht: (2024)
Offline Reinforcement Learning with Penalized Action Noise Injection
von: Oh, JunHyeok, et al.
Veröffentlicht: (2025)
von: Oh, JunHyeok, et al.
Veröffentlicht: (2025)
Scalable Offline Model-Based RL with Action Chunks
von: Park, Kwanyoung, et al.
Veröffentlicht: (2025)
von: Park, Kwanyoung, et al.
Veröffentlicht: (2025)
Offline Learning of Controllable Diverse Behaviors
von: Petitbois, Mathieu, et al.
Veröffentlicht: (2025)
von: Petitbois, Mathieu, et al.
Veröffentlicht: (2025)
Uncertainty-based Offline Variational Bayesian Reinforcement Learning for Robustness under Diverse Data Corruptions
von: Yang, Rui, et al.
Veröffentlicht: (2024)
von: Yang, Rui, et al.
Veröffentlicht: (2024)
Learning Causally Invariant Reward Functions from Diverse Demonstrations
von: Ovinnikov, Ivan, et al.
Veröffentlicht: (2024)
von: Ovinnikov, Ivan, et al.
Veröffentlicht: (2024)
Offline-to-online Reinforcement Learning for Image-based Grasping with Scarce Demonstrations
von: Chan, Bryan, et al.
Veröffentlicht: (2024)
von: Chan, Bryan, et al.
Veröffentlicht: (2024)
Scaling Offline Model-Based RL via Jointly-Optimized World-Action Model Pretraining
von: Cheng, Jie, et al.
Veröffentlicht: (2024)
von: Cheng, Jie, et al.
Veröffentlicht: (2024)
Preferred-Action-Optimized Diffusion Policies for Offline Reinforcement Learning
von: Zhang, Tianle, et al.
Veröffentlicht: (2024)
von: Zhang, Tianle, et al.
Veröffentlicht: (2024)
Offline Diversity Maximization Under Imitation Constraints
von: Vlastelica, Marin, et al.
Veröffentlicht: (2023)
von: Vlastelica, Marin, et al.
Veröffentlicht: (2023)
Offline Behavioral Data Selection
von: Lei, Shiye, et al.
Veröffentlicht: (2025)
von: Lei, Shiye, et al.
Veröffentlicht: (2025)
Offline Reinforcement Learning with OOD State Correction and OOD Action Suppression
von: Mao, Yixiu, et al.
Veröffentlicht: (2024)
von: Mao, Yixiu, et al.
Veröffentlicht: (2024)
Two-Step Offline Preference-Based Reinforcement Learning with Constrained Actions
von: Xu, Yinglun, et al.
Veröffentlicht: (2023)
von: Xu, Yinglun, et al.
Veröffentlicht: (2023)
Guided Flow Policy: Learning from High-Value Actions in Offline Reinforcement Learning
von: Tiofack, Franki Nguimatsia, et al.
Veröffentlicht: (2025)
von: Tiofack, Franki Nguimatsia, et al.
Veröffentlicht: (2025)
Tackling Data Corruption in Offline Reinforcement Learning via Sequence Modeling
von: Xu, Jiawei, et al.
Veröffentlicht: (2024)
von: Xu, Jiawei, et al.
Veröffentlicht: (2024)
Disentangling Policy from Offline Task Representation Learning via Adversarial Data Augmentation
von: Jia, Chengxing, et al.
Veröffentlicht: (2024)
von: Jia, Chengxing, et al.
Veröffentlicht: (2024)
Diversity By Design: Leveraging Distribution Matching for Offline Model-Based Optimization
von: Yao, Michael S., et al.
Veröffentlicht: (2025)
von: Yao, Michael S., et al.
Veröffentlicht: (2025)
Counterfactual Identifiability via Dynamic Optimal Transport
von: Ribeiro, Fabio De Sousa, et al.
Veröffentlicht: (2025)
von: Ribeiro, Fabio De Sousa, et al.
Veröffentlicht: (2025)
Skill Learning via Policy Diversity Yields Identifiable Representations for Reinforcement Learning
von: Reizinger, Patrik, et al.
Veröffentlicht: (2025)
von: Reizinger, Patrik, et al.
Veröffentlicht: (2025)
Automatic Reward Shaping from Confounded Offline Data
von: Li, Mingxuan, et al.
Veröffentlicht: (2025)
von: Li, Mingxuan, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
How to Leverage Diverse Demonstrations in Offline Imitation Learning
von: Yue, Sheng, et al.
Veröffentlicht: (2024) -
Prediction-Intervention Games and Invariant Sets
von: Kühne, Linus, et al.
Veröffentlicht: (2026) -
Many Experiments, Few Repetitions, Unpaired Data, and Sparse Effects: Is Causal Inference Possible?
von: Schur, Felix, et al.
Veröffentlicht: (2026) -
Constrained Latent Action Policies for Model-Based Offline Reinforcement Learning
von: Alles, Marvin, et al.
Veröffentlicht: (2024) -
HIQL: Offline Goal-Conditioned RL with Latent States as Actions
von: Park, Seohong, et al.
Veröffentlicht: (2023)