Identifying Latent Actions and Dynamics from Offline Data via Demonstrator Diversity
Fuente:
arXiv
Guardado en:
| Autor principal: | Schur, Felix |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
How to Leverage Diverse Demonstrations in Offline Imitation Learning
por: Yue, Sheng, et al.
Publicado: (2024)
por: Yue, Sheng, et al.
Publicado: (2024)
Prediction-Intervention Games and Invariant Sets
por: Kühne, Linus, et al.
Publicado: (2026)
por: Kühne, Linus, et al.
Publicado: (2026)
Many Experiments, Few Repetitions, Unpaired Data, and Sparse Effects: Is Causal Inference Possible?
por: Schur, Felix, et al.
Publicado: (2026)
por: Schur, Felix, et al.
Publicado: (2026)
Constrained Latent Action Policies for Model-Based Offline Reinforcement Learning
por: Alles, Marvin, et al.
Publicado: (2024)
por: Alles, Marvin, et al.
Publicado: (2024)
HIQL: Offline Goal-Conditioned RL with Latent States as Actions
por: Park, Seohong, et al.
Publicado: (2023)
por: Park, Seohong, et al.
Publicado: (2023)
Leveraging Offline Data in Linear Latent Contextual Bandits
por: Kausik, Chinmaya, et al.
Publicado: (2024)
por: Kausik, Chinmaya, et al.
Publicado: (2024)
CLAM: Continuous Latent Action Models for Robot Learning from Unlabeled Demonstrations
por: Liang, Anthony, et al.
Publicado: (2025)
por: Liang, Anthony, et al.
Publicado: (2025)
SAMG: Offline-to-Online Reinforcement Learning via State-Action-Conditional Offline Model Guidance
por: Zhang, Liyu, et al.
Publicado: (2024)
por: Zhang, Liyu, et al.
Publicado: (2024)
Penalizing Infeasible Actions and Reward Scaling in Reinforcement Learning with Offline Data
por: Kim, Jeonghye, et al.
Publicado: (2025)
por: Kim, Jeonghye, et al.
Publicado: (2025)
UNIQ: Offline Inverse Q-learning for Avoiding Undesirable Demonstrations
por: Hoang, Huy, et al.
Publicado: (2024)
por: Hoang, Huy, et al.
Publicado: (2024)
SPRINQL: Sub-optimal Demonstrations driven Offline Imitation Learning
por: Hoang, Huy, et al.
Publicado: (2024)
por: Hoang, Huy, et al.
Publicado: (2024)
Towards Robust Offline Reinforcement Learning under Diverse Data Corruption
por: Yang, Rui, et al.
Publicado: (2023)
por: Yang, Rui, et al.
Publicado: (2023)
DAWM: Diffusion Action World Models for Offline Reinforcement Learning via Action-Inferred Transitions
por: Li, Zongyue, et al.
Publicado: (2025)
por: Li, Zongyue, et al.
Publicado: (2025)
Transferring Causal Effects using Proxies
por: Iglesias-Alonso, Manuel, et al.
Publicado: (2025)
por: Iglesias-Alonso, Manuel, et al.
Publicado: (2025)
Action-Free Offline-to-Online RL via Discretised State Policies
por: Neggatu, Natinael Solomon, et al.
Publicado: (2026)
por: Neggatu, Natinael Solomon, et al.
Publicado: (2026)
Imitation Learning from Suboptimal Demonstrations via Meta-Learning An Action Ranker
por: Fan, Jiangdong, et al.
Publicado: (2024)
por: Fan, Jiangdong, et al.
Publicado: (2024)
Latent Adversarial Regularization for Offline Preference Optimization
por: Jiang, Enyi, et al.
Publicado: (2026)
por: Jiang, Enyi, et al.
Publicado: (2026)
On Sample-Efficient Offline Reinforcement Learning: Data Diversity, Posterior Sampling, and Beyond
por: Nguyen-Tang, Thanh, et al.
Publicado: (2024)
por: Nguyen-Tang, Thanh, et al.
Publicado: (2024)
Robust Policy Expansion for Offline-to-Online RL under Diverse Data Corruption
por: He, Longxiang, et al.
Publicado: (2025)
por: He, Longxiang, et al.
Publicado: (2025)
Learning What to Do and What Not To Do: Offline Imitation from Expert and Undesirable Demonstrations
por: Hoang, Huy, et al.
Publicado: (2025)
por: Hoang, Huy, et al.
Publicado: (2025)
A Unified Linear Programming Framework for Offline Reward Learning from Human Demonstrations and Feedback
por: Kim, Kihyun, et al.
Publicado: (2024)
por: Kim, Kihyun, et al.
Publicado: (2024)
Offline Reinforcement Learning with Penalized Action Noise Injection
por: Oh, JunHyeok, et al.
Publicado: (2025)
por: Oh, JunHyeok, et al.
Publicado: (2025)
Scalable Offline Model-Based RL with Action Chunks
por: Park, Kwanyoung, et al.
Publicado: (2025)
por: Park, Kwanyoung, et al.
Publicado: (2025)
Offline Learning of Controllable Diverse Behaviors
por: Petitbois, Mathieu, et al.
Publicado: (2025)
por: Petitbois, Mathieu, et al.
Publicado: (2025)
Uncertainty-based Offline Variational Bayesian Reinforcement Learning for Robustness under Diverse Data Corruptions
por: Yang, Rui, et al.
Publicado: (2024)
por: Yang, Rui, et al.
Publicado: (2024)
Learning Causally Invariant Reward Functions from Diverse Demonstrations
por: Ovinnikov, Ivan, et al.
Publicado: (2024)
por: Ovinnikov, Ivan, et al.
Publicado: (2024)
Offline-to-online Reinforcement Learning for Image-based Grasping with Scarce Demonstrations
por: Chan, Bryan, et al.
Publicado: (2024)
por: Chan, Bryan, et al.
Publicado: (2024)
Scaling Offline Model-Based RL via Jointly-Optimized World-Action Model Pretraining
por: Cheng, Jie, et al.
Publicado: (2024)
por: Cheng, Jie, et al.
Publicado: (2024)
Preferred-Action-Optimized Diffusion Policies for Offline Reinforcement Learning
por: Zhang, Tianle, et al.
Publicado: (2024)
por: Zhang, Tianle, et al.
Publicado: (2024)
Offline Diversity Maximization Under Imitation Constraints
por: Vlastelica, Marin, et al.
Publicado: (2023)
por: Vlastelica, Marin, et al.
Publicado: (2023)
Offline Behavioral Data Selection
por: Lei, Shiye, et al.
Publicado: (2025)
por: Lei, Shiye, et al.
Publicado: (2025)
Offline Reinforcement Learning with OOD State Correction and OOD Action Suppression
por: Mao, Yixiu, et al.
Publicado: (2024)
por: Mao, Yixiu, et al.
Publicado: (2024)
Two-Step Offline Preference-Based Reinforcement Learning with Constrained Actions
por: Xu, Yinglun, et al.
Publicado: (2023)
por: Xu, Yinglun, et al.
Publicado: (2023)
Guided Flow Policy: Learning from High-Value Actions in Offline Reinforcement Learning
por: Tiofack, Franki Nguimatsia, et al.
Publicado: (2025)
por: Tiofack, Franki Nguimatsia, et al.
Publicado: (2025)
Tackling Data Corruption in Offline Reinforcement Learning via Sequence Modeling
por: Xu, Jiawei, et al.
Publicado: (2024)
por: Xu, Jiawei, et al.
Publicado: (2024)
Disentangling Policy from Offline Task Representation Learning via Adversarial Data Augmentation
por: Jia, Chengxing, et al.
Publicado: (2024)
por: Jia, Chengxing, et al.
Publicado: (2024)
Diversity By Design: Leveraging Distribution Matching for Offline Model-Based Optimization
por: Yao, Michael S., et al.
Publicado: (2025)
por: Yao, Michael S., et al.
Publicado: (2025)
Counterfactual Identifiability via Dynamic Optimal Transport
por: Ribeiro, Fabio De Sousa, et al.
Publicado: (2025)
por: Ribeiro, Fabio De Sousa, et al.
Publicado: (2025)
Skill Learning via Policy Diversity Yields Identifiable Representations for Reinforcement Learning
por: Reizinger, Patrik, et al.
Publicado: (2025)
por: Reizinger, Patrik, et al.
Publicado: (2025)
Automatic Reward Shaping from Confounded Offline Data
por: Li, Mingxuan, et al.
Publicado: (2025)
por: Li, Mingxuan, et al.
Publicado: (2025)
Ejemplares similares
-
How to Leverage Diverse Demonstrations in Offline Imitation Learning
por: Yue, Sheng, et al.
Publicado: (2024) -
Prediction-Intervention Games and Invariant Sets
por: Kühne, Linus, et al.
Publicado: (2026) -
Many Experiments, Few Repetitions, Unpaired Data, and Sparse Effects: Is Causal Inference Possible?
por: Schur, Felix, et al.
Publicado: (2026) -
Constrained Latent Action Policies for Model-Based Offline Reinforcement Learning
por: Alles, Marvin, et al.
Publicado: (2024) -
HIQL: Offline Goal-Conditioned RL with Latent States as Actions
por: Park, Seohong, et al.
Publicado: (2023)