Markov Balance Satisfaction Improves Performance in Strictly Batch Offline Imitation Learning
Fuente:
arXiv
Saved in:
| Main Authors: | Agrawal, Rishabh, Dahlin, Nathan, Jain, Rahul, Nayyar, Ashutosh |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Balance Equation-based Distributionally Robust Offline Imitation Learning
by: Agrawal, Rishabh, et al.
Published: (2025)
by: Agrawal, Rishabh, et al.
Published: (2025)
When Dynamics Shift, Robust Task Inference Wins: Offline Imitation Learning with Behavior Foundation Models Revisited
by: Agrawal, Rishabh, et al.
Published: (2026)
by: Agrawal, Rishabh, et al.
Published: (2026)
Posterior Sampling-based Online Learning for Episodic POMDPs
by: Tang, Dengwang, et al.
Published: (2023)
by: Tang, Dengwang, et al.
Published: (2023)
Adaptive Few-Shot Learning (AFSL): Tackling Data Scarcity with Stability, Robustness, and Versatility
by: Agrawal, Rishabh
Published: (2025)
by: Agrawal, Rishabh
Published: (2025)
DITTO: Offline Imitation Learning with World Models
by: DeMoss, Branton, et al.
Published: (2023)
by: DeMoss, Branton, et al.
Published: (2023)
Using Non-Expert Data to Robustify Imitation Learning via Offline Reinforcement Learning
by: Huang, Kevin, et al.
Published: (2025)
by: Huang, Kevin, et al.
Published: (2025)
Offline Imitation Learning with Model-based Reverse Augmentation
by: Shao, Jie-Jing, et al.
Published: (2024)
by: Shao, Jie-Jing, et al.
Published: (2024)
How to Leverage Diverse Demonstrations in Offline Imitation Learning
by: Yue, Sheng, et al.
Published: (2024)
by: Yue, Sheng, et al.
Published: (2024)
FAST-Q: Fast-track Exploration with Adversarially Balanced State Representations for Counterfactual Action Estimation in Offline Reinforcement Learning
by: Agrawal, Pulkit, et al.
Published: (2025)
by: Agrawal, Pulkit, et al.
Published: (2025)
Quantifying Generalisation in Imitation Learning
by: Gavenski, Nathan, et al.
Published: (2025)
by: Gavenski, Nathan, et al.
Published: (2025)
OLLIE: Imitation Learning from Offline Pretraining to Online Finetuning
by: Yue, Sheng, et al.
Published: (2024)
by: Yue, Sheng, et al.
Published: (2024)
SPRINQL: Sub-optimal Demonstrations driven Offline Imitation Learning
by: Hoang, Huy, et al.
Published: (2024)
by: Hoang, Huy, et al.
Published: (2024)
Offline Imitation Learning Through Graph Search and Retrieval
by: Yin, Zhao-Heng, et al.
Published: (2024)
by: Yin, Zhao-Heng, et al.
Published: (2024)
SEABO: A Simple Search-Based Method for Offline Imitation Learning
by: Lyu, Jiafei, et al.
Published: (2024)
by: Lyu, Jiafei, et al.
Published: (2024)
Align Your Intents: Offline Imitation Learning via Optimal Transport
by: Bobrin, Maksim, et al.
Published: (2024)
by: Bobrin, Maksim, et al.
Published: (2024)
A Dual Approach to Imitation Learning from Observations with Offline Datasets
by: Sikchi, Harshit, et al.
Published: (2024)
by: Sikchi, Harshit, et al.
Published: (2024)
Offline Diversity Maximization Under Imitation Constraints
by: Vlastelica, Marin, et al.
Published: (2023)
by: Vlastelica, Marin, et al.
Published: (2023)
PROF: An LLM-based Reward Code Preference Optimization Framework for Offline Imitation Learning
by: Sun, Shengjie, et al.
Published: (2025)
by: Sun, Shengjie, et al.
Published: (2025)
SafeMIL: Learning Offline Safe Imitation Policy from Non-Preferred Trajectories
by: Burnwal, Returaj, et al.
Published: (2025)
by: Burnwal, Returaj, et al.
Published: (2025)
Learning What to Do and What Not To Do: Offline Imitation from Expert and Undesirable Demonstrations
by: Hoang, Huy, et al.
Published: (2025)
by: Hoang, Huy, et al.
Published: (2025)
Beyond Mimicry: Toward Lifelong Adaptability in Imitation Learning
by: Gavenski, Nathan, et al.
Published: (2026)
by: Gavenski, Nathan, et al.
Published: (2026)
Complement Submodular Information Measures for Balanced and Robust Data Selection
by: Iyer, Rishabh
Published: (2026)
by: Iyer, Rishabh
Published: (2026)
Measurement Scheduling for ICU Patients with Offline Reinforcement Learning
by: Ji, Zongliang, et al.
Published: (2024)
by: Ji, Zongliang, et al.
Published: (2024)
A Bayesian Learning Algorithm for Unknown Zero-sum Stochastic Games with an Arbitrary Opponent
by: Jafarnia-Jahromi, Mehdi, et al.
Published: (2021)
by: Jafarnia-Jahromi, Mehdi, et al.
Published: (2021)
OSIL: Learning Offline Safe Imitation Policies with Safety Inferred from Non-preferred Trajectories
by: Burnwal, Returaj, et al.
Published: (2026)
by: Burnwal, Returaj, et al.
Published: (2026)
A Survey of Imitation Learning Methods, Environments and Metrics
by: Gavenski, Nathan, et al.
Published: (2024)
by: Gavenski, Nathan, et al.
Published: (2024)
Context-Sensitive Abstractions for Reinforcement Learning with Parameterized Actions
by: Nayyar, Rashmeet Kaur, et al.
Published: (2025)
by: Nayyar, Rashmeet Kaur, et al.
Published: (2025)
Imitation Learning Datasets: A Toolkit For Creating Datasets, Training Agents and Benchmarking
by: Gavenski, Nathan, et al.
Published: (2024)
by: Gavenski, Nathan, et al.
Published: (2024)
Offline Imitation of Badminton Player Behavior via Experiential Contexts and Brownian Motion
by: Wang, Kuang-Da, et al.
Published: (2024)
by: Wang, Kuang-Da, et al.
Published: (2024)
Offline Imitation from Observation via Primal Wasserstein State Occupancy Matching
by: Yan, Kai, et al.
Published: (2023)
by: Yan, Kai, et al.
Published: (2023)
Towards Generalisable Imitation Learning Through Conditioned Transition Estimation and Online Behaviour Alignment
by: Gavenski, Nathan, et al.
Published: (2026)
by: Gavenski, Nathan, et al.
Published: (2026)
Pure Exploration for Constrained Best Mixed Arm Identification with a Fixed Budget
by: Tang, Dengwang, et al.
Published: (2024)
by: Tang, Dengwang, et al.
Published: (2024)
Efficient Online Learning with Offline Datasets for Infinite Horizon MDPs: A Bayesian Approach
by: Tang, Dengwang, et al.
Published: (2023)
by: Tang, Dengwang, et al.
Published: (2023)
Dual-Force: Enhanced Offline Diversity Maximization under Imitation Constraints
by: Kolev, Pavel, et al.
Published: (2025)
by: Kolev, Pavel, et al.
Published: (2025)
Explorative Imitation Learning: A Path Signature Approach for Continuous Environments
by: Gavenski, Nathan, et al.
Published: (2024)
by: Gavenski, Nathan, et al.
Published: (2024)
Improving Offline Reinforcement Learning with Inaccurate Simulators
by: Hou, Yiwen, et al.
Published: (2024)
by: Hou, Yiwen, et al.
Published: (2024)
Seesaw: Accelerating Training by Balancing Learning Rate and Batch Size Scheduling
by: Meterez, Alexandru, et al.
Published: (2025)
by: Meterez, Alexandru, et al.
Published: (2025)
From Imitation to Optimization: A Comparative Study of Offline Learning for Autonomous Driving
by: Guillen-Perez, Antonio
Published: (2025)
by: Guillen-Perez, Antonio
Published: (2025)
Batch Bayesian Active Learning with Partial Batch Label Sampling
by: Hu, Kangping, et al.
Published: (2025)
by: Hu, Kangping, et al.
Published: (2025)
Dimensional Balance Improves Large Scale Spatiotemporal Prediction Performance
by: Chen, Jing, et al.
Published: (2026)
by: Chen, Jing, et al.
Published: (2026)
Similar Items
-
Balance Equation-based Distributionally Robust Offline Imitation Learning
by: Agrawal, Rishabh, et al.
Published: (2025) -
When Dynamics Shift, Robust Task Inference Wins: Offline Imitation Learning with Behavior Foundation Models Revisited
by: Agrawal, Rishabh, et al.
Published: (2026) -
Posterior Sampling-based Online Learning for Episodic POMDPs
by: Tang, Dengwang, et al.
Published: (2023) -
Adaptive Few-Shot Learning (AFSL): Tackling Data Scarcity with Stability, Robustness, and Versatility
by: Agrawal, Rishabh
Published: (2025) -
DITTO: Offline Imitation Learning with World Models
by: DeMoss, Branton, et al.
Published: (2023)