Offline Imitation Learning by Controlling the Effective Planning Horizon
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Ahn, Hee-Jun, Shim, Seong-Woong, Lee, Byung-Jun |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Prior-Guided Diffusion Planning for Offline Reinforcement Learning
von: Ki, Donghyeon, et al.
Veröffentlicht: (2025)
von: Ki, Donghyeon, et al.
Veröffentlicht: (2025)
NBDI: A Simple and Effective Termination Condition for Skill Extraction from Task-Agnostic Demonstrations
von: Kim, Myunsoo, et al.
Veröffentlicht: (2025)
von: Kim, Myunsoo, et al.
Veröffentlicht: (2025)
Adaptive Non-uniform Timestep Sampling for Accelerating Diffusion Model Training
von: Kim, Myunsoo, et al.
Veröffentlicht: (2024)
von: Kim, Myunsoo, et al.
Veröffentlicht: (2024)
Offline Reinforcement Learning with Penalized Action Noise Injection
von: Oh, JunHyeok, et al.
Veröffentlicht: (2025)
von: Oh, JunHyeok, et al.
Veröffentlicht: (2025)
Semi-gradient DICE for Offline Constrained Reinforcement Learning
von: Kim, Woosung, et al.
Veröffentlicht: (2025)
von: Kim, Woosung, et al.
Veröffentlicht: (2025)
Actor-Critic without Actor
von: Ki, Donghyeon, et al.
Veröffentlicht: (2025)
von: Ki, Donghyeon, et al.
Veröffentlicht: (2025)
Direct Soft-Policy Sampling via Langevin Dynamics
von: Ki, Donghyeon, et al.
Veröffentlicht: (2026)
von: Ki, Donghyeon, et al.
Veröffentlicht: (2026)
Score-Based One-step MeanFlow Policy Optimization
von: Kim, Kyungyoon, et al.
Veröffentlicht: (2026)
von: Kim, Kyungyoon, et al.
Veröffentlicht: (2026)
FairDICE: Fairness-Driven Offline Multi-Objective Reinforcement Learning
von: Kim, Woosung, et al.
Veröffentlicht: (2025)
von: Kim, Woosung, et al.
Veröffentlicht: (2025)
Beyond RAG vs. Long-Context: Learning Distraction-Aware Retrieval for Efficient Knowledge Grounding
von: Shim, Seongwoong, et al.
Veröffentlicht: (2025)
von: Shim, Seongwoong, et al.
Veröffentlicht: (2025)
OffSim: Offline Simulator for Model-based Offline Inverse Reinforcement Learning
von: Ahn, Woo-Jin, et al.
Veröffentlicht: (2025)
von: Ahn, Woo-Jin, et al.
Veröffentlicht: (2025)
Offline Reinforcement Learning with Universal Horizon Models
von: Chung, Hojun, et al.
Veröffentlicht: (2026)
von: Chung, Hojun, et al.
Veröffentlicht: (2026)
Learning for Long-Horizon Planning via Neuro-Symbolic Abductive Imitation
von: Shao, Jie-Jing, et al.
Veröffentlicht: (2024)
von: Shao, Jie-Jing, et al.
Veröffentlicht: (2024)
Planning Transformer: Long-Horizon Offline Reinforcement Learning with Planning Tokens
von: Clinton, Joseph, et al.
Veröffentlicht: (2024)
von: Clinton, Joseph, et al.
Veröffentlicht: (2024)
Offline Imitation Learning with Variational Counterfactual Reasoning
von: He, Bowei, et al.
Veröffentlicht: (2023)
von: He, Bowei, et al.
Veröffentlicht: (2023)
Efficient Offline Reinforcement Learning: First Imitate, then Improve
von: Jelley, Adam, et al.
Veröffentlicht: (2024)
von: Jelley, Adam, et al.
Veröffentlicht: (2024)
DITTO: Offline Imitation Learning with World Models
von: DeMoss, Branton, et al.
Veröffentlicht: (2023)
von: DeMoss, Branton, et al.
Veröffentlicht: (2023)
TABX: A High-Throughput Sandbox Battle Simulator for Multi-Agent Reinforcement Learning
von: Lee, Hayeong, et al.
Veröffentlicht: (2026)
von: Lee, Hayeong, et al.
Veröffentlicht: (2026)
Guided Data Augmentation for Offline Reinforcement Learning and Imitation Learning
von: Corrado, Nicholas E., et al.
Veröffentlicht: (2023)
von: Corrado, Nicholas E., et al.
Veröffentlicht: (2023)
Robust Offline Imitation Learning from Diverse Auxiliary Data
von: Ghosh, Udita, et al.
Veröffentlicht: (2024)
von: Ghosh, Udita, et al.
Veröffentlicht: (2024)
Zero-Shot Offline Imitation Learning via Optimal Transport
von: Rupf, Thomas, et al.
Veröffentlicht: (2024)
von: Rupf, Thomas, et al.
Veröffentlicht: (2024)
Phase-Amplitude Reduction-Based Imitation Learning
von: Yamamori, Satoshi, et al.
Veröffentlicht: (2024)
von: Yamamori, Satoshi, et al.
Veröffentlicht: (2024)
Horizon Reduction as Information Loss in Offline Reinforcement Learning
von: Nidadala, Uday Kumar, et al.
Veröffentlicht: (2025)
von: Nidadala, Uday Kumar, et al.
Veröffentlicht: (2025)
Offline Imitation Learning with Model-based Reverse Augmentation
von: Shao, Jie-Jing, et al.
Veröffentlicht: (2024)
von: Shao, Jie-Jing, et al.
Veröffentlicht: (2024)
How to Leverage Diverse Demonstrations in Offline Imitation Learning
von: Yue, Sheng, et al.
Veröffentlicht: (2024)
von: Yue, Sheng, et al.
Veröffentlicht: (2024)
SPIRE: Synergistic Planning, Imitation, and Reinforcement Learning for Long-Horizon Manipulation
von: Zhou, Zihan, et al.
Veröffentlicht: (2024)
von: Zhou, Zihan, et al.
Veröffentlicht: (2024)
Self-Supervised Interpretable End-to-End Learning via Latent Functional Modularity
von: Seong, Hyunki, et al.
Veröffentlicht: (2024)
von: Seong, Hyunki, et al.
Veröffentlicht: (2024)
From Novelty to Imitation: Self-Distilled Rewards for Offline Reinforcement Learning
von: Chaudhary, Gaurav, et al.
Veröffentlicht: (2025)
von: Chaudhary, Gaurav, et al.
Veröffentlicht: (2025)
Offline Imitation Learning Through Graph Search and Retrieval
von: Yin, Zhao-Heng, et al.
Veröffentlicht: (2024)
von: Yin, Zhao-Heng, et al.
Veröffentlicht: (2024)
OLLIE: Imitation Learning from Offline Pretraining to Online Finetuning
von: Yue, Sheng, et al.
Veröffentlicht: (2024)
von: Yue, Sheng, et al.
Veröffentlicht: (2024)
SPRINQL: Sub-optimal Demonstrations driven Offline Imitation Learning
von: Hoang, Huy, et al.
Veröffentlicht: (2024)
von: Hoang, Huy, et al.
Veröffentlicht: (2024)
Balance Equation-based Distributionally Robust Offline Imitation Learning
von: Agrawal, Rishabh, et al.
Veröffentlicht: (2025)
von: Agrawal, Rishabh, et al.
Veröffentlicht: (2025)
Long-Horizon Visual Imitation Learning via Plan and Code Reflection
von: Chen, Quan, et al.
Veröffentlicht: (2025)
von: Chen, Quan, et al.
Veröffentlicht: (2025)
C-GAIL: Stabilizing Generative Adversarial Imitation Learning with Control Theory
von: Luo, Tianjiao, et al.
Veröffentlicht: (2024)
von: Luo, Tianjiao, et al.
Veröffentlicht: (2024)
On the Sample Complexity of Imitation Learning for Smoothed Model Predictive Control
von: Pfrommer, Daniel, et al.
Veröffentlicht: (2023)
von: Pfrommer, Daniel, et al.
Veröffentlicht: (2023)
Diffusion-Based Offline RL for Improved Decision-Making in Augmented ARC Task
von: Kim, Yunho, et al.
Veröffentlicht: (2024)
von: Kim, Yunho, et al.
Veröffentlicht: (2024)
Inverse Q-Learning Done Right: Offline Imitation Learning in $Q^π$-Realizable MDPs
von: Moulin, Antoine, et al.
Veröffentlicht: (2025)
von: Moulin, Antoine, et al.
Veröffentlicht: (2025)
SEABO: A Simple Search-Based Method for Offline Imitation Learning
von: Lyu, Jiafei, et al.
Veröffentlicht: (2024)
von: Lyu, Jiafei, et al.
Veröffentlicht: (2024)
Align Your Intents: Offline Imitation Learning via Optimal Transport
von: Bobrin, Maksim, et al.
Veröffentlicht: (2024)
von: Bobrin, Maksim, et al.
Veröffentlicht: (2024)
Offline Imitation Learning from Multiple Baselines with Applications to Compiler Optimization
von: Marinov, Teodor V., et al.
Veröffentlicht: (2024)
von: Marinov, Teodor V., et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Prior-Guided Diffusion Planning for Offline Reinforcement Learning
von: Ki, Donghyeon, et al.
Veröffentlicht: (2025) -
NBDI: A Simple and Effective Termination Condition for Skill Extraction from Task-Agnostic Demonstrations
von: Kim, Myunsoo, et al.
Veröffentlicht: (2025) -
Adaptive Non-uniform Timestep Sampling for Accelerating Diffusion Model Training
von: Kim, Myunsoo, et al.
Veröffentlicht: (2024) -
Offline Reinforcement Learning with Penalized Action Noise Injection
von: Oh, JunHyeok, et al.
Veröffentlicht: (2025) -
Semi-gradient DICE for Offline Constrained Reinforcement Learning
von: Kim, Woosung, et al.
Veröffentlicht: (2025)