IN-RIL: Interleaved Reinforcement and Imitation Learning for Policy Fine-Tuning
Fuente:
arXiv
Saved in:
| Main Authors: | Gao, Dechen, Wang, Hang, Zhou, Hanchu, Ammar, Nejib, Mishra, Shatadal, Moradipari, Ahmadreza, Soltani, Iman, Zhang, Junshan |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Network-Efficient World Model Token Streaming
by: Mishra, Shatadal, et al.
Published: (2026)
by: Mishra, Shatadal, et al.
Published: (2026)
CarDreamer: Open-Source Learning Platform for World Model based Autonomous Driving
by: Gao, Dechen, et al.
Published: (2024)
by: Gao, Dechen, et al.
Published: (2024)
VITA: Vision-to-Action Flow Matching Policy
by: Gao, Dechen, et al.
Published: (2025)
by: Gao, Dechen, et al.
Published: (2025)
Ego-centric Learning of Communicative World Models for Autonomous Driving
by: Wang, Hang, et al.
Published: (2025)
by: Wang, Hang, et al.
Published: (2025)
MarineFormer: A Spatio-Temporal Attention Model for USV Navigation in Dynamic Marine Environments
by: Kazemi, Ehsan, et al.
Published: (2024)
by: Kazemi, Ehsan, et al.
Published: (2024)
Tolerance of Reinforcement Learning Controllers against Deviations in Cyber Physical Systems
by: Zhang, Changjian, et al.
Published: (2024)
by: Zhang, Changjian, et al.
Published: (2024)
On-Policy Distillation of Language Models for Autonomous Vehicle Motion Planning
by: Afsharrad, Amirhossein, et al.
Published: (2026)
by: Afsharrad, Amirhossein, et al.
Published: (2026)
Gaze on the Prize: Shaping Visual Attention with Return-Guided Contrastive Learning
by: Lee, Andrew, et al.
Published: (2025)
by: Lee, Andrew, et al.
Published: (2025)
Agentic AI for Trip Planning Optimization Application
by: Chen, Tiejin, et al.
Published: (2026)
by: Chen, Tiejin, et al.
Published: (2026)
Look, Focus, Act: Efficient and Robust Robot Learning via Human Gaze and Foveated Vision Transformers
by: Chuang, Ian, et al.
Published: (2025)
by: Chuang, Ian, et al.
Published: (2025)
Symbolic Imitation Learning: From Black-Box to Explainable Driving Policies
by: Sharifi, Iman, et al.
Published: (2023)
by: Sharifi, Iman, et al.
Published: (2023)
Physics-informed Imitative Reinforcement Learning for Real-world Driving
by: Zhou, Hang, et al.
Published: (2024)
by: Zhou, Hang, et al.
Published: (2024)
GenAI-based Multi-Agent Reinforcement Learning towards Distributed Agent Intelligence: A Generative-RL Agent Perspective
by: Wang, Hang, et al.
Published: (2025)
by: Wang, Hang, et al.
Published: (2025)
SAIL: Faster-than-Demonstration Execution of Imitation Learning Policies
by: Arachchige, Nadun Ranawaka, et al.
Published: (2025)
by: Arachchige, Nadun Ranawaka, et al.
Published: (2025)
Fine-Tuning Large Language Models for Cooperative Tactical Deconfliction of Small Unmanned Aerial Systems
by: Sharifi, Iman, et al.
Published: (2026)
by: Sharifi, Iman, et al.
Published: (2026)
EI-Drive: A Platform for Cooperative Perception with Realistic Communication Models
by: Zhou, Hanchu, et al.
Published: (2024)
by: Zhou, Hanchu, et al.
Published: (2024)
RLIF: Interactive Imitation Learning as Reinforcement Learning
by: Luo, Jianlan, et al.
Published: (2023)
by: Luo, Jianlan, et al.
Published: (2023)
Out-of-Distribution Recovery with Object-Centric Keypoint Inverse Policy for Visuomotor Imitation Learning
by: Gao, George Jiayuan, et al.
Published: (2024)
by: Gao, George Jiayuan, et al.
Published: (2024)
Fine-tuning is Not Enough: A Parallel Framework for Collaborative Imitation and Reinforcement Learning in End-to-end Autonomous Driving
by: Lian, Zhexi, et al.
Published: (2026)
by: Lian, Zhexi, et al.
Published: (2026)
Cosmos Policy: Fine-Tuning Video Models for Visuomotor Control and Planning
by: Kim, Moo Jin, et al.
Published: (2026)
by: Kim, Moo Jin, et al.
Published: (2026)
Integrating Offline Pre-Training with Online Fine-Tuning: A Reinforcement Learning Approach for Robot Social Navigation
by: Su, Run, et al.
Published: (2025)
by: Su, Run, et al.
Published: (2025)
RILe: Reinforced Imitation Learning
by: Albaba, Mert, et al.
Published: (2024)
by: Albaba, Mert, et al.
Published: (2024)
NuPlanQA: A Large-Scale Dataset and Benchmark for Multi-View Driving Scene Understanding in Multi-Modal Large Language Models
by: Park, Sung-Yeon, et al.
Published: (2025)
by: Park, Sung-Yeon, et al.
Published: (2025)
Improved Bayesian Regret Bounds for Thompson Sampling in Reinforcement Learning
by: Moradipari, Ahmadreza, et al.
Published: (2023)
by: Moradipari, Ahmadreza, et al.
Published: (2023)
ConRFT: A Reinforced Fine-tuning Method for VLA Models via Consistency Policy
by: Chen, Yuhui, et al.
Published: (2025)
by: Chen, Yuhui, et al.
Published: (2025)
I-CTRL: Imitation to Control Humanoid Robots Through Constrained Reinforcement Learning
by: Yan, Yashuai, et al.
Published: (2024)
by: Yan, Yashuai, et al.
Published: (2024)
Training Robots without Robots: Deep Imitation Learning for Master-to-Robot Policy Transfer
by: Kim, Heecheol, et al.
Published: (2022)
by: Kim, Heecheol, et al.
Published: (2022)
A Novel Framework for Learning Stochastic Representations for Sequence Generation and Recognition
by: Hwang, Jungsik, et al.
Published: (2024)
by: Hwang, Jungsik, et al.
Published: (2024)
Reinforced Imitative Trajectory Planning for Urban Automated Driving
by: Zeng, Di, et al.
Published: (2024)
by: Zeng, Di, et al.
Published: (2024)
HannesImitation: Grasping with the Hannes Prosthetic Hand via Imitation Learning
by: Alessi, Carlo, et al.
Published: (2025)
by: Alessi, Carlo, et al.
Published: (2025)
Adapt Your Body: Mitigating Proprioception Shifts in Imitation Learning
by: Kuang, Fuhang, et al.
Published: (2025)
by: Kuang, Fuhang, et al.
Published: (2025)
Trust Region Inverse Reinforcement Learning: Explicit Dual Ascent using Local Policy Updates
by: Diwan, Anish, et al.
Published: (2026)
by: Diwan, Anish, et al.
Published: (2026)
SKIL: Semantic Keypoint Imitation Learning for Generalizable Data-efficient Manipulation
by: Wang, Shengjie, et al.
Published: (2025)
by: Wang, Shengjie, et al.
Published: (2025)
Dual RL: Unification and New Methods for Reinforcement and Imitation Learning
by: Sikchi, Harshit, et al.
Published: (2023)
by: Sikchi, Harshit, et al.
Published: (2023)
TDMPBC: Self-Imitative Reinforcement Learning for Humanoid Robot Control
by: Zhuang, Zifeng, et al.
Published: (2025)
by: Zhuang, Zifeng, et al.
Published: (2025)
Implicit Kinodynamic Motion Retargeting for Human-to-humanoid Imitation Learning
by: Chen, Xingyu, et al.
Published: (2025)
by: Chen, Xingyu, et al.
Published: (2025)
LAMARL: LLM-Aided Multi-Agent Reinforcement Learning for Cooperative Policy Generation
by: Zhu, Guobin, et al.
Published: (2025)
by: Zhu, Guobin, et al.
Published: (2025)
Latent Diffusion Planning for Imitation Learning
by: Xie, Amber, et al.
Published: (2025)
by: Xie, Amber, et al.
Published: (2025)
Safe CoR: A Dual-Expert Approach to Integrating Imitation Learning and Safe Reinforcement Learning Using Constraint Rewards
by: Kwon, Hyeokjin, et al.
Published: (2024)
by: Kwon, Hyeokjin, et al.
Published: (2024)
DiffTORI: Differentiable Trajectory Optimization for Deep Reinforcement and Imitation Learning
by: Wan, Weikang, et al.
Published: (2024)
by: Wan, Weikang, et al.
Published: (2024)
Similar Items
-
Network-Efficient World Model Token Streaming
by: Mishra, Shatadal, et al.
Published: (2026) -
CarDreamer: Open-Source Learning Platform for World Model based Autonomous Driving
by: Gao, Dechen, et al.
Published: (2024) -
VITA: Vision-to-Action Flow Matching Policy
by: Gao, Dechen, et al.
Published: (2025) -
Ego-centric Learning of Communicative World Models for Autonomous Driving
by: Wang, Hang, et al.
Published: (2025) -
MarineFormer: A Spatio-Temporal Attention Model for USV Navigation in Dynamic Marine Environments
by: Kazemi, Ehsan, et al.
Published: (2024)