Beyond Imitation: Reinforcement Learning for Active Latent Planning
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zheng, Zhi, Lee, Wee Sun |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Reasoning-CV: Fine-tuning Powerful Reasoning LLMs for Knowledge-Assisted Claim Verification
von: Zheng, Zhi, et al.
Veröffentlicht: (2025)
von: Zheng, Zhi, et al.
Veröffentlicht: (2025)
Latent Diffusion Planning for Imitation Learning
von: Xie, Amber, et al.
Veröffentlicht: (2025)
von: Xie, Amber, et al.
Veröffentlicht: (2025)
Continual Reinforcement Learning by Planning with Online World Models
von: Liu, Zichen, et al.
Veröffentlicht: (2025)
von: Liu, Zichen, et al.
Veröffentlicht: (2025)
Reinforced Imitative Trajectory Planning for Urban Automated Driving
von: Zeng, Di, et al.
Veröffentlicht: (2024)
von: Zeng, Di, et al.
Veröffentlicht: (2024)
SofT-GRPO: Surpassing Discrete-Token LLM Reinforcement Learning via Gumbel-Reparameterized Soft-Thinking Policy Optimization
von: Zheng, Zhi, et al.
Veröffentlicht: (2025)
von: Zheng, Zhi, et al.
Veröffentlicht: (2025)
PLANRL: A Motion Planning and Imitation Learning Framework to Bootstrap Reinforcement Learning
von: Bhaskar, Amisha, et al.
Veröffentlicht: (2024)
von: Bhaskar, Amisha, et al.
Veröffentlicht: (2024)
Imitation Game: A Model-based and Imitation Learning Deep Reinforcement Learning Hybrid
von: Veith, Eric MSP, et al.
Veröffentlicht: (2024)
von: Veith, Eric MSP, et al.
Veröffentlicht: (2024)
Imitation Bootstrapped Reinforcement Learning
von: Hu, Hengyuan, et al.
Veröffentlicht: (2023)
von: Hu, Hengyuan, et al.
Veröffentlicht: (2023)
On the Empirical Complexity of Reasoning and Planning in LLMs
von: Kang, Liwei, et al.
Veröffentlicht: (2024)
von: Kang, Liwei, et al.
Veröffentlicht: (2024)
DARIL: When Imitation Learning outperforms Reinforcement Learning in Surgical Action Planning
von: Boels, Maxence, et al.
Veröffentlicht: (2025)
von: Boels, Maxence, et al.
Veröffentlicht: (2025)
RLIF: Interactive Imitation Learning as Reinforcement Learning
von: Luo, Jianlan, et al.
Veröffentlicht: (2023)
von: Luo, Jianlan, et al.
Veröffentlicht: (2023)
RILe: Reinforced Imitation Learning
von: Albaba, Mert, et al.
Veröffentlicht: (2024)
von: Albaba, Mert, et al.
Veröffentlicht: (2024)
Zero-shot Imitation Learning by Latent Topology Mapping
von: Jacobson, Maxwell J., et al.
Veröffentlicht: (2026)
von: Jacobson, Maxwell J., et al.
Veröffentlicht: (2026)
Dual RL: Unification and New Methods for Reinforcement and Imitation Learning
von: Sikchi, Harshit, et al.
Veröffentlicht: (2023)
von: Sikchi, Harshit, et al.
Veröffentlicht: (2023)
I-CTRL: Imitation to Control Humanoid Robots Through Constrained Reinforcement Learning
von: Yan, Yashuai, et al.
Veröffentlicht: (2024)
von: Yan, Yashuai, et al.
Veröffentlicht: (2024)
Personalized Dynamic Difficulty Adjustment -- Imitation Learning Meets Reinforcement Learning
von: Fuchs, Ronja, et al.
Veröffentlicht: (2024)
von: Fuchs, Ronja, et al.
Veröffentlicht: (2024)
Reinforcement Learning via Implicit Imitation Guidance
von: Dong, Perry, et al.
Veröffentlicht: (2025)
von: Dong, Perry, et al.
Veröffentlicht: (2025)
Imitating Cost-Constrained Behaviors in Reinforcement Learning
von: Shao, Qian, et al.
Veröffentlicht: (2024)
von: Shao, Qian, et al.
Veröffentlicht: (2024)
LatentMimic: Terrain-Adaptive Locomotion via Latent Space Imitation
von: Wang, Zhiquan, et al.
Veröffentlicht: (2026)
von: Wang, Zhiquan, et al.
Veröffentlicht: (2026)
SPIRE: Synergistic Planning, Imitation, and Reinforcement Learning for Long-Horizon Manipulation
von: Zhou, Zihan, et al.
Veröffentlicht: (2024)
von: Zhou, Zihan, et al.
Veröffentlicht: (2024)
Safe CoR: A Dual-Expert Approach to Integrating Imitation Learning and Safe Reinforcement Learning Using Constraint Rewards
von: Kwon, Hyeokjin, et al.
Veröffentlicht: (2024)
von: Kwon, Hyeokjin, et al.
Veröffentlicht: (2024)
Beyond Mimicry: Toward Lifelong Adaptability in Imitation Learning
von: Gavenski, Nathan, et al.
Veröffentlicht: (2026)
von: Gavenski, Nathan, et al.
Veröffentlicht: (2026)
Combining Reinforcement Learning and Optimal Transport for the Traveling Salesman Problem
von: Goh, Yong Liang, et al.
Veröffentlicht: (2022)
von: Goh, Yong Liang, et al.
Veröffentlicht: (2022)
ImitationNet: Unsupervised Human-to-Robot Motion Retargeting via Shared Latent Space
von: Yan, Yashuai, et al.
Veröffentlicht: (2023)
von: Yan, Yashuai, et al.
Veröffentlicht: (2023)
Model Predictive Adversarial Imitation Learning for Planning from Observation
von: Han, Tyler, et al.
Veröffentlicht: (2025)
von: Han, Tyler, et al.
Veröffentlicht: (2025)
Visual Hindsight Self-Imitation Learning for Interactive Navigation
von: Kim, Kibeom, et al.
Veröffentlicht: (2023)
von: Kim, Kibeom, et al.
Veröffentlicht: (2023)
FinFlowRL: An Imitation-Reinforcement Learning Framework for Adaptive Stochastic Control in Finance
von: Li, Yang, et al.
Veröffentlicht: (2025)
von: Li, Yang, et al.
Veröffentlicht: (2025)
IN-RIL: Interleaved Reinforcement and Imitation Learning for Policy Fine-Tuning
von: Gao, Dechen, et al.
Veröffentlicht: (2025)
von: Gao, Dechen, et al.
Veröffentlicht: (2025)
An Optimal Discriminator Weighted Imitation Perspective for Reinforcement Learning
von: Xu, Haoran, et al.
Veröffentlicht: (2025)
von: Xu, Haoran, et al.
Veröffentlicht: (2025)
Blending Imitation and Reinforcement Learning for Robust Policy Improvement
von: Liu, Xuefeng, et al.
Veröffentlicht: (2023)
von: Liu, Xuefeng, et al.
Veröffentlicht: (2023)
Differentiable Tree Search Network
von: Mittal, Dixant, et al.
Veröffentlicht: (2024)
von: Mittal, Dixant, et al.
Veröffentlicht: (2024)
Extending Epistemic Uncertainty Beyond Parameters Would Assist in Designing Reliable LLMs
von: Nguyen-Hien, T. Duy, et al.
Veröffentlicht: (2025)
von: Nguyen-Hien, T. Duy, et al.
Veröffentlicht: (2025)
Sample-efficient Adversarial Imitation Learning
von: Jung, Dahuin, et al.
Veröffentlicht: (2023)
von: Jung, Dahuin, et al.
Veröffentlicht: (2023)
Confounded Causal Imitation Learning with Instrumental Variables
von: Zeng, Yan, et al.
Veröffentlicht: (2025)
von: Zeng, Yan, et al.
Veröffentlicht: (2025)
Rewind-IL: Online Failure Detection and State Respawning for Imitation Learning
von: Zheng, Gehan, et al.
Veröffentlicht: (2026)
von: Zheng, Gehan, et al.
Veröffentlicht: (2026)
Mxplainer: Explain and Learn Insights by Imitating Mahjong Agents
von: Li, Lingfeng, et al.
Veröffentlicht: (2025)
von: Li, Lingfeng, et al.
Veröffentlicht: (2025)
TACO: Temporal Latent Action-Driven Contrastive Loss for Visual Reinforcement Learning
von: Zheng, Ruijie, et al.
Veröffentlicht: (2023)
von: Zheng, Ruijie, et al.
Veröffentlicht: (2023)
Thoughts-as-Planning: Latent World Models for Chain-of-Thoughts Optimization via Reinforcement Planning
von: Liu, Dong, et al.
Veröffentlicht: (2026)
von: Liu, Dong, et al.
Veröffentlicht: (2026)
Active Legibility in Multiagent Reinforcement Learning
von: Liu, Yanyu, et al.
Veröffentlicht: (2024)
von: Liu, Yanyu, et al.
Veröffentlicht: (2024)
Learning for Long-Horizon Planning via Neuro-Symbolic Abductive Imitation
von: Shao, Jie-Jing, et al.
Veröffentlicht: (2024)
von: Shao, Jie-Jing, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Reasoning-CV: Fine-tuning Powerful Reasoning LLMs for Knowledge-Assisted Claim Verification
von: Zheng, Zhi, et al.
Veröffentlicht: (2025) -
Latent Diffusion Planning for Imitation Learning
von: Xie, Amber, et al.
Veröffentlicht: (2025) -
Continual Reinforcement Learning by Planning with Online World Models
von: Liu, Zichen, et al.
Veröffentlicht: (2025) -
Reinforced Imitative Trajectory Planning for Urban Automated Driving
von: Zeng, Di, et al.
Veröffentlicht: (2024) -
SofT-GRPO: Surpassing Discrete-Token LLM Reinforcement Learning via Gumbel-Reparameterized Soft-Thinking Policy Optimization
von: Zheng, Zhi, et al.
Veröffentlicht: (2025)