From Prior to Pro: Efficient Skill Mastery via Distribution Contractive RL Finetuning
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Sun, Zhanyi, Song, Shuran |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Latent Policy Barrier: Learning Robust Visuomotor Policies by Staying In-Distribution
von: Sun, Zhanyi, et al.
Veröffentlicht: (2025)
von: Sun, Zhanyi, et al.
Veröffentlicht: (2025)
Posterior Behavioral Cloning: Pretraining BC Policies for Efficient RL Finetuning
von: Wagenmaker, Andrew, et al.
Veröffentlicht: (2025)
von: Wagenmaker, Andrew, et al.
Veröffentlicht: (2025)
Residual Off-Policy RL for Finetuning Behavior Cloning Policies
von: Ankile, Lars, et al.
Veröffentlicht: (2025)
von: Ankile, Lars, et al.
Veröffentlicht: (2025)
RL-VLM-F: Reinforcement Learning from Vision Language Foundation Model Feedback
von: Wang, Yufei, et al.
Veröffentlicht: (2024)
von: Wang, Yufei, et al.
Veröffentlicht: (2024)
EquiBot: SIM(3)-Equivariant Diffusion Policy for Generalizable and Data Efficient Learning
von: Yang, Jingyun, et al.
Veröffentlicht: (2024)
von: Yang, Jingyun, et al.
Veröffentlicht: (2024)
OGPO: Sample Efficient Full-Finetuning of Generative Control Policies
von: Patil, Sarvesh, et al.
Veröffentlicht: (2026)
von: Patil, Sarvesh, et al.
Veröffentlicht: (2026)
Dynamics-Guided Diffusion Model for Sensor-less Robot Manipulator Design
von: Xu, Xiaomeng, et al.
Veröffentlicht: (2024)
von: Xu, Xiaomeng, et al.
Veröffentlicht: (2024)
Compliant Residual DAgger: Improving Real-World Contact-Rich Manipulation with Human Corrections
von: Xu, Xiaomeng, et al.
Veröffentlicht: (2025)
von: Xu, Xiaomeng, et al.
Veröffentlicht: (2025)
Dual-Granularity Contrastive Reward via Generated Episodic Guidance for Efficient Embodied RL
von: Liu, Xin, et al.
Veröffentlicht: (2026)
von: Liu, Xin, et al.
Veröffentlicht: (2026)
Learn Where Outcomes Diverge: Efficient VLA RL via Probabilistic Chunk Masking
von: Bagaria, Vaidehi, et al.
Veröffentlicht: (2026)
von: Bagaria, Vaidehi, et al.
Veröffentlicht: (2026)
Q-Guided Stein Variational Model Predictive Control via RL-informed Policy Prior
von: Cai, Shizhe, et al.
Veröffentlicht: (2025)
von: Cai, Shizhe, et al.
Veröffentlicht: (2025)
DoughNet: A Visual Predictive Model for Topological Manipulation of Deformable Objects
von: Bauer, Dominik, et al.
Veröffentlicht: (2024)
von: Bauer, Dominik, et al.
Veröffentlicht: (2024)
Towards Efficient and Expressive Offline RL via Flow-Anchored Noise-conditioned Q-Learning
von: Lee, Sungyoung, et al.
Veröffentlicht: (2026)
von: Lee, Sungyoung, et al.
Veröffentlicht: (2026)
Efficient Language-instructed Skill Acquisition via Reward-Policy Co-Evolution
von: Huang, Changxin, et al.
Veröffentlicht: (2024)
von: Huang, Changxin, et al.
Veröffentlicht: (2024)
From Imitation to Refinement -- Residual RL for Precise Assembly
von: Ankile, Lars, et al.
Veröffentlicht: (2024)
von: Ankile, Lars, et al.
Veröffentlicht: (2024)
Solving Robotics Tasks with Prior Demonstration via Exploration-Efficient Deep Reinforcement Learning
von: Shen, Chengyandan, et al.
Veröffentlicht: (2025)
von: Shen, Chengyandan, et al.
Veröffentlicht: (2025)
Contractive Dynamical Imitation Policies for Efficient Out-of-Sample Recovery
von: Abyaneh, Amin, et al.
Veröffentlicht: (2024)
von: Abyaneh, Amin, et al.
Veröffentlicht: (2024)
Tactile-based Object Retrieval From Granular Media
von: Xu, Jingxi, et al.
Veröffentlicht: (2024)
von: Xu, Jingxi, et al.
Veröffentlicht: (2024)
Contractive Diffusion Policies: Robust Action Diffusion via Contractive Score-Based Sampling with Differential Equations
von: Abyaneh, Amin, et al.
Veröffentlicht: (2026)
von: Abyaneh, Amin, et al.
Veröffentlicht: (2026)
Refined Policy Distillation: From VLA Generalists to RL Experts
von: Jülg, Tobias, et al.
Veröffentlicht: (2025)
von: Jülg, Tobias, et al.
Veröffentlicht: (2025)
FLoRA: Sample-Efficient Preference-based RL via Low-Rank Style Adaptation of Reward Functions
von: Marta, Daniel, et al.
Veröffentlicht: (2025)
von: Marta, Daniel, et al.
Veröffentlicht: (2025)
Computationally Efficient RL under Linear Bellman Completeness for Deterministic Dynamics
von: Wu, Runzhe, et al.
Veröffentlicht: (2024)
von: Wu, Runzhe, et al.
Veröffentlicht: (2024)
SPRINT: Efficient Spectral Priors for Humanoid Athletic Sprints
von: Wei, Yantong, et al.
Veröffentlicht: (2026)
von: Wei, Yantong, et al.
Veröffentlicht: (2026)
TWISTED-RL: Hierarchical Skilled Agents for Knot-Tying without Human Demonstrations
von: Freund, Guy, et al.
Veröffentlicht: (2026)
von: Freund, Guy, et al.
Veröffentlicht: (2026)
TMRL: Diffusion Timestep-Modulated Pretraining Enables Exploration for Efficient Policy Finetuning
von: Hong, Matthew M., et al.
Veröffentlicht: (2026)
von: Hong, Matthew M., et al.
Veröffentlicht: (2026)
Push Smarter, Not Harder: Hierarchical RL-Diffusion Policy for Efficient Nonprehensile Manipulation
von: Caro, Steven, et al.
Veröffentlicht: (2025)
von: Caro, Steven, et al.
Veröffentlicht: (2025)
VendiRL: A Framework for Self-Supervised Reinforcement Learning of Diversely Diverse Skills
von: Lintunen, Erik M.
Veröffentlicht: (2025)
von: Lintunen, Erik M.
Veröffentlicht: (2025)
Disentangled Unsupervised Skill Discovery for Efficient Hierarchical Reinforcement Learning
von: Hu, Jiaheng, et al.
Veröffentlicht: (2024)
von: Hu, Jiaheng, et al.
Veröffentlicht: (2024)
SkillBlender: Towards Versatile Humanoid Whole-Body Loco-Manipulation via Skill Blending
von: Kuang, Yuxuan, et al.
Veröffentlicht: (2025)
von: Kuang, Yuxuan, et al.
Veröffentlicht: (2025)
RL Token: Bootstrapping Online RL with Vision-Language-Action Models
von: Xu, Charles, et al.
Veröffentlicht: (2026)
von: Xu, Charles, et al.
Veröffentlicht: (2026)
DexMachina: Functional Retargeting for Bimanual Dexterous Manipulation
von: Mandi, Zhao, et al.
Veröffentlicht: (2025)
von: Mandi, Zhao, et al.
Veröffentlicht: (2025)
RIFT: Group-Relative RL Fine-Tuning for Realistic and Controllable Traffic Simulation
von: Chen, Keyu, et al.
Veröffentlicht: (2025)
von: Chen, Keyu, et al.
Veröffentlicht: (2025)
Grasp, See, and Place: Efficient Unknown Object Rearrangement with Policy Structure Prior
von: Xu, Kechun, et al.
Veröffentlicht: (2024)
von: Xu, Kechun, et al.
Veröffentlicht: (2024)
Prompt-Driven Domain Adaptation for End-to-End Autonomous Driving via In-Context RL
von: Khurram, Aleesha, et al.
Veröffentlicht: (2025)
von: Khurram, Aleesha, et al.
Veröffentlicht: (2025)
SECRM-2D: RL-Based Efficient and Comfortable Route-Following Autonomous Driving with Analytic Safety Guarantees
von: Shi, Tianyu, et al.
Veröffentlicht: (2024)
von: Shi, Tianyu, et al.
Veröffentlicht: (2024)
Foundational Policy Acquisition via Multitask Learning for Motor Skill Generation
von: Yamamori, Satoshi, et al.
Veröffentlicht: (2023)
von: Yamamori, Satoshi, et al.
Veröffentlicht: (2023)
Online Distribution Shift Detection via Recency Prediction
von: Luo, Rachel, et al.
Veröffentlicht: (2022)
von: Luo, Rachel, et al.
Veröffentlicht: (2022)
Slug Mobile: Test-Bench for RL Testing
von: Morris, Jonathan Wellington, et al.
Veröffentlicht: (2024)
von: Morris, Jonathan Wellington, et al.
Veröffentlicht: (2024)
Periodic Skill Discovery
von: Park, Jonghae, et al.
Veröffentlicht: (2025)
von: Park, Jonghae, et al.
Veröffentlicht: (2025)
Uni-Skill: Building Self-Evolving Skill Repository for Generalizable Robotic Manipulation
von: Xie, Senwei, et al.
Veröffentlicht: (2026)
von: Xie, Senwei, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Latent Policy Barrier: Learning Robust Visuomotor Policies by Staying In-Distribution
von: Sun, Zhanyi, et al.
Veröffentlicht: (2025) -
Posterior Behavioral Cloning: Pretraining BC Policies for Efficient RL Finetuning
von: Wagenmaker, Andrew, et al.
Veröffentlicht: (2025) -
Residual Off-Policy RL for Finetuning Behavior Cloning Policies
von: Ankile, Lars, et al.
Veröffentlicht: (2025) -
RL-VLM-F: Reinforcement Learning from Vision Language Foundation Model Feedback
von: Wang, Yufei, et al.
Veröffentlicht: (2024) -
EquiBot: SIM(3)-Equivariant Diffusion Policy for Generalizable and Data Efficient Learning
von: Yang, Jingyun, et al.
Veröffentlicht: (2024)