Drift-Based Policy Optimization: Native One-Step Policy Learning for Online Robot Control
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Gao, Yuxuan, Shen, Yedong, Zhang, Shiqi, Yu, Wenhao, Duan, Yifan, pan, Jia, Wu, Jiajia, Deng, Jiajun, Zhang, Yanyong |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
STDArm: Transferring Visuomotor Policies From Static Data Training to Dynamic Robot Manipulation
von: Duan, Yifan, et al.
Veröffentlicht: (2025)
von: Duan, Yifan, et al.
Veröffentlicht: (2025)
MT-PCR: Leveraging Modality Transformation for Large-Scale Point Cloud Registration with Limited Overlap
von: Wu, Yilong, et al.
Veröffentlicht: (2025)
von: Wu, Yilong, et al.
Veröffentlicht: (2025)
GA-GS: Generation-Assisted Gaussian Splatting for Static Scene Reconstruction
von: Shen, Yedong, et al.
Veröffentlicht: (2026)
von: Shen, Yedong, et al.
Veröffentlicht: (2026)
Implicit Drifting Policy: One-Step Action Generation via Conditional Expert Geometry
von: Yang, Zemin, et al.
Veröffentlicht: (2026)
von: Yang, Zemin, et al.
Veröffentlicht: (2026)
Drifting Field Policy: A One-Step Generative Policy via Wasserstein Gradient Flow
von: Koo, Juil, et al.
Veröffentlicht: (2026)
von: Koo, Juil, et al.
Veröffentlicht: (2026)
LDP: A Local Diffusion Planner for Efficient Robot Navigation and Collision Avoidance
von: Yu, Wenhao, et al.
Veröffentlicht: (2024)
von: Yu, Wenhao, et al.
Veröffentlicht: (2024)
One Step Is Enough: Dispersive MeanFlow Policy Optimization
von: Zou, Guowei, et al.
Veröffentlicht: (2026)
von: Zou, Guowei, et al.
Veröffentlicht: (2026)
PGOV3D: Open-Vocabulary 3D Semantic Segmentation with Partial-to-Global Curriculum
von: Zhang, Shiqi, et al.
Veröffentlicht: (2025)
von: Zhang, Shiqi, et al.
Veröffentlicht: (2025)
Positive-Only Drifting Policy Optimization
von: Zhang, Qi
Veröffentlicht: (2026)
von: Zhang, Qi
Veröffentlicht: (2026)
Latent Policy Steering through One-Step Flow Policies
von: Im, Hokyun, et al.
Veröffentlicht: (2026)
von: Im, Hokyun, et al.
Veröffentlicht: (2026)
Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization
von: Preiss, James A., et al.
Veröffentlicht: (2025)
von: Preiss, James A., et al.
Veröffentlicht: (2025)
SeedPolicy: Horizon Scaling via Self-Evolving Diffusion Policy for Robot Manipulation
von: Gui, Youqiang, et al.
Veröffentlicht: (2026)
von: Gui, Youqiang, et al.
Veröffentlicht: (2026)
Steady-State Drifting Equilibrium Analysis of Single-Track Two-Wheeled Robots for Controller Design
von: Jing, Feilong, et al.
Veröffentlicht: (2025)
von: Jing, Feilong, et al.
Veröffentlicht: (2025)
One-Step Flow Policy: Self-Distillation for Fast Visuomotor Policies
von: Li, Shaolong, et al.
Veröffentlicht: (2026)
von: Li, Shaolong, et al.
Veröffentlicht: (2026)
Policy Contrastive Decoding for Robotic Foundation Models
von: Wu, Shihan, et al.
Veröffentlicht: (2025)
von: Wu, Shihan, et al.
Veröffentlicht: (2025)
Uni-O4: Unifying Online and Offline Deep Reinforcement Learning with Multi-Step On-Policy Optimization
von: Lei, Kun, et al.
Veröffentlicht: (2023)
von: Lei, Kun, et al.
Veröffentlicht: (2023)
One-Step Diffusion Policy: Fast Visuomotor Policies via Diffusion Distillation
von: Wang, Zhendong, et al.
Veröffentlicht: (2024)
von: Wang, Zhendong, et al.
Veröffentlicht: (2024)
Model-Based Policy Adaptation for Closed-Loop End-to-End Autonomous Driving
von: Lin, Haohong, et al.
Veröffentlicht: (2025)
von: Lin, Haohong, et al.
Veröffentlicht: (2025)
Query-Centric Diffusion Policy for Generalizable Robotic Assembly
von: Xu, Ziyi, et al.
Veröffentlicht: (2025)
von: Xu, Ziyi, et al.
Veröffentlicht: (2025)
Flow Policy Gradients for Robot Control
von: Yi, Brent, et al.
Veröffentlicht: (2026)
von: Yi, Brent, et al.
Veröffentlicht: (2026)
An Real-Sim-Real (RSR) Loop Framework for Generalizable Robotic Policy Transfer with Differentiable Simulation
von: Shi, Lu, et al.
Veröffentlicht: (2025)
von: Shi, Lu, et al.
Veröffentlicht: (2025)
AdaWorldPolicy: World-Model-Driven Diffusion Policy with Online Adaptive Learning for Robotic Manipulation
von: Yuan, Ge, et al.
Veröffentlicht: (2026)
von: Yuan, Ge, et al.
Veröffentlicht: (2026)
MHRC: Closed-loop Decentralized Multi-Heterogeneous Robot Collaboration with Large Language Models
von: Yu, Wenhao, et al.
Veröffentlicht: (2024)
von: Yu, Wenhao, et al.
Veröffentlicht: (2024)
Towards Online Safety Corrections for Robotic Manipulation Policies
von: Spalter, Ariana, et al.
Veröffentlicht: (2024)
von: Spalter, Ariana, et al.
Veröffentlicht: (2024)
NFPO: Stabilized Policy Optimization of Normalizing Flow for Robotic Policy Learning
von: Shi, Diyuan, et al.
Veröffentlicht: (2026)
von: Shi, Diyuan, et al.
Veröffentlicht: (2026)
ImaginationPolicy: Towards Generalizable, Precise and Reliable End-to-End Policy for Robotic Manipulation
von: Lu, Dekun, et al.
Veröffentlicht: (2025)
von: Lu, Dekun, et al.
Veröffentlicht: (2025)
Learning Surgical Robotic Manipulation with 3D Spatial Priors
von: Sheng, Yu, et al.
Veröffentlicht: (2026)
von: Sheng, Yu, et al.
Veröffentlicht: (2026)
Cortical Policy: A Dual-Stream View Transformer for Robotic Manipulation
von: Zhang, Xuening, et al.
Veröffentlicht: (2026)
von: Zhang, Xuening, et al.
Veröffentlicht: (2026)
OCC-VO: Dense Mapping via 3D Occupancy-Based Visual Odometry for Autonomous Driving
von: Li, Heng, et al.
Veröffentlicht: (2023)
von: Li, Heng, et al.
Veröffentlicht: (2023)
StereoPolicy: Improving Robotic Manipulation Policies via Stereo Perception
von: Han, Evans, et al.
Veröffentlicht: (2026)
von: Han, Evans, et al.
Veröffentlicht: (2026)
Learning Robust Control Policies for Inverted Pose on Miniature Blimp Robots
von: Yang, Yuanlin, et al.
Veröffentlicht: (2026)
von: Yang, Yuanlin, et al.
Veröffentlicht: (2026)
Learning Native Continuation for Action Chunking Flow Policies
von: Liu, Yufeng, et al.
Veröffentlicht: (2026)
von: Liu, Yufeng, et al.
Veröffentlicht: (2026)
Effective Tuning Strategies for Generalist Robot Manipulation Policies
von: Zhang, Wenbo, et al.
Veröffentlicht: (2024)
von: Zhang, Wenbo, et al.
Veröffentlicht: (2024)
Integrating Online Learning and Connectivity Maintenance for Communication-Aware Multi-Robot Coordination
von: Yang, Yupeng, et al.
Veröffentlicht: (2024)
von: Yang, Yupeng, et al.
Veröffentlicht: (2024)
Learning Quadruped Locomotion Policies using Logical Rules
von: DeFazio, David, et al.
Veröffentlicht: (2021)
von: DeFazio, David, et al.
Veröffentlicht: (2021)
FAVLA: A Force-Adaptive Fast-Slow VLA model for Contact-Rich Robotic Manipulation
von: Li, Yao, et al.
Veröffentlicht: (2026)
von: Li, Yao, et al.
Veröffentlicht: (2026)
TRANSIC: Sim-to-Real Policy Transfer by Learning from Online Correction
von: Jiang, Yunfan, et al.
Veröffentlicht: (2024)
von: Jiang, Yunfan, et al.
Veröffentlicht: (2024)
GPO: Growing Policy Optimization for Legged Robot Locomotion and Whole-Body Control
von: Liao, Shuhao, et al.
Veröffentlicht: (2026)
von: Liao, Shuhao, et al.
Veröffentlicht: (2026)
When would Vision-Proprioception Policies Fail in Robotic Manipulation?
von: Lu, Jingxian, et al.
Veröffentlicht: (2026)
von: Lu, Jingxian, et al.
Veröffentlicht: (2026)
Fast Visuomotor Policy for Robotic Manipulation
von: Jia, Jingkai, et al.
Veröffentlicht: (2025)
von: Jia, Jingkai, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
STDArm: Transferring Visuomotor Policies From Static Data Training to Dynamic Robot Manipulation
von: Duan, Yifan, et al.
Veröffentlicht: (2025) -
MT-PCR: Leveraging Modality Transformation for Large-Scale Point Cloud Registration with Limited Overlap
von: Wu, Yilong, et al.
Veröffentlicht: (2025) -
GA-GS: Generation-Assisted Gaussian Splatting for Static Scene Reconstruction
von: Shen, Yedong, et al.
Veröffentlicht: (2026) -
Implicit Drifting Policy: One-Step Action Generation via Conditional Expert Geometry
von: Yang, Zemin, et al.
Veröffentlicht: (2026) -
Drifting Field Policy: A One-Step Generative Policy via Wasserstein Gradient Flow
von: Koo, Juil, et al.
Veröffentlicht: (2026)