NFPO: Stabilized Policy Optimization of Normalizing Flow for Robotic Policy Learning
Fuente:
arXiv
Saved in:
| Main Authors: | Shi, Diyuan, Tang, Yiqi, Zhuang, Zifeng, Wang, Donglin |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Discovering Self-Protective Falling Policy for Humanoid Robot via Deep Reinforcement Learning
by: Shi, Diyuan, et al.
Published: (2025)
by: Shi, Diyuan, et al.
Published: (2025)
TDMPBC: Self-Imitative Reinforcement Learning for Humanoid Robot Control
by: Zhuang, Zifeng, et al.
Published: (2025)
by: Zhuang, Zifeng, et al.
Published: (2025)
PolicyFlow: Policy Optimization with Continuous Normalizing Flow in Reinforcement Learning
by: Yang, Shunpeng, et al.
Published: (2026)
by: Yang, Shunpeng, et al.
Published: (2026)
ReinboT: Amplifying Robot Visual-Language Manipulation with Reinforcement Learning
by: Zhang, Hongyin, et al.
Published: (2025)
by: Zhang, Hongyin, et al.
Published: (2025)
Unlock Reliable Skill Inference for Quadruped Adaptive Behavior by Skill Graph
by: Zhang, Hongyin, et al.
Published: (2023)
by: Zhang, Hongyin, et al.
Published: (2023)
Efficient Online RL Fine Tuning with Offline Pre-trained Policy Only
by: Xiao, Wei, et al.
Published: (2025)
by: Xiao, Wei, et al.
Published: (2025)
Learning Robotic Policy with Imagined Transition: Mitigating the Trade-off between Robustness and Optimality
by: Xiao, Wei, et al.
Published: (2025)
by: Xiao, Wei, et al.
Published: (2025)
World-Value-Action Model: Implicit Planning for Vision-Language-Action Systems
by: Li, Runze, et al.
Published: (2026)
by: Li, Runze, et al.
Published: (2026)
Adaptive Diffusion Policy Optimization for Robotic Manipulation
by: Jiang, Huiyun, et al.
Published: (2025)
by: Jiang, Huiyun, et al.
Published: (2025)
Flow Policy Gradients for Robot Control
by: Yi, Brent, et al.
Published: (2026)
by: Yi, Brent, et al.
Published: (2026)
Diffusion Stabilizer Policy for Automated Surgical Robot Manipulations
by: Ho, Chonlam, et al.
Published: (2025)
by: Ho, Chonlam, et al.
Published: (2025)
Normalizing Flows are Capable Models for Bi-manual Visuomotor Policy
by: Li, Jialong, et al.
Published: (2025)
by: Li, Jialong, et al.
Published: (2025)
ManiCM: Real-time 3D Diffusion Policy via Consistency Model for Robotic Manipulation
by: Lu, Guanxing, et al.
Published: (2024)
by: Lu, Guanxing, et al.
Published: (2024)
Riemannian Flow Matching Policy for Robot Motion Learning
by: Braun, Max, et al.
Published: (2024)
by: Braun, Max, et al.
Published: (2024)
Drift-Based Policy Optimization: Native One-Step Policy Learning for Online Robot Control
by: Gao, Yuxuan, et al.
Published: (2026)
by: Gao, Yuxuan, et al.
Published: (2026)
Dynamic Adaptive Legged Locomotion Policy via Decoupling Reaction Force Control and Gait Control
by: Wang, Renjie, et al.
Published: (2025)
by: Wang, Renjie, et al.
Published: (2025)
FlowCorrect: Efficient Interactive Correction of Generative Flow Policies for Robotic Manipulation
by: Welte, Edgar, et al.
Published: (2026)
by: Welte, Edgar, et al.
Published: (2026)
Robot Trains Robot: Automatic Real-World Policy Adaptation and Learning for Humanoids
by: Hu, Kaizhe, et al.
Published: (2025)
by: Hu, Kaizhe, et al.
Published: (2025)
MP1: MeanFlow Tames Policy Learning in 1-step for Robotic Manipulation
by: Sheng, Juyi, et al.
Published: (2025)
by: Sheng, Juyi, et al.
Published: (2025)
Raising Body Ownership in End-to-End Visuomotor Policy Learning via Robot-Centric Pooling
by: Zhuang, Zheyu, et al.
Published: (2024)
by: Zhuang, Zheyu, et al.
Published: (2024)
FLOWER: Democratizing Generalist Robot Policies with Efficient Vision-Language-Action Flow Policies
by: Reuss, Moritz, et al.
Published: (2025)
by: Reuss, Moritz, et al.
Published: (2025)
FlowPolicy: Enabling Fast and Robust 3D Flow-based Policy via Consistency Flow Matching for Robot Manipulation
by: Zhang, Qinglun, et al.
Published: (2024)
by: Zhang, Qinglun, et al.
Published: (2024)
Learning Robotic Manipulation Policies from Point Clouds with Conditional Flow Matching
by: Chisari, Eugenio, et al.
Published: (2024)
by: Chisari, Eugenio, et al.
Published: (2024)
One Step Is Enough: Dispersive MeanFlow Policy Optimization
by: Zou, Guowei, et al.
Published: (2026)
by: Zou, Guowei, et al.
Published: (2026)
ManiFlow: A General Robot Manipulation Policy via Consistency Flow Training
by: Yan, Ge, et al.
Published: (2025)
by: Yan, Ge, et al.
Published: (2025)
GPO: Growing Policy Optimization for Legged Robot Locomotion and Whole-Body Control
by: Liao, Shuhao, et al.
Published: (2026)
by: Liao, Shuhao, et al.
Published: (2026)
FlowRAM: Grounding Flow Matching Policy with Region-Aware Mamba Framework for Robotic Manipulation
by: Wang, Sen, et al.
Published: (2025)
by: Wang, Sen, et al.
Published: (2025)
Spatial RoboGrasp: Generalized Robotic Grasping Control Policy
by: Huang, Yiqi, et al.
Published: (2025)
by: Huang, Yiqi, et al.
Published: (2025)
Posterior Optimization with Clipped Objective for Bridging Efficiency and Stability in Generative Policy Learning
by: Chen, Yuhui, et al.
Published: (2026)
by: Chen, Yuhui, et al.
Published: (2026)
Where-to-Learn: Analytical Policy Gradient Directed Exploration for On-Policy Robotic Reinforcement Learning
by: Chang, Leixin, et al.
Published: (2026)
by: Chang, Leixin, et al.
Published: (2026)
Learning Human-Aware Robot Policies for Adaptive Assistance
by: Qin, Jason, et al.
Published: (2024)
by: Qin, Jason, et al.
Published: (2024)
RoboGrasp: A Universal Grasping Policy for Robust Robotic Control
by: Huang, Yiqi, et al.
Published: (2025)
by: Huang, Yiqi, et al.
Published: (2025)
Latent Policy Steering with Embodiment-Agnostic Pretrained World Models
by: Wang, Yiqi, et al.
Published: (2025)
by: Wang, Yiqi, et al.
Published: (2025)
Masked Generative Policy for Robotic Control
by: Zhuang, Lipeng, et al.
Published: (2025)
by: Zhuang, Lipeng, et al.
Published: (2025)
SafeFlowMPC: Predictive and Safe Trajectory Planning for Robot Manipulators with Learning-based Policies
by: Oelerich, Thies, et al.
Published: (2026)
by: Oelerich, Thies, et al.
Published: (2026)
Multitask Reinforcement Learning for Quadcopter Attitude Stabilization and Tracking using Graph Policy
by: Liu, Yu Tang, et al.
Published: (2025)
by: Liu, Yu Tang, et al.
Published: (2025)
RoboRouter: Training-Free Policy Routing for Robotic Manipulation
by: Chen, Yiteng, et al.
Published: (2026)
by: Chen, Yiteng, et al.
Published: (2026)
Boundary-to-Region Supervision for Offline Safe Reinforcement Learning
by: Su, Huikang, et al.
Published: (2025)
by: Su, Huikang, et al.
Published: (2025)
Design of Restricted Normalizing Flow towards Arbitrary Stochastic Policy with Computational Efficiency
by: Kobayashi, Taisuke, et al.
Published: (2024)
by: Kobayashi, Taisuke, et al.
Published: (2024)
Robot Fleet Learning via Policy Merging
by: Wang, Lirui, et al.
Published: (2023)
by: Wang, Lirui, et al.
Published: (2023)
Similar Items
-
Discovering Self-Protective Falling Policy for Humanoid Robot via Deep Reinforcement Learning
by: Shi, Diyuan, et al.
Published: (2025) -
TDMPBC: Self-Imitative Reinforcement Learning for Humanoid Robot Control
by: Zhuang, Zifeng, et al.
Published: (2025) -
PolicyFlow: Policy Optimization with Continuous Normalizing Flow in Reinforcement Learning
by: Yang, Shunpeng, et al.
Published: (2026) -
ReinboT: Amplifying Robot Visual-Language Manipulation with Reinforcement Learning
by: Zhang, Hongyin, et al.
Published: (2025) -
Unlock Reliable Skill Inference for Quadruped Adaptive Behavior by Skill Graph
by: Zhang, Hongyin, et al.
Published: (2023)