Mean Flow Policy with Instantaneous Velocity Constraint for One-step Action Generation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhan, Guojian, Tao, Letian, Wang, Pengcheng, Wang, Yixiao, Li, Yiheng, Chen, Yuxin, Li, Hongyang, Tomizuka, Masayoshi, Li, Shengbo Eben |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
DADP: Domain Adaptive Diffusion Policy
von: Wang, Pengcheng, et al.
Veröffentlicht: (2026)
von: Wang, Pengcheng, et al.
Veröffentlicht: (2026)
Mind Your Entropy: From Maximum Entropy to Trajectory Entropy-Constrained RL
von: Zhan, Guojian, et al.
Veröffentlicht: (2025)
von: Zhan, Guojian, et al.
Veröffentlicht: (2025)
Bootstrap Off-policy with World Model
von: Zhan, Guojian, et al.
Veröffentlicht: (2025)
von: Zhan, Guojian, et al.
Veröffentlicht: (2025)
Canonical Form of Datatic Description in Control Systems
von: Zhan, Guojian, et al.
Veröffentlicht: (2024)
von: Zhan, Guojian, et al.
Veröffentlicht: (2024)
The Feasibility Theory of Constrained Reinforcement Learning: A Tutorial Study
von: Yang, Yujie, et al.
Veröffentlicht: (2024)
von: Yang, Yujie, et al.
Veröffentlicht: (2024)
Controllability Test for Nonlinear Datatic Systems
von: Yang, Yujie, et al.
Veröffentlicht: (2024)
von: Yang, Yujie, et al.
Veröffentlicht: (2024)
Transferable Latent-to-Latent Locomotion Policy for Efficient and Versatile Motion Control of Diverse Legged Robots
von: Zheng, Ziang, et al.
Veröffentlicht: (2025)
von: Zheng, Ziang, et al.
Veröffentlicht: (2025)
Residual Policy Gradient: A Reward View of KL-regularized Objective
von: Wang, Pengcheng, et al.
Veröffentlicht: (2025)
von: Wang, Pengcheng, et al.
Veröffentlicht: (2025)
Real-Time Generative Policy via Langevin-Guided Flow Matching for Autonomous Driving
von: Zhu, Tianze, et al.
Veröffentlicht: (2026)
von: Zhu, Tianze, et al.
Veröffentlicht: (2026)
Jump-Start Reinforcement Learning with Self-Evolving Priors for Extreme Monopedal Locomotion
von: Zheng, Ziang, et al.
Veröffentlicht: (2025)
von: Zheng, Ziang, et al.
Veröffentlicht: (2025)
Residual-MPPI: Online Policy Customization for Continuous Control
von: Wang, Pengcheng, et al.
Veröffentlicht: (2024)
von: Wang, Pengcheng, et al.
Veröffentlicht: (2024)
MeanCache: From Instantaneous to Average Velocity for Accelerating Flow Matching Inference
von: Gao, Huanlin, et al.
Veröffentlicht: (2026)
von: Gao, Huanlin, et al.
Veröffentlicht: (2026)
An Explicit Discrete-Time Dynamic Vehicle Model with Assured Numerical Stability
von: Zhan, Guojian, et al.
Veröffentlicht: (2024)
von: Zhan, Guojian, et al.
Veröffentlicht: (2024)
Distributional Soft Actor-Critic with Harmonic Gradient for Safe and Efficient Autonomous Driving in Multi-lane Scenarios
von: Zhang, Feihong, et al.
Veröffentlicht: (2025)
von: Zhang, Feihong, et al.
Veröffentlicht: (2025)
CLAW: Composable Language-Annotated Whole-body Motion Generation
von: Cao, Jianuo, et al.
Veröffentlicht: (2026)
von: Cao, Jianuo, et al.
Veröffentlicht: (2026)
Zeroth-Order Actor-Critic: An Evolutionary Framework for Sequential Decision Problems
von: Lei, Yuheng, et al.
Veröffentlicht: (2022)
von: Lei, Yuheng, et al.
Veröffentlicht: (2022)
Off-policy Reinforcement Learning with Model-based Exploration Augmentation
von: Wang, Likun, et al.
Veröffentlicht: (2025)
von: Wang, Likun, et al.
Veröffentlicht: (2025)
LanguageMPC: Large Language Models as Decision Makers for Autonomous Driving
von: Sha, Hao, et al.
Veröffentlicht: (2023)
von: Sha, Hao, et al.
Veröffentlicht: (2023)
FDPP: Fine-tune Diffusion Policy with Human Preference
von: Chen, Yuxin, et al.
Veröffentlicht: (2025)
von: Chen, Yuxin, et al.
Veröffentlicht: (2025)
One-step Latent-free Image Generation with Pixel Mean Flows
von: Lu, Yiyang, et al.
Veröffentlicht: (2026)
von: Lu, Yiyang, et al.
Veröffentlicht: (2026)
STAPO: Stabilizing Reinforcement Learning for LLMs by Silencing Rare Spurious Tokens
von: Liu, Shiqi, et al.
Veröffentlicht: (2026)
von: Liu, Shiqi, et al.
Veröffentlicht: (2026)
IntMeanFlow: Few-step Speech Generation with Integral Velocity Distillation
von: Wang, Wei, et al.
Veröffentlicht: (2025)
von: Wang, Wei, et al.
Veröffentlicht: (2025)
Mean Flows for One-step Generative Modeling
von: Geng, Zhengyang, et al.
Veröffentlicht: (2025)
von: Geng, Zhengyang, et al.
Veröffentlicht: (2025)
Residual Q-Learning: Offline and Online Policy Customization without Value
von: Li, Chenran, et al.
Veröffentlicht: (2023)
von: Li, Chenran, et al.
Veröffentlicht: (2023)
Score-Based One-step MeanFlow Policy Optimization
von: Kim, Kyungyoon, et al.
Veröffentlicht: (2026)
von: Kim, Kyungyoon, et al.
Veröffentlicht: (2026)
Exchange Policy Optimization Algorithm for Semi-Infinite Safe Reinforcement Learning
von: Zhang, Jiaming, et al.
Veröffentlicht: (2025)
von: Zhang, Jiaming, et al.
Veröffentlicht: (2025)
Efficient Reinforcement Learning of Task Planners for Robotic Palletization through Iterative Action Masking Learning
von: Wu, Zheng, et al.
Veröffentlicht: (2024)
von: Wu, Zheng, et al.
Veröffentlicht: (2024)
Physics-Aware Robotic Palletization with Online Masking Inference
von: Zhang, Tianqi, et al.
Veröffentlicht: (2025)
von: Zhang, Tianqi, et al.
Veröffentlicht: (2025)
MP1: MeanFlow Tames Policy Learning in 1-step for Robotic Manipulation
von: Sheng, Juyi, et al.
Veröffentlicht: (2025)
von: Sheng, Juyi, et al.
Veröffentlicht: (2025)
One Step Is Enough: Dispersive MeanFlow Policy Optimization
von: Zou, Guowei, et al.
Veröffentlicht: (2026)
von: Zou, Guowei, et al.
Veröffentlicht: (2026)
Bridging the Sim-to-Real Gap with Dynamic Compliance Tuning for Industrial Insertion
von: Zhang, Xiang, et al.
Veröffentlicht: (2023)
von: Zhang, Xiang, et al.
Veröffentlicht: (2023)
Conformal Symplectic Optimization for Stable Reinforcement Learning
von: Lyu, Yao, et al.
Veröffentlicht: (2024)
von: Lyu, Yao, et al.
Veröffentlicht: (2024)
Immiscible Diffusion: Accelerating Diffusion Training with Noise Assignment
von: Li, Yiheng, et al.
Veröffentlicht: (2024)
von: Li, Yiheng, et al.
Veröffentlicht: (2024)
Improved Immiscible Diffusion: Accelerate Diffusion Training by Reducing Its Miscibility
von: Li, Yiheng, et al.
Veröffentlicht: (2025)
von: Li, Yiheng, et al.
Veröffentlicht: (2025)
Pre-training on Synthetic Driving Data for Trajectory Prediction
von: Li, Yiheng, et al.
Veröffentlicht: (2023)
von: Li, Yiheng, et al.
Veröffentlicht: (2023)
TwinFlow: Realizing One-step Generation on Large Models with Self-adversarial Flows
von: Cheng, Zhenglin, et al.
Veröffentlicht: (2025)
von: Cheng, Zhenglin, et al.
Veröffentlicht: (2025)
Sparse Diffusion Policy: A Sparse, Reusable, and Flexible Policy for Robot Learning
von: Wang, Yixiao, et al.
Veröffentlicht: (2024)
von: Wang, Yixiao, et al.
Veröffentlicht: (2024)
One-Step Generative Policies with Q-Learning: A Reformulation of MeanFlow
von: Wang, Zeyuan, et al.
Veröffentlicht: (2025)
von: Wang, Zeyuan, et al.
Veröffentlicht: (2025)
Algorithm Design and Comparative Test of Natural Gradient Gaussian Approximation Filter
von: Cao, Wenhan, et al.
Veröffentlicht: (2025)
von: Cao, Wenhan, et al.
Veröffentlicht: (2025)
On the Equilibrium between Feasible Zone and Uncertain Model in Safe Exploration
von: Yang, Yujie, et al.
Veröffentlicht: (2026)
von: Yang, Yujie, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
DADP: Domain Adaptive Diffusion Policy
von: Wang, Pengcheng, et al.
Veröffentlicht: (2026) -
Mind Your Entropy: From Maximum Entropy to Trajectory Entropy-Constrained RL
von: Zhan, Guojian, et al.
Veröffentlicht: (2025) -
Bootstrap Off-policy with World Model
von: Zhan, Guojian, et al.
Veröffentlicht: (2025) -
Canonical Form of Datatic Description in Control Systems
von: Zhan, Guojian, et al.
Veröffentlicht: (2024) -
The Feasibility Theory of Constrained Reinforcement Learning: A Tutorial Study
von: Yang, Yujie, et al.
Veröffentlicht: (2024)