Flow Matching Policy Gradients
Fuente:
arXiv
Guardado en:
| Autores principales: | McAllister, David, Ge, Songwei, Yi, Brent, Kim, Chung Min, Weber, Ethan, Choi, Hongsuk, Feng, Haiwen, Kanazawa, Angjoo |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
PyRoki: A Modular Toolkit for Robot Kinematic Optimization
por: Kim, Chung Min, et al.
Publicado: (2025)
por: Kim, Chung Min, et al.
Publicado: (2025)
Flow Policy Gradients for Robot Control
por: Yi, Brent, et al.
Publicado: (2026)
por: Yi, Brent, et al.
Publicado: (2026)
Visual Imitation Enables Contextual Humanoid Control
por: Allshire, Arthur, et al.
Publicado: (2025)
por: Allshire, Arthur, et al.
Publicado: (2025)
Viser: Imperative, Web-based 3D Visualization in Python
por: Yi, Brent, et al.
Publicado: (2025)
por: Yi, Brent, et al.
Publicado: (2025)
Decentralized Diffusion Models
por: McAllister, David, et al.
Publicado: (2025)
por: McAllister, David, et al.
Publicado: (2025)
Eye, Robot: Learning to Look to Act with a BC-RL Perception-Action Loop
por: Kerr, Justin, et al.
Publicado: (2025)
por: Kerr, Justin, et al.
Publicado: (2025)
Rethinking Score Distillation as a Bridge Between Image Distributions
por: McAllister, David, et al.
Publicado: (2024)
por: McAllister, David, et al.
Publicado: (2024)
Latent-Conditioned Policy Gradient for Multi-Objective Deep Reinforcement Learning
por: Kanazawa, Takuya, et al.
Publicado: (2023)
por: Kanazawa, Takuya, et al.
Publicado: (2023)
Robot See Robot Do: Imitating Articulated Object Manipulation with Monocular 4D Reconstruction
por: Kerr, Justin, et al.
Publicado: (2024)
por: Kerr, Justin, et al.
Publicado: (2024)
Finite Difference Flow Optimization for RL Post-Training of Text-to-Image Models
por: McAllister, David, et al.
Publicado: (2026)
por: McAllister, David, et al.
Publicado: (2026)
Drifting Field Policy: A One-Step Generative Policy via Wasserstein Gradient Flow
por: Koo, Juil, et al.
Publicado: (2026)
por: Koo, Juil, et al.
Publicado: (2026)
Reconstructing People, Places, and Cameras
por: Müller, Lea, et al.
Publicado: (2024)
por: Müller, Lea, et al.
Publicado: (2024)
Reinforcement Fine-Tuning of Flow-Matching Policies for Vision-Language-Action Models
por: Lyu, Mingyang, et al.
Publicado: (2025)
por: Lyu, Mingyang, et al.
Publicado: (2025)
WarmPrior: Straightening Flow-Matching Policies with Temporal Priors
por: Kang, Sinjae, et al.
Publicado: (2026)
por: Kang, Sinjae, et al.
Publicado: (2026)
Riemannian Flow Matching Policy for Robot Motion Learning
por: Braun, Max, et al.
Publicado: (2024)
por: Braun, Max, et al.
Publicado: (2024)
Fast and Robust Visuomotor Riemannian Flow Matching Policy
por: Ding, Haoran, et al.
Publicado: (2024)
por: Ding, Haoran, et al.
Publicado: (2024)
ReinFlow: Fine-tuning Flow Matching Policy with Online Reinforcement Learning
por: Zhang, Tonghe, et al.
Publicado: (2025)
por: Zhang, Tonghe, et al.
Publicado: (2025)
A Case for Conversational Cataloging.
por: McAllister, Caryl, et al.
Publicado: (1983)
por: McAllister, Caryl, et al.
Publicado: (1983)
CoLA-Flow Policy: Temporally Coherent Imitation Learning via Continuous Latent Action Flow Matching for Robotic Manipulation
por: Songwei, Wu, et al.
Publicado: (2026)
por: Songwei, Wu, et al.
Publicado: (2026)
Perceptive Humanoid Parkour: Chaining Dynamic Human Skills via Motion Matching
por: Wu, Zhen, et al.
Publicado: (2026)
por: Wu, Zhen, et al.
Publicado: (2026)
FocalPolicy: Frequency-Optimized Chunking and Locally Anchored Flow Matching for Coherent Visuomotor Policy
por: He, Qian, et al.
Publicado: (2026)
por: He, Qian, et al.
Publicado: (2026)
VFP: Variational Flow-Matching Policy for Multi-Modal Robot Manipulation
por: Zhai, Xuanran, et al.
Publicado: (2025)
por: Zhai, Xuanran, et al.
Publicado: (2025)
NeuralRemaster: Phase-Preserving Diffusion for Structure-Aligned Generation
por: Zeng, Yu, et al.
Publicado: (2025)
por: Zeng, Yu, et al.
Publicado: (2025)
Real-Time Generative Policy via Langevin-Guided Flow Matching for Autonomous Driving
por: Zhu, Tianze, et al.
Publicado: (2026)
por: Zhu, Tianze, et al.
Publicado: (2026)
Grasp Synthesis Matching From Rigid To Soft Robot Grippers Using Conditional Flow Matching
por: Parulekar, Tanisha, et al.
Publicado: (2026)
por: Parulekar, Tanisha, et al.
Publicado: (2026)
Mitigating Suboptimality of Deterministic Policy Gradients in Complex Q-functions
por: Jain, Ayush, et al.
Publicado: (2024)
por: Jain, Ayush, et al.
Publicado: (2024)
On-Line Library Housekeeping Systems. A Survey
por: McAllister, Caryl
Publicado: (1971)
por: McAllister, Caryl
Publicado: (1971)
Flow with the Force Field: Learning 3D Compliant Flow Matching Policies from Force and Demonstration-Guided Simulation Data
por: Li, Tianyu, et al.
Publicado: (2025)
por: Li, Tianyu, et al.
Publicado: (2025)
Score and Distribution Matching Policy: Advanced Accelerated Visuomotor Policies via Matched Distillation
por: Jia, Bofang, et al.
Publicado: (2024)
por: Jia, Bofang, et al.
Publicado: (2024)
Agent-to-Sim: Learning Interactive Behavior Models from Casual Longitudinal Videos
por: Yang, Gengshan, et al.
Publicado: (2024)
por: Yang, Gengshan, et al.
Publicado: (2024)
The Lie We Tell: Correcting the Euclidean Fallacy in Vision Language Action Policies via Score Matching on Tangent Space
por: Chuang, Bing-Cheng, et al.
Publicado: (2026)
por: Chuang, Bing-Cheng, et al.
Publicado: (2026)
FLAG: Flow Policy MaxEnt-RL by Latent Augmented Guidance
por: Kim, Sungha, et al.
Publicado: (2026)
por: Kim, Sungha, et al.
Publicado: (2026)
PolicyFlow: Policy Optimization with Continuous Normalizing Flow in Reinforcement Learning
por: Yang, Shunpeng, et al.
Publicado: (2026)
por: Yang, Shunpeng, et al.
Publicado: (2026)
Coordinated Humanoid Manipulation with Choice Policies
por: Qi, Haozhi, et al.
Publicado: (2025)
por: Qi, Haozhi, et al.
Publicado: (2025)
Generalized Advantage Estimation for Distributional Policy Gradients
por: Shaik, Shahil, et al.
Publicado: (2025)
por: Shaik, Shahil, et al.
Publicado: (2025)
Harnessing Bounded-Support Evolution Strategies for Policy Refinement
por: Hirschowitz, Ethan, et al.
Publicado: (2025)
por: Hirschowitz, Ethan, et al.
Publicado: (2025)
OmniRetarget: Interaction-Preserving Data Generation for Humanoid Whole-Body Loco-Manipulation and Scene Interaction
por: Yang, Lujie, et al.
Publicado: (2025)
por: Yang, Lujie, et al.
Publicado: (2025)
TWIST2: Scalable, Portable, and Holistic Humanoid Data Collection System
por: Ze, Yanjie, et al.
Publicado: (2025)
por: Ze, Yanjie, et al.
Publicado: (2025)
Dynamic Objects Relocalization in Changing Environments with Flow Matching
por: Argenziano, Francesco, et al.
Publicado: (2025)
por: Argenziano, Francesco, et al.
Publicado: (2025)
Quantile-Coupled Flow Matching for Distributional Reinforcement Learning
por: Groom, Michael, et al.
Publicado: (2026)
por: Groom, Michael, et al.
Publicado: (2026)
Ejemplares similares
-
PyRoki: A Modular Toolkit for Robot Kinematic Optimization
por: Kim, Chung Min, et al.
Publicado: (2025) -
Flow Policy Gradients for Robot Control
por: Yi, Brent, et al.
Publicado: (2026) -
Visual Imitation Enables Contextual Humanoid Control
por: Allshire, Arthur, et al.
Publicado: (2025) -
Viser: Imperative, Web-based 3D Visualization in Python
por: Yi, Brent, et al.
Publicado: (2025) -
Decentralized Diffusion Models
por: McAllister, David, et al.
Publicado: (2025)