FLOWER: Democratizing Generalist Robot Policies with Efficient Vision-Language-Action Flow Policies
Fuente:
arXiv
Saved in:
| Main Authors: | Reuss, Moritz, Zhou, Hongyi, Rühle, Marcel, Yağmurlu, Ömer Erdinç, Otto, Fabian, Lioutikov, Rudolf |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Scaling Robot Policy Learning via Zero-Shot Labeling with Foundation Models
by: Blank, Nils, et al.
Published: (2024)
by: Blank, Nils, et al.
Published: (2024)
Multimodal Diffusion Transformer: Learning Versatile Behavior from Multimodal Goals
by: Reuss, Moritz, et al.
Published: (2024)
by: Reuss, Moritz, et al.
Published: (2024)
BEAST: Efficient Tokenization of B-Splines Encoded Action Sequences for Imitation Learning
by: Zhou, Hongyi, et al.
Published: (2025)
by: Zhou, Hongyi, et al.
Published: (2025)
Efficient Diffusion Transformer Policies with Mixture of Expert Denoisers for Multitask Learning
by: Reuss, Moritz, et al.
Published: (2024)
by: Reuss, Moritz, et al.
Published: (2024)
Open the Black Box: Step-based Policy Updates for Temporally-Correlated Episodic Reinforcement Learning
by: Li, Ge, et al.
Published: (2024)
by: Li, Ge, et al.
Published: (2024)
NaviTrace: Evaluating Embodied Navigation of Vision-Language Models
by: Windecker, Tim, et al.
Published: (2025)
by: Windecker, Tim, et al.
Published: (2025)
TOP-ERL: Transformer-based Off-Policy Episodic Reinforcement Learning
by: Li, Ge, et al.
Published: (2024)
by: Li, Ge, et al.
Published: (2024)
Variational Distillation of Diffusion Policies into Mixture of Experts
by: Zhou, Hongyi, et al.
Published: (2024)
by: Zhou, Hongyi, et al.
Published: (2024)
Towards Diverse Behaviors: A Benchmark for Imitation Learning with Human Demonstrations
by: Jia, Xiaogang, et al.
Published: (2024)
by: Jia, Xiaogang, et al.
Published: (2024)
PointMapPolicy: Structured Point Cloud Processing for Multi-Modal Imitation Learning
by: Jia, Xiaogang, et al.
Published: (2025)
by: Jia, Xiaogang, et al.
Published: (2025)
Dita: Scaling Diffusion Transformer for Generalist Vision-Language-Action Policy
by: Hou, Zhi, et al.
Published: (2025)
by: Hou, Zhi, et al.
Published: (2025)
BMP: Bridging the Gap between B-Spline and Movement Primitives
by: Liao, Weiran, et al.
Published: (2024)
by: Liao, Weiran, et al.
Published: (2024)
NanoVLA: Routing Decoupled Vision-Language Understanding for Nano-sized Generalist Robotic Policies
by: Chen, Jiahong, et al.
Published: (2025)
by: Chen, Jiahong, et al.
Published: (2025)
SINGER: An Onboard Generalist Vision-Language Navigation Policy for Drones
by: Adang, Maximilian, et al.
Published: (2025)
by: Adang, Maximilian, et al.
Published: (2025)
InternVLA-M1: A Spatially Guided Vision-Language-Action Framework for Generalist Robot Policy
by: Chen, Xinyi, et al.
Published: (2025)
by: Chen, Xinyi, et al.
Published: (2025)
Green-VLA: Staged Vision-Language-Action Model for Generalist Robots
by: Apanasevich, I., et al.
Published: (2026)
by: Apanasevich, I., et al.
Published: (2026)
HiMoE-VLA: Hierarchical Mixture-of-Experts for Generalist Vision-Language-Action Policies
by: Du, Zhiying, et al.
Published: (2025)
by: Du, Zhiying, et al.
Published: (2025)
X-IL: Exploring the Design Space of Imitation Learning Policies
by: Jia, Xiaogang, et al.
Published: (2025)
by: Jia, Xiaogang, et al.
Published: (2025)
Octo: An Open-Source Generalist Robot Policy
by: Octo Model Team, et al.
Published: (2024)
by: Octo Model Team, et al.
Published: (2024)
Robust Finetuning of Vision-Language-Action Robot Policies via Parameter Merging
by: Yadav, Yajat, et al.
Published: (2025)
by: Yadav, Yajat, et al.
Published: (2025)
Gaze2Act: Gaze-Conditioned Vision-Language-Action Policies for Interactive Robot Manipulation
by: Zuo, Kuangji, et al.
Published: (2026)
by: Zuo, Kuangji, et al.
Published: (2026)
Toward Embodiment Equivariant Vision-Language-Action Policy
by: Chen, Anzhe, et al.
Published: (2025)
by: Chen, Anzhe, et al.
Published: (2025)
What Matters in Building Vision-Language-Action Models for Generalist Robots
by: Li, Xinghang, et al.
Published: (2024)
by: Li, Xinghang, et al.
Published: (2024)
Effective Tuning Strategies for Generalist Robot Manipulation Policies
by: Zhang, Wenbo, et al.
Published: (2024)
by: Zhang, Wenbo, et al.
Published: (2024)
PolaRiS: Scalable Real-to-Sim Evaluations for Generalist Robot Policies
by: Jain, Arhan, et al.
Published: (2025)
by: Jain, Arhan, et al.
Published: (2025)
VITA: Vision-to-Action Flow Matching Policy
by: Gao, Dechen, et al.
Published: (2025)
by: Gao, Dechen, et al.
Published: (2025)
Graph-Fused Vision-Language-Action for Policy Reasoning in Multi-Arm Robotic Manipulation
by: Li, Shunlei, et al.
Published: (2025)
by: Li, Shunlei, et al.
Published: (2025)
Reinforcement Fine-Tuning of Flow-Matching Policies for Vision-Language-Action Models
by: Lyu, Mingyang, et al.
Published: (2025)
by: Lyu, Mingyang, et al.
Published: (2025)
RoboArena: Distributed Real-World Evaluation of Generalist Robot Policies
by: Atreya, Pranav, et al.
Published: (2025)
by: Atreya, Pranav, et al.
Published: (2025)
AutoEval: Autonomous Evaluation of Generalist Robot Manipulation Policies in the Real World
by: Zhou, Zhiyuan, et al.
Published: (2025)
by: Zhou, Zhiyuan, et al.
Published: (2025)
Video Prediction Policy: A Generalist Robot Policy with Predictive Visual Representations
by: Hu, Yucheng, et al.
Published: (2024)
by: Hu, Yucheng, et al.
Published: (2024)
A Taxonomy for Evaluating Generalist Robot Manipulation Policies
by: Gao, Jensen, et al.
Published: (2025)
by: Gao, Jensen, et al.
Published: (2025)
Beyond Sight: Finetuning Generalist Robot Policies with Heterogeneous Sensors via Language Grounding
by: Jones, Joshua, et al.
Published: (2025)
by: Jones, Joshua, et al.
Published: (2025)
Learning while Deploying: Fleet-Scale Reinforcement Learning for Generalist Robot Policies
by: Wang, Yi, et al.
Published: (2026)
by: Wang, Yi, et al.
Published: (2026)
Asynchronous Fast-Slow Vision-Language-Action Policies for Whole-Body Robotic Manipulation
by: Zou, Teqiang, et al.
Published: (2025)
by: Zou, Teqiang, et al.
Published: (2025)
Turning Video Models into Generalist Robot Policies
by: Li, Sizhe Lester, et al.
Published: (2026)
by: Li, Sizhe Lester, et al.
Published: (2026)
OmniGuide: Universal Guidance Fields for Enhancing Generalist Robot Policies
by: Song, Yunzhou, et al.
Published: (2026)
by: Song, Yunzhou, et al.
Published: (2026)
RLDG: Robotic Generalist Policy Distillation via Reinforcement Learning
by: Xu, Charles, et al.
Published: (2024)
by: Xu, Charles, et al.
Published: (2024)
Beyond Visuals: Investigating Force Feedback in Extended Reality for Robot Data Collection
by: Li, Xueyin, et al.
Published: (2025)
by: Li, Xueyin, et al.
Published: (2025)
Generalist Robot Manipulation beyond Action Labeled Data
by: Spiridonov, Alexander, et al.
Published: (2025)
by: Spiridonov, Alexander, et al.
Published: (2025)
Similar Items
-
Scaling Robot Policy Learning via Zero-Shot Labeling with Foundation Models
by: Blank, Nils, et al.
Published: (2024) -
Multimodal Diffusion Transformer: Learning Versatile Behavior from Multimodal Goals
by: Reuss, Moritz, et al.
Published: (2024) -
BEAST: Efficient Tokenization of B-Splines Encoded Action Sequences for Imitation Learning
by: Zhou, Hongyi, et al.
Published: (2025) -
Efficient Diffusion Transformer Policies with Mixture of Expert Denoisers for Multitask Learning
by: Reuss, Moritz, et al.
Published: (2024) -
Open the Black Box: Step-based Policy Updates for Temporally-Correlated Episodic Reinforcement Learning
by: Li, Ge, et al.
Published: (2024)