LaMP: Learning Vision-Language-Action Policies with 3D Scene Flow as Latent Motion Prior
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wang, Xinkai, Wang, Chenyi, Xu, Yifu, Ye, Mingzhe, Zhang, Fu-Cheng, Tian, Jialin, Zhan, Xinyu, Zhu, Lifeng, Lu, Cewu, Yang, Lixin |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
LaMP: Language-Motion Pretraining for Motion Generation, Retrieval, and Captioning
von: Li, Zhe, et al.
Veröffentlicht: (2024)
von: Li, Zhe, et al.
Veröffentlicht: (2024)
Motion Before Action: Diffusing Object Motion as Manipulation Condition
von: Su, Yue, et al.
Veröffentlicht: (2024)
von: Su, Yue, et al.
Veröffentlicht: (2024)
LaMP: When Large Language Models Meet Personalization
von: Salemi, Alireza, et al.
Veröffentlicht: (2023)
von: Salemi, Alireza, et al.
Veröffentlicht: (2023)
LaMP-Val: Large Language Models Empower Personalized Valuation in Auction
von: Sun, Jie, et al.
Veröffentlicht: (2024)
von: Sun, Jie, et al.
Veröffentlicht: (2024)
Dense Policy: Bidirectional Autoregressive Learning of Actions
von: Su, Yue, et al.
Veröffentlicht: (2025)
von: Su, Yue, et al.
Veröffentlicht: (2025)
LaMP-Cap: Personalized Figure Caption Generation With Multimodal Figure Profiles
von: Ng, Ho Yin 'Sam', et al.
Veröffentlicht: (2025)
von: Ng, Ho Yin 'Sam', et al.
Veröffentlicht: (2025)
LaMP-QA: A Benchmark for Personalized Long-form Question Answering
von: Salemi, Alireza, et al.
Veröffentlicht: (2025)
von: Salemi, Alireza, et al.
Veröffentlicht: (2025)
LIDEA: Human-to-Robot Imitation Learning via Implicit Feature Distillation and Explicit Geometry Alignment
von: Xu, Yifu, et al.
Veröffentlicht: (2026)
von: Xu, Yifu, et al.
Veröffentlicht: (2026)
VFM-Recon: Unlocking Cross-Domain Scene-Level Neural Reconstruction with Scale-Aligned Foundation Priors
von: Ming, Yuhang, et al.
Veröffentlicht: (2026)
von: Ming, Yuhang, et al.
Veröffentlicht: (2026)
PriorVLA: Prior-Preserving Adaptation for Vision-Language-Action Models
von: Guo, Xinyu, et al.
Veröffentlicht: (2026)
von: Guo, Xinyu, et al.
Veröffentlicht: (2026)
X-Imitator: Spatial-Aware Imitation Learning via Bidirectional Action-Pose Interaction
von: Xiong, Kai, et al.
Veröffentlicht: (2026)
von: Xiong, Kai, et al.
Veröffentlicht: (2026)
VITA: Vision-to-Action Flow Matching Policy
von: Gao, Dechen, et al.
Veröffentlicht: (2025)
von: Gao, Dechen, et al.
Veröffentlicht: (2025)
LatBot: Distilling Universal Latent Actions for Vision-Language-Action Models
von: Li, Zuolei, et al.
Veröffentlicht: (2025)
von: Li, Zuolei, et al.
Veröffentlicht: (2025)
COIN: Control-Inpainting Diffusion Prior for Human and Camera Motion Estimation
von: Li, Jiefeng, et al.
Veröffentlicht: (2024)
von: Li, Jiefeng, et al.
Veröffentlicht: (2024)
Multi-view Hand Reconstruction with a Point-Embedded Transformer
von: Yang, Lixin, et al.
Veröffentlicht: (2024)
von: Yang, Lixin, et al.
Veröffentlicht: (2024)
L1 Sample Flow for Efficient Visuomotor Learning
von: Song, Weixi, et al.
Veröffentlicht: (2025)
von: Song, Weixi, et al.
Veröffentlicht: (2025)
MP-SfM: Monocular Surface Priors for Robust Structure-from-Motion
von: Pataki, Zador, et al.
Veröffentlicht: (2025)
von: Pataki, Zador, et al.
Veröffentlicht: (2025)
Latent Reasoning VLA: Latent Thinking and Prediction for Vision-Language-Action Models
von: Bai, Shuanghao, et al.
Veröffentlicht: (2026)
von: Bai, Shuanghao, et al.
Veröffentlicht: (2026)
FlowMP: Learning Motion Fields for Robot Planning with Conditional Flow Matching
von: Nguyen, Khang, et al.
Veröffentlicht: (2025)
von: Nguyen, Khang, et al.
Veröffentlicht: (2025)
Latent Spherical Flow Policy for Reinforcement Learning with Combinatorial Actions
von: Kong, Lingkai, et al.
Veröffentlicht: (2026)
von: Kong, Lingkai, et al.
Veröffentlicht: (2026)
Embodiment-Agnostic Action Planning via Object-Part Scene Flow
von: Tang, Weiliang, et al.
Veröffentlicht: (2024)
von: Tang, Weiliang, et al.
Veröffentlicht: (2024)
ReMP: Reusable Motion Prior for Multi-domain 3D Human Pose Estimation and Motion Inbetweening
von: Jang, Hojun, et al.
Veröffentlicht: (2024)
von: Jang, Hojun, et al.
Veröffentlicht: (2024)
MP1: MeanFlow Tames Policy Learning in 1-step for Robotic Manipulation
von: Sheng, Juyi, et al.
Veröffentlicht: (2025)
von: Sheng, Juyi, et al.
Veröffentlicht: (2025)
FlowCoMotion: Text-to-Motion Generation via Token-Latent Flow Modeling
von: Guan, Dawei, et al.
Veröffentlicht: (2026)
von: Guan, Dawei, et al.
Veröffentlicht: (2026)
LaMo: Self-Supervised Latent Motion Priors for Physical Realism in Video Generation
von: Jiang, Bo, et al.
Veröffentlicht: (2026)
von: Jiang, Bo, et al.
Veröffentlicht: (2026)
WarmPrior: Straightening Flow-Matching Policies with Temporal Priors
von: Kang, Sinjae, et al.
Veröffentlicht: (2026)
von: Kang, Sinjae, et al.
Veröffentlicht: (2026)
Latent Policy Steering through One-Step Flow Policies
von: Im, Hokyun, et al.
Veröffentlicht: (2026)
von: Im, Hokyun, et al.
Veröffentlicht: (2026)
OAKINK2: A Dataset of Bimanual Hands-Object Manipulation in Complex Task Completion
von: Zhan, Xinyu, et al.
Veröffentlicht: (2024)
von: Zhan, Xinyu, et al.
Veröffentlicht: (2024)
FlowMamba: Learning Point Cloud Scene Flow with Global Motion Propagation
von: Lin, Min, et al.
Veröffentlicht: (2024)
von: Lin, Min, et al.
Veröffentlicht: (2024)
Seeing Space and Motion: Enhancing Latent Actions with Geometric and Dynamic Awareness for Vision-Language-Action Models
von: Cai, Zhejia, et al.
Veröffentlicht: (2025)
von: Cai, Zhejia, et al.
Veröffentlicht: (2025)
Cross-Hand Latent Representation for Vision-Language-Action Models
von: Jiang, Guangqi, et al.
Veröffentlicht: (2026)
von: Jiang, Guangqi, et al.
Veröffentlicht: (2026)
DISCO: Embodied Navigation and Interaction via Differentiable Scene Semantics and Dual-level Control
von: Xu, Xinyu, et al.
Veröffentlicht: (2024)
von: Xu, Xinyu, et al.
Veröffentlicht: (2024)
Transferable Latent-to-Latent Locomotion Policy for Efficient and Versatile Motion Control of Diverse Legged Robots
von: Zheng, Ziang, et al.
Veröffentlicht: (2025)
von: Zheng, Ziang, et al.
Veröffentlicht: (2025)
Aligning Latent Spaces with Flow Priors
von: Li, Yizhuo, et al.
Veröffentlicht: (2025)
von: Li, Yizhuo, et al.
Veröffentlicht: (2025)
RotVLA: Rotational Latent Action for Vision-Language-Action Model
von: Li, Qiwei, et al.
Veröffentlicht: (2026)
von: Li, Qiwei, et al.
Veröffentlicht: (2026)
On the Identifiability of Latent Action Policies
von: Lachapelle, Sébastien
Veröffentlicht: (2025)
von: Lachapelle, Sébastien
Veröffentlicht: (2025)
DipMe: Haptic Recognition of Granular Media for Tangible Interactive Applications
von: Wang, Xinkai, et al.
Veröffentlicht: (2024)
von: Wang, Xinkai, et al.
Veröffentlicht: (2024)
PALUM: Part-based Attention Learning for Unified Motion Retargeting
von: Liu, Siqi, et al.
Veröffentlicht: (2026)
von: Liu, Siqi, et al.
Veröffentlicht: (2026)
Co-Evolving Latent Action World Models
von: Wang, Yucen, et al.
Veröffentlicht: (2025)
von: Wang, Yucen, et al.
Veröffentlicht: (2025)
Mean Flow Policy with Instantaneous Velocity Constraint for One-step Action Generation
von: Zhan, Guojian, et al.
Veröffentlicht: (2026)
von: Zhan, Guojian, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
LaMP: Language-Motion Pretraining for Motion Generation, Retrieval, and Captioning
von: Li, Zhe, et al.
Veröffentlicht: (2024) -
Motion Before Action: Diffusing Object Motion as Manipulation Condition
von: Su, Yue, et al.
Veröffentlicht: (2024) -
LaMP: When Large Language Models Meet Personalization
von: Salemi, Alireza, et al.
Veröffentlicht: (2023) -
LaMP-Val: Large Language Models Empower Personalized Valuation in Auction
von: Sun, Jie, et al.
Veröffentlicht: (2024) -
Dense Policy: Bidirectional Autoregressive Learning of Actions
von: Su, Yue, et al.
Veröffentlicht: (2025)