UniVLA: Learning to Act Anywhere with Task-centric Latent Actions
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Bu, Qingwen, Yang, Yanting, Cai, Jisong, Gao, Shenyuan, Ren, Guanghui, Yao, Maoqing, Luo, Ping, Li, Hongyang |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Towards Synergistic, Generalized, and Efficient Dual-System for Robotic Manipulation
von: Bu, Qingwen, et al.
Veröffentlicht: (2024)
von: Bu, Qingwen, et al.
Veröffentlicht: (2024)
ACoT-VLA: Action Chain-of-Thought for Vision-Language-Action Models
von: Zhong, Linqing, et al.
Veröffentlicht: (2026)
von: Zhong, Linqing, et al.
Veröffentlicht: (2026)
Seeing Farther and Smarter: Value-Guided Multi-Path Reflection for VLM Policy Optimization
von: Yang, Yanting, et al.
Veröffentlicht: (2026)
von: Yang, Yanting, et al.
Veröffentlicht: (2026)
Libra-VLA: Achieving Learning Equilibrium via Asynchronous Coarse-to-Fine Dual-System
von: Wei, Yifei, et al.
Veröffentlicht: (2026)
von: Wei, Yifei, et al.
Veröffentlicht: (2026)
WholeBodyVLA: Towards Unified Latent VLA for Whole-Body Loco-Manipulation Control
von: Jiang, Haoran, et al.
Veröffentlicht: (2025)
von: Jiang, Haoran, et al.
Veröffentlicht: (2025)
GuidedVLA: Specifying Task-Relevant Factors via Plug-and-Play Action Attention Specialization
von: Jia, Xiaosong, et al.
Veröffentlicht: (2026)
von: Jia, Xiaosong, et al.
Veröffentlicht: (2026)
AdaWorld: Learning Adaptable World Models with Latent Actions
von: Gao, Shenyuan, et al.
Veröffentlicht: (2025)
von: Gao, Shenyuan, et al.
Veröffentlicht: (2025)
Is Diversity All You Need for Scalable Robotic Manipulation?
von: Shi, Modi, et al.
Veröffentlicht: (2025)
von: Shi, Modi, et al.
Veröffentlicht: (2025)
AnywhereVLA: Language-Conditioned Exploration and Mobile Manipulation
von: Gubernatorov, Konstantin, et al.
Veröffentlicht: (2025)
von: Gubernatorov, Konstantin, et al.
Veröffentlicht: (2025)
Joint-Aligned Latent Action: Towards Scalable VLA Pretraining in the Wild
von: Luo, Hao, et al.
Veröffentlicht: (2026)
von: Luo, Hao, et al.
Veröffentlicht: (2026)
Adversarial Data Collection: Human-Collaborative Perturbations for Efficient and Robust Robotic Imitation Learning
von: Huang, Siyuan, et al.
Veröffentlicht: (2025)
von: Huang, Siyuan, et al.
Veröffentlicht: (2025)
Latent Reasoning VLA: Latent Thinking and Prediction for Vision-Language-Action Models
von: Bai, Shuanghao, et al.
Veröffentlicht: (2026)
von: Bai, Shuanghao, et al.
Veröffentlicht: (2026)
Unified Embodied VLM Reasoning with Robotic Action via Autoregressive Discretized Pre-training
von: Liu, Yi, et al.
Veröffentlicht: (2025)
von: Liu, Yi, et al.
Veröffentlicht: (2025)
Imagine2Act: Leveraging Object-Action Motion Consistency from Imagined Goals for Robotic Manipulation
von: Heng, Liang, et al.
Veröffentlicht: (2025)
von: Heng, Liang, et al.
Veröffentlicht: (2025)
ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models
von: Li, Puhao, et al.
Veröffentlicht: (2025)
von: Li, Puhao, et al.
Veröffentlicht: (2025)
SwitchVLA: Execution-Aware Task Switching for Vision-Language-Action Models
von: Li, Meng, et al.
Veröffentlicht: (2025)
von: Li, Meng, et al.
Veröffentlicht: (2025)
Learning to Act Robustly with View-Invariant Latent Actions
von: Jeong, Youngjoon, et al.
Veröffentlicht: (2026)
von: Jeong, Youngjoon, et al.
Veröffentlicht: (2026)
Hume: Introducing System-2 Thinking in Visual-Language-Action Model
von: Song, Haoming, et al.
Veröffentlicht: (2025)
von: Song, Haoming, et al.
Veröffentlicht: (2025)
EWMBench: Evaluating Scene, Motion, and Semantic Quality in Embodied World Models
von: Yue, Hu, et al.
Veröffentlicht: (2025)
von: Yue, Hu, et al.
Veröffentlicht: (2025)
RotVLA: Rotational Latent Action for Vision-Language-Action Model
von: Li, Qiwei, et al.
Veröffentlicht: (2026)
von: Li, Qiwei, et al.
Veröffentlicht: (2026)
UniDriveVLA: Unifying Understanding, Perception, and Action Planning for Autonomous Driving
von: Li, Yongkang, et al.
Veröffentlicht: (2026)
von: Li, Yongkang, et al.
Veröffentlicht: (2026)
VLA-JEPA: Enhancing Vision-Language-Action Model with Latent World Model
von: Sun, Jingwen, et al.
Veröffentlicht: (2026)
von: Sun, Jingwen, et al.
Veröffentlicht: (2026)
LAOF: Robust Latent Action Learning with Optical Flow Constraints
von: Bu, Xizhou, et al.
Veröffentlicht: (2025)
von: Bu, Xizhou, et al.
Veröffentlicht: (2025)
EnerVerse-AC: Envisioning Embodied Environments with Action Condition
von: Jiang, Yuxin, et al.
Veröffentlicht: (2025)
von: Jiang, Yuxin, et al.
Veröffentlicht: (2025)
UniAct: Unified Motion Generation and Action Streaming for Humanoid Robots
von: Jiang, Nan, et al.
Veröffentlicht: (2025)
von: Jiang, Nan, et al.
Veröffentlicht: (2025)
Closed-Loop Visuomotor Control with Generative Expectation for Robotic Manipulation
von: Bu, Qingwen, et al.
Veröffentlicht: (2024)
von: Bu, Qingwen, et al.
Veröffentlicht: (2024)
UniLACT: Depth-Aware RGB Latent Action Learning for Vision-Language-Action Models
von: Govind, Manish Kumar, et al.
Veröffentlicht: (2026)
von: Govind, Manish Kumar, et al.
Veröffentlicht: (2026)
AT-VLA: Adaptive Tactile Injection for Enhanced Feedback Reaction in Vision-Language-Action Models
von: Li, Xiaoqi, et al.
Veröffentlicht: (2026)
von: Li, Xiaoqi, et al.
Veröffentlicht: (2026)
ForeAct: Steering Your VLA with Efficient Visual Foresight Planning
von: Zhang, Zhuoyang, et al.
Veröffentlicht: (2026)
von: Zhang, Zhuoyang, et al.
Veröffentlicht: (2026)
Agility Meets Stability: Versatile Humanoid Control with Heterogeneous Data
von: Pan, Yixuan, et al.
Veröffentlicht: (2025)
von: Pan, Yixuan, et al.
Veröffentlicht: (2025)
RedVLA: Physical Red Teaming for Vision-Language-Action Models
von: Zhang, Yuhao, et al.
Veröffentlicht: (2026)
von: Zhang, Yuhao, et al.
Veröffentlicht: (2026)
Continually Evolving Skill Knowledge in Vision Language Action Model
von: Wu, Yuxuan, et al.
Veröffentlicht: (2025)
von: Wu, Yuxuan, et al.
Veröffentlicht: (2025)
Discrete Diffusion VLA: Bringing Discrete Diffusion to Action Decoding in Vision-Language-Action Policies
von: Liang, Zhixuan, et al.
Veröffentlicht: (2025)
von: Liang, Zhixuan, et al.
Veröffentlicht: (2025)
See Once, Then Act: Vision-Language-Action Model with Task Learning from One-Shot Video Demonstrations
von: Chen, Guangyan, et al.
Veröffentlicht: (2025)
von: Chen, Guangyan, et al.
Veröffentlicht: (2025)
NS-VLA: Towards Neuro-Symbolic Vision-Language-Action Models
von: Zhu, Ziyue, et al.
Veröffentlicht: (2026)
von: Zhu, Ziyue, et al.
Veröffentlicht: (2026)
RoVLA: Multi-Consistency Constraints for Robust Vision-Language-Action Models
von: Luo, Jingzhou, et al.
Veröffentlicht: (2026)
von: Luo, Jingzhou, et al.
Veröffentlicht: (2026)
DreamAvoid: Critical-Phase Test-Time Dreaming to Avoid Failures in VLA Policies
von: Fan, Xianzhe, et al.
Veröffentlicht: (2026)
von: Fan, Xianzhe, et al.
Veröffentlicht: (2026)
StreamVLA: Breaking the Reason-Act Cycle via Completion-State Gating
von: Chen, Tongqing, et al.
Veröffentlicht: (2026)
von: Chen, Tongqing, et al.
Veröffentlicht: (2026)
Act2Goal: From World Model To General Goal-conditioned Policy
von: Zhou, Pengfei, et al.
Veröffentlicht: (2025)
von: Zhou, Pengfei, et al.
Veröffentlicht: (2025)
AR-VLA: True Autoregressive Action Expert for Vision-Language-Action Models
von: Hu, Yutong, et al.
Veröffentlicht: (2026)
von: Hu, Yutong, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Towards Synergistic, Generalized, and Efficient Dual-System for Robotic Manipulation
von: Bu, Qingwen, et al.
Veröffentlicht: (2024) -
ACoT-VLA: Action Chain-of-Thought for Vision-Language-Action Models
von: Zhong, Linqing, et al.
Veröffentlicht: (2026) -
Seeing Farther and Smarter: Value-Guided Multi-Path Reflection for VLM Policy Optimization
von: Yang, Yanting, et al.
Veröffentlicht: (2026) -
Libra-VLA: Achieving Learning Equilibrium via Asynchronous Coarse-to-Fine Dual-System
von: Wei, Yifei, et al.
Veröffentlicht: (2026) -
WholeBodyVLA: Towards Unified Latent VLA for Whole-Body Loco-Manipulation Control
von: Jiang, Haoran, et al.
Veröffentlicht: (2025)