Action with Visual Primitives
Fuente:
arXiv
Saved in:
| Main Authors: | Guo, Weilong, Wang, Yuchen, Zhou, Renping, Zhang, Yunfeng, Fang, Rui, Pang, Yuyang, Xu, Wenda, Huang, Gao |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
FODMP: Fast One-Step Diffusion of Movement Primitives Generation for Time-Dependent Robot Actions
by: Shi, Xirui, et al.
Published: (2026)
by: Shi, Xirui, et al.
Published: (2026)
FRMD: Fast Robot Motion Diffusion with Consistency-Distilled Movement Primitives for Smooth Action Generation
by: Shi, Xirui, et al.
Published: (2025)
by: Shi, Xirui, et al.
Published: (2025)
Prediction with Action: Visual Policy Learning via Joint Denoising Process
by: Guo, Yanjiang, et al.
Published: (2024)
by: Guo, Yanjiang, et al.
Published: (2024)
A Vision-Language-Action-Critic Model for Robotic Real-World Reinforcement Learning
by: Zhai, Shaopeng, et al.
Published: (2025)
by: Zhai, Shaopeng, et al.
Published: (2025)
When to Trust Imagination: Adaptive Action Execution for World Action Models
by: Wang, Rui, et al.
Published: (2026)
by: Wang, Rui, et al.
Published: (2026)
Traj2Action: A Co-Denoising Framework for Trajectory-Guided Human-to-Robot Skill Transfer
by: Zhou, Han, et al.
Published: (2025)
by: Zhou, Han, et al.
Published: (2025)
From Observation to Action: Latent Action-based Primitive Segmentation for VLA Pre-training in Industrial Settings
by: Zhang, Jiajie, et al.
Published: (2025)
by: Zhang, Jiajie, et al.
Published: (2025)
Scenarios Engineering driven Autonomous Transportation in Open-Pit Mines
by: Teng, Siyu, et al.
Published: (2024)
by: Teng, Siyu, et al.
Published: (2024)
Inspection Planning Primitives with Implicit Models
by: You, Jingyang, et al.
Published: (2025)
by: You, Jingyang, et al.
Published: (2025)
ResWM: Residual-Action World Model for Visual RL
by: Zhang, Jseen, et al.
Published: (2026)
by: Zhang, Jseen, et al.
Published: (2026)
Pure Vision Language Action (VLA) Models: A Comprehensive Survey
by: Zhang, Dapeng, et al.
Published: (2025)
by: Zhang, Dapeng, et al.
Published: (2025)
Expertise need not monopolize: Action-Specialized Mixture of Experts for Vision-Language-Action Learning
by: Shen, Weijie, et al.
Published: (2025)
by: Shen, Weijie, et al.
Published: (2025)
Movement Primitives in Robotics: A Comprehensive Survey
by: Gutierrez, Nolan B., et al.
Published: (2025)
by: Gutierrez, Nolan B., et al.
Published: (2025)
KineVLA: Towards Kinematics-Aware Vision-Language-Action Models with Bi-Level Action Decomposition
by: Han, Gaoge, et al.
Published: (2026)
by: Han, Gaoge, et al.
Published: (2026)
Learning Primitive Embodied World Models: Towards Scalable Robotic Learning
by: Sun, Qiao, et al.
Published: (2025)
by: Sun, Qiao, et al.
Published: (2025)
Embodiment-Aware Generalist Specialist Distillation for Unified Humanoid Whole-Body Control
by: Peng, Quanquan, et al.
Published: (2026)
by: Peng, Quanquan, et al.
Published: (2026)
FUNCanon: Learning Pose-Aware Action Primitives via Functional Object Canonicalization for Generalizable Robotic Manipulation
by: Xu, Hongli, et al.
Published: (2025)
by: Xu, Hongli, et al.
Published: (2025)
Action-aware Dynamic Pruning for Efficient Vision-Language-Action Manipulation
by: Pei, Xiaohuan, et al.
Published: (2025)
by: Pei, Xiaohuan, et al.
Published: (2025)
SpatialVLA: Exploring Spatial Representations for Visual-Language-Action Model
by: Qu, Delin, et al.
Published: (2025)
by: Qu, Delin, et al.
Published: (2025)
FusionPlanner: A Multi-task Motion Planner for Mining Trucks via Multi-sensor Fusion
by: Teng, Siyu, et al.
Published: (2023)
by: Teng, Siyu, et al.
Published: (2023)
ManipDreamer: Boosting Robotic Manipulation World Model with Action Tree and Visual Guidance
by: Li, Ying, et al.
Published: (2025)
by: Li, Ying, et al.
Published: (2025)
HACMan++: Spatially-Grounded Motion Primitives for Manipulation
by: Jiang, Bowen, et al.
Published: (2024)
by: Jiang, Bowen, et al.
Published: (2024)
MeshA*: Efficient Path Planning With Motion Primitives
by: Agranovskiy, Marat, et al.
Published: (2024)
by: Agranovskiy, Marat, et al.
Published: (2024)
Action-to-Action Flow Matching
by: Jia, Jindou, et al.
Published: (2026)
by: Jia, Jindou, et al.
Published: (2026)
Learning Native Continuation for Action Chunking Flow Policies
by: Liu, Yufeng, et al.
Published: (2026)
by: Liu, Yufeng, et al.
Published: (2026)
Towards Backdoor-Based Ownership Verification for Vision-Language-Action Models
by: Sun, Ming, et al.
Published: (2026)
by: Sun, Ming, et al.
Published: (2026)
Hume: Introducing System-2 Thinking in Visual-Language-Action Model
by: Song, Haoming, et al.
Published: (2025)
by: Song, Haoming, et al.
Published: (2025)
GraspGF: Learning Score-based Grasping Primitive for Human-assisting Dexterous Grasping
by: Wu, Tianhao, et al.
Published: (2023)
by: Wu, Tianhao, et al.
Published: (2023)
Obstacle Avoidance using Dynamic Movement Primitives and Reinforcement Learning
by: Urbaniak, Dominik, et al.
Published: (2025)
by: Urbaniak, Dominik, et al.
Published: (2025)
Sim-and-Human Co-training for Data-Efficient and Generalizable Robotic Manipulation
by: Fang, Kaipeng, et al.
Published: (2026)
by: Fang, Kaipeng, et al.
Published: (2026)
SA-VLA: Spatially-Aware Flow-Matching for Vision-Language-Action Reinforcement Learning
by: Pan, Xu, et al.
Published: (2026)
by: Pan, Xu, et al.
Published: (2026)
QDTraj: Exploration of Diverse Trajectory Primitives for Articulated Objects Robotic Manipulation
by: Kappel, Mathilde, et al.
Published: (2026)
by: Kappel, Mathilde, et al.
Published: (2026)
Wonderful Team: Zero-Shot Physical Task Planning with Visual LLMs
by: Wang, Zidan, et al.
Published: (2024)
by: Wang, Zidan, et al.
Published: (2024)
DeMaVLA: A Vision-Language-Action Foundation Model for Generalizable Deformable Manipulation
by: Su, Taiyi, et al.
Published: (2026)
by: Su, Taiyi, et al.
Published: (2026)
LADEV: A Language-Driven Testing and Evaluation Platform for Vision-Language-Action Models in Robotic Manipulation
by: Wang, Zhijie, et al.
Published: (2024)
by: Wang, Zhijie, et al.
Published: (2024)
PRIME: Scaffolding Manipulation Tasks with Behavior Primitives for Data-Efficient Imitation Learning
by: Gao, Tian, et al.
Published: (2024)
by: Gao, Tian, et al.
Published: (2024)
AttenA+: Rectifying Action Inequality in Robotic Foundation Models
by: Peng, Daojie, et al.
Published: (2026)
by: Peng, Daojie, et al.
Published: (2026)
ActionFlow: A Pipelined Action Acceleration for Vision Language Models on Edge
by: Dai, Yuntao, et al.
Published: (2025)
by: Dai, Yuntao, et al.
Published: (2025)
Differentiable Motion Manifold Primitives for Reactive Motion Generation under Kinodynamic Constraints
by: Lee, Yonghyeon
Published: (2024)
by: Lee, Yonghyeon
Published: (2024)
WorldVLA: Towards Autoregressive Action World Model
by: Cen, Jun, et al.
Published: (2025)
by: Cen, Jun, et al.
Published: (2025)
Similar Items
-
FODMP: Fast One-Step Diffusion of Movement Primitives Generation for Time-Dependent Robot Actions
by: Shi, Xirui, et al.
Published: (2026) -
FRMD: Fast Robot Motion Diffusion with Consistency-Distilled Movement Primitives for Smooth Action Generation
by: Shi, Xirui, et al.
Published: (2025) -
Prediction with Action: Visual Policy Learning via Joint Denoising Process
by: Guo, Yanjiang, et al.
Published: (2024) -
A Vision-Language-Action-Critic Model for Robotic Real-World Reinforcement Learning
by: Zhai, Shaopeng, et al.
Published: (2025) -
When to Trust Imagination: Adaptive Action Execution for World Action Models
by: Wang, Rui, et al.
Published: (2026)