Spatial-Temporal Aware Visuomotor Diffusion Policy Learning
Fuente:
arXiv
Saved in:
| Main Authors: | Liu, Zhenyang, Wang, Yikai, Wang, Kuanning, Liang, Longfei, Xue, Xiangyang, Fu, Yanwei |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
ActiveVLA: Injecting Active Perception into Vision-Language-Action Models for Precise 3D Robotic Manipulation
by: Liu, Zhenyang, et al.
Published: (2026)
by: Liu, Zhenyang, et al.
Published: (2026)
OFlow: Injecting Object-Aware Temporal Flow Matching for Robust Robotic Manipulation
by: Wang, Kuanning, et al.
Published: (2026)
by: Wang, Kuanning, et al.
Published: (2026)
A Neural Representation Framework with LLM-Driven Spatial Reasoning for Open-Vocabulary 3D Visual Grounding
by: Liu, Zhenyang, et al.
Published: (2025)
by: Liu, Zhenyang, et al.
Published: (2025)
SCOOP'D: Learning Mixed-Liquid-Solid Scooping via Sim2Real Generative Policy
by: Wang, Kuanning, et al.
Published: (2025)
by: Wang, Kuanning, et al.
Published: (2025)
DemoGen: Synthetic Demonstration Generation for Data-Efficient Visuomotor Policy Learning
by: Xue, Zhengrong, et al.
Published: (2025)
by: Xue, Zhengrong, et al.
Published: (2025)
VADF: Vision-Adaptive Diffusion Policy Framework for Efficient Robotic Manipulation
by: Yu, Xinglei, et al.
Published: (2026)
by: Yu, Xinglei, et al.
Published: (2026)
OCRA: Object-Centric Learning with 3D and Tactile Priors for Human-to-Robot Action Transfer
by: Wang, Kuanning, et al.
Published: (2026)
by: Wang, Kuanning, et al.
Published: (2026)
History-Aware Visuomotor Policy Learning via Point Tracking
by: Chen, Jingjing, et al.
Published: (2025)
by: Chen, Jingjing, et al.
Published: (2025)
TriVLA: A Triple-System-Based Unified Vision-Language-Action Model with Episodic World Modeling for General Robot Control
by: Liu, Zhenyang, et al.
Published: (2025)
by: Liu, Zhenyang, et al.
Published: (2025)
Diffusion Policy: Visuomotor Policy Learning via Action Diffusion
by: Chi, Cheng, et al.
Published: (2023)
by: Chi, Cheng, et al.
Published: (2023)
Spatial Policy: Guiding Visuomotor Robotic Manipulation with Spatial-Aware Modeling and Reasoning
by: Liu, Yijun, et al.
Published: (2025)
by: Liu, Yijun, et al.
Published: (2025)
ReasonGrounder: LVLM-Guided Hierarchical Feature Splatting for Open-Vocabulary 3D Visual Grounding and Reasoning
by: Liu, Zhenyang, et al.
Published: (2025)
by: Liu, Zhenyang, et al.
Published: (2025)
Sequential Multi-Object Grasping with One Dexterous Hand
by: He, Sicheng, et al.
Published: (2025)
by: He, Sicheng, et al.
Published: (2025)
H$^3$DP: Triply-Hierarchical Diffusion Policy for Visuomotor Learning
by: Lu, Yiyang, et al.
Published: (2025)
by: Lu, Yiyang, et al.
Published: (2025)
SeFA-Policy: Fast and Accurate Visuomotor Policy Learning with Selective Flow Alignment
by: Xue, Rong, et al.
Published: (2025)
by: Xue, Rong, et al.
Published: (2025)
Choose What to Observe: Task-Aware Semantic-Geometric Representations for Visuomotor Policy
by: Ding, Haoran, et al.
Published: (2026)
by: Ding, Haoran, et al.
Published: (2026)
RAG-6DPose: Retrieval-Augmented 6D Pose Estimation via Leveraging CAD as Knowledge Base
by: Wang, Kuanning, et al.
Published: (2025)
by: Wang, Kuanning, et al.
Published: (2025)
Referring-Aware Visuomotor Policy Learning for Closed-Loop Manipulation
by: Ma, Jiahua, et al.
Published: (2026)
by: Ma, Jiahua, et al.
Published: (2026)
One-Step Diffusion Policy: Fast Visuomotor Policies via Diffusion Distillation
by: Wang, Zhendong, et al.
Published: (2024)
by: Wang, Zhendong, et al.
Published: (2024)
A Latency-Aware Framework for Visuomotor Policy Learning on Industrial Robots
by: Ruan, Daniel, et al.
Published: (2026)
by: Ruan, Daniel, et al.
Published: (2026)
AimBot: A Simple Auxiliary Visual Cue to Enhance Spatial Awareness of Visuomotor Policies
by: Dai, Yinpei, et al.
Published: (2025)
by: Dai, Yinpei, et al.
Published: (2025)
DINOv3-Diffusion Policy: Self-Supervised Large Visual Model for Visuomotor Diffusion Policy Learning
by: Egbe, ThankGod, et al.
Published: (2025)
by: Egbe, ThankGod, et al.
Published: (2025)
Preference Aligned Visuomotor Diffusion Policies for Deformable Object Manipulation
by: Moletta, Marco, et al.
Published: (2026)
by: Moletta, Marco, et al.
Published: (2026)
Hyper-DP3: Frequency-Aware Right-Sizing of 3D Diffusion Policies for Visuomotor Control
by: Zhang, Jinhao, et al.
Published: (2026)
by: Zhang, Jinhao, et al.
Published: (2026)
Hybrid-Diffusion Models: Combining Open-loop Routines with Visuomotor Diffusion Policies
by: Van Haastregt, Jonne, et al.
Published: (2025)
by: Van Haastregt, Jonne, et al.
Published: (2025)
Do You Need Proprioceptive States in Visuomotor Policies?
by: Zhao, Juntu, et al.
Published: (2025)
by: Zhao, Juntu, et al.
Published: (2025)
UMI-on-Air: Embodiment-Aware Guidance for Embodiment-Agnostic Visuomotor Policies
by: Gupta, Harsh, et al.
Published: (2025)
by: Gupta, Harsh, et al.
Published: (2025)
Scalable Multi-Task Reinforcement Learning for Generalizable Spatial Intelligence in Visuomotor Agents
by: Cai, Shaofei, et al.
Published: (2025)
by: Cai, Shaofei, et al.
Published: (2025)
3D Equivariant Visuomotor Policy Learning via Spherical Projection
by: Hu, Boce, et al.
Published: (2025)
by: Hu, Boce, et al.
Published: (2025)
Responsive Noise-Relaying Diffusion Policy: Responsive and Efficient Visuomotor Control
by: Chen, Zhuoqun, et al.
Published: (2025)
by: Chen, Zhuoqun, et al.
Published: (2025)
Falcon: Fast Visuomotor Policies via Partial Denoising
by: Chen, Haojun, et al.
Published: (2025)
by: Chen, Haojun, et al.
Published: (2025)
Latent Policy Barrier: Learning Robust Visuomotor Policies by Staying In-Distribution
by: Sun, Zhanyi, et al.
Published: (2025)
by: Sun, Zhanyi, et al.
Published: (2025)
Real-Time Operator Takeover for Visuomotor Diffusion Policy Training
by: Moletta, Marco, et al.
Published: (2025)
by: Moletta, Marco, et al.
Published: (2025)
Mind the Gap: Learning Implicit Impedance in Visuomotor Policies via Intent-Execution Mismatch
by: Xu, Cuijie, et al.
Published: (2026)
by: Xu, Cuijie, et al.
Published: (2026)
DexKnot: Generalizable Visuomotor Policy Learning for Dexterous Bag-Knotting Manipulation
by: Zhang, Jiayuan, et al.
Published: (2026)
by: Zhang, Jiayuan, et al.
Published: (2026)
A Study on Enhancing the Generalization Ability of Visuomotor Policies via Data Augmentation
by: Wang, Hanwen
Published: (2025)
by: Wang, Hanwen
Published: (2025)
DemoSpeedup: Accelerating Visuomotor Policies via Entropy-Guided Demonstration Acceleration
by: Guo, Lingxiao, et al.
Published: (2025)
by: Guo, Lingxiao, et al.
Published: (2025)
3D Diffusion Policy: Generalizable Visuomotor Policy Learning via Simple 3D Representations
by: Ze, Yanjie, et al.
Published: (2024)
by: Ze, Yanjie, et al.
Published: (2024)
DIPOLE: Fusing Vision and Geometry for Robust Visuomotor Generalization
by: Tang, Yikai, et al.
Published: (2025)
by: Tang, Yikai, et al.
Published: (2025)
Learning to Pick: A Visuomotor Policy for Clustered Strawberry Picking
by: Fei, Zhenghao, et al.
Published: (2025)
by: Fei, Zhenghao, et al.
Published: (2025)
Similar Items
-
ActiveVLA: Injecting Active Perception into Vision-Language-Action Models for Precise 3D Robotic Manipulation
by: Liu, Zhenyang, et al.
Published: (2026) -
OFlow: Injecting Object-Aware Temporal Flow Matching for Robust Robotic Manipulation
by: Wang, Kuanning, et al.
Published: (2026) -
A Neural Representation Framework with LLM-Driven Spatial Reasoning for Open-Vocabulary 3D Visual Grounding
by: Liu, Zhenyang, et al.
Published: (2025) -
SCOOP'D: Learning Mixed-Liquid-Solid Scooping via Sim2Real Generative Policy
by: Wang, Kuanning, et al.
Published: (2025) -
DemoGen: Synthetic Demonstration Generation for Data-Efficient Visuomotor Policy Learning
by: Xue, Zhengrong, et al.
Published: (2025)