Geometry-aware 4D Video Generation for Robot Manipulation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Liu, Zeyi, Li, Shuang, Cousineau, Eric, Feng, Siyuan, Burchfiel, Benjamin, Song, Shuran |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
ManiWAV: Learning Robot Manipulation from In-the-Wild Audio-Visual Data
von: Liu, Zeyi, et al.
Veröffentlicht: (2024)
von: Liu, Zeyi, et al.
Veröffentlicht: (2024)
ContactHandover: Contact-Guided Robot-to-Human Object Handover
von: Wang, Zixi, et al.
Veröffentlicht: (2024)
von: Wang, Zixi, et al.
Veröffentlicht: (2024)
ManipDreamer3D : Synthesizing Plausible Robotic Manipulation Video with Occupancy-aware 3D Trajectory
von: Li, Ying, et al.
Veröffentlicht: (2025)
von: Li, Ying, et al.
Veröffentlicht: (2025)
VideoVLA: Video Generators Can Be Generalizable Robot Manipulators
von: Shen, Yichao, et al.
Veröffentlicht: (2025)
von: Shen, Yichao, et al.
Veröffentlicht: (2025)
Manipulation as in Simulation: Enabling Accurate Geometry Perception in Robots
von: Liu, Minghuan, et al.
Veröffentlicht: (2025)
von: Liu, Minghuan, et al.
Veröffentlicht: (2025)
Robotic Manipulation by Imitating Generated Videos Without Physical Demonstrations
von: Patel, Shivansh, et al.
Veröffentlicht: (2025)
von: Patel, Shivansh, et al.
Veröffentlicht: (2025)
TidyBot++: An Open-Source Holonomic Mobile Manipulator for Robot Learning
von: Wu, Jimmy, et al.
Veröffentlicht: (2024)
von: Wu, Jimmy, et al.
Veröffentlicht: (2024)
Rectified Point Flow: Generic Point Cloud Pose Estimation
von: Sun, Tao, et al.
Veröffentlicht: (2025)
von: Sun, Tao, et al.
Veröffentlicht: (2025)
Universal Manipulation Interface: In-The-Wild Robot Teaching Without In-The-Wild Robots
von: Chi, Cheng, et al.
Veröffentlicht: (2024)
von: Chi, Cheng, et al.
Veröffentlicht: (2024)
RoboVIP: Multi-View Video Generation with Visual Identity Prompting Augments Robot Manipulation
von: Wang, Boyang, et al.
Veröffentlicht: (2026)
von: Wang, Boyang, et al.
Veröffentlicht: (2026)
IRASim: A Fine-Grained World Model for Robot Manipulation
von: Zhu, Fangqi, et al.
Veröffentlicht: (2024)
von: Zhu, Fangqi, et al.
Veröffentlicht: (2024)
BEHAVIOR Robot Suite: Streamlining Real-World Whole-Body Manipulation for Everyday Household Activities
von: Jiang, Yunfan, et al.
Veröffentlicht: (2025)
von: Jiang, Yunfan, et al.
Veröffentlicht: (2025)
This&That: Language-Gesture Controlled Video Generation for Robot Planning
von: Wang, Boyang, et al.
Veröffentlicht: (2024)
von: Wang, Boyang, et al.
Veröffentlicht: (2024)
UAD: Unsupervised Affordance Distillation for Generalization in Robotic Manipulation
von: Tang, Yihe, et al.
Veröffentlicht: (2025)
von: Tang, Yihe, et al.
Veröffentlicht: (2025)
PointWorld: Scaling 3D World Models for In-The-Wild Robotic Manipulation
von: Huang, Wenlong, et al.
Veröffentlicht: (2026)
von: Huang, Wenlong, et al.
Veröffentlicht: (2026)
AnyPlace: Learning Generalized Object Placement for Robot Manipulation
von: Zhao, Yuchi, et al.
Veröffentlicht: (2025)
von: Zhao, Yuchi, et al.
Veröffentlicht: (2025)
SkiP: When to Skip and When to Refine for Efficient Robot Manipulation
von: Dai, Mingtong, et al.
Veröffentlicht: (2026)
von: Dai, Mingtong, et al.
Veröffentlicht: (2026)
SKT: Integrating State-Aware Keypoint Trajectories with Vision-Language Models for Robotic Garment Manipulation
von: Li, Xin, et al.
Veröffentlicht: (2024)
von: Li, Xin, et al.
Veröffentlicht: (2024)
Dream2Flow: Bridging Video Generation and Open-World Manipulation with 3D Object Flow
von: Dharmarajan, Karthik, et al.
Veröffentlicht: (2025)
von: Dharmarajan, Karthik, et al.
Veröffentlicht: (2025)
VideoArtGS: Building Digital Twins of Articulated Objects from Monocular Video
von: Liu, Yu, et al.
Veröffentlicht: (2025)
von: Liu, Yu, et al.
Veröffentlicht: (2025)
Semantically Controllable Augmentations for Generalizable Robot Learning
von: Chen, Zoey, et al.
Veröffentlicht: (2024)
von: Chen, Zoey, et al.
Veröffentlicht: (2024)
GEM-4D: Geometry-Enhanced Video World Models for Robot Manipulation
von: Zhou, Kaichen, et al.
Veröffentlicht: (2026)
von: Zhou, Kaichen, et al.
Veröffentlicht: (2026)
FlowHOI: Flow-based Semantics-Grounded Generation of Hand-Object Interactions for Dexterous Robot Manipulation
von: Zeng, Huajian, et al.
Veröffentlicht: (2026)
von: Zeng, Huajian, et al.
Veröffentlicht: (2026)
Distracted Robot: How Visual Clutter Undermine Robotic Manipulation
von: Rasouli, Amir, et al.
Veröffentlicht: (2025)
von: Rasouli, Amir, et al.
Veröffentlicht: (2025)
Unified Video Action Model
von: Li, Shuang, et al.
Veröffentlicht: (2025)
von: Li, Shuang, et al.
Veröffentlicht: (2025)
Adaptive Compliance Policy: Learning Approximate Compliance for Diffusion Guided Control
von: Hou, Yifan, et al.
Veröffentlicht: (2024)
von: Hou, Yifan, et al.
Veröffentlicht: (2024)
RoBridge: A Hierarchical Architecture Bridging Cognition and Execution for General Robotic Manipulation
von: Zhang, Kaidong, et al.
Veröffentlicht: (2025)
von: Zhang, Kaidong, et al.
Veröffentlicht: (2025)
VG4D: Vision-Language Model Goes 4D Video Recognition
von: Deng, Zhichao, et al.
Veröffentlicht: (2024)
von: Deng, Zhichao, et al.
Veröffentlicht: (2024)
GWM: Towards Scalable Gaussian World Models for Robotic Manipulation
von: Lu, Guanxing, et al.
Veröffentlicht: (2025)
von: Lu, Guanxing, et al.
Veröffentlicht: (2025)
H2R-Grounder: A Paired-Data-Free Paradigm for Translating Human Interaction Videos into Physically Grounded Robot Videos
von: Ci, Hai, et al.
Veröffentlicht: (2025)
von: Ci, Hai, et al.
Veröffentlicht: (2025)
Robots Pre-train Robots: Manipulation-Centric Robotic Representation from Large-Scale Robot Datasets
von: Jiang, Guangqi, et al.
Veröffentlicht: (2024)
von: Jiang, Guangqi, et al.
Veröffentlicht: (2024)
RoomTour3D: Geometry-Aware Video-Instruction Tuning for Embodied Navigation
von: Han, Mingfei, et al.
Veröffentlicht: (2024)
von: Han, Mingfei, et al.
Veröffentlicht: (2024)
ManiSoft: Towards Vision-Language Manipulation for Soft Continuum Robotics
von: Wei, Ziyu, et al.
Veröffentlicht: (2026)
von: Wei, Ziyu, et al.
Veröffentlicht: (2026)
Visual IRL for Human-Like Robotic Manipulation
von: Asali, Ehsan, et al.
Veröffentlicht: (2024)
von: Asali, Ehsan, et al.
Veröffentlicht: (2024)
HomeRobot: Open-Vocabulary Mobile Manipulation
von: Yenamandra, Sriram, et al.
Veröffentlicht: (2023)
von: Yenamandra, Sriram, et al.
Veröffentlicht: (2023)
NovaFlow: Zero-Shot Manipulation via Actionable Flow from Generated Videos
von: Li, Hongyu, et al.
Veröffentlicht: (2025)
von: Li, Hongyu, et al.
Veröffentlicht: (2025)
CL3R: 3D Reconstruction and Contrastive Learning for Enhanced Robotic Manipulation Representations
von: Cui, Wenbo, et al.
Veröffentlicht: (2025)
von: Cui, Wenbo, et al.
Veröffentlicht: (2025)
DynaRend: Learning 3D Dynamics via Masked Future Rendering for Robotic Manipulation
von: Tian, Jingyi, et al.
Veröffentlicht: (2025)
von: Tian, Jingyi, et al.
Veröffentlicht: (2025)
A Real-to-Sim-to-Real Approach to Robotic Manipulation with VLM-Generated Iterative Keypoint Rewards
von: Patel, Shivansh, et al.
Veröffentlicht: (2025)
von: Patel, Shivansh, et al.
Veröffentlicht: (2025)
AIC MLLM: Autonomous Interactive Correction MLLM for Robust Robotic Manipulation
von: Xiong, Chuyan, et al.
Veröffentlicht: (2024)
von: Xiong, Chuyan, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
ManiWAV: Learning Robot Manipulation from In-the-Wild Audio-Visual Data
von: Liu, Zeyi, et al.
Veröffentlicht: (2024) -
ContactHandover: Contact-Guided Robot-to-Human Object Handover
von: Wang, Zixi, et al.
Veröffentlicht: (2024) -
ManipDreamer3D : Synthesizing Plausible Robotic Manipulation Video with Occupancy-aware 3D Trajectory
von: Li, Ying, et al.
Veröffentlicht: (2025) -
VideoVLA: Video Generators Can Be Generalizable Robot Manipulators
von: Shen, Yichao, et al.
Veröffentlicht: (2025) -
Manipulation as in Simulation: Enabling Accurate Geometry Perception in Robots
von: Liu, Minghuan, et al.
Veröffentlicht: (2025)