AnchorDream: Repurposing Video Diffusion for Embodiment-Aware Robot Data Synthesis
Fuente:
arXiv
Saved in:
| Main Authors: | Ye, Junjie, Xue, Rong, Van Hoorick, Basile, Tokmakov, Pavel, Irshad, Muhammad Zubair, Wang, Yue, Guizilini, Vitor |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
RoboDream: Compositional World Models for Scalable Robot Data Synthesis
by: Ye, Junjie, et al.
Published: (2026)
by: Ye, Junjie, et al.
Published: (2026)
Fiducial Exoskeletons: Image-Centric Robot State Estimation
by: Smith, Cameron, et al.
Published: (2026)
by: Smith, Cameron, et al.
Published: (2026)
AnyView: Synthesizing Any Novel View in Dynamic Scenes
by: Van Hoorick, Basile, et al.
Published: (2026)
by: Van Hoorick, Basile, et al.
Published: (2026)
SIRE: SE(3) Intrinsic Rigidity Embeddings
by: Smith, Cameron, et al.
Published: (2025)
by: Smith, Cameron, et al.
Published: (2025)
GRIN: Zero-Shot Metric Depth with Pixel-Level Diffusion
by: Guizilini, Vitor, et al.
Published: (2024)
by: Guizilini, Vitor, et al.
Published: (2024)
Zero-Shot Novel View and Depth Synthesis with Multi-View Geometric Diffusion
by: Guizilini, Vitor, et al.
Published: (2025)
by: Guizilini, Vitor, et al.
Published: (2025)
Generative Camera Dolly: Extreme Monocular Dynamic Novel View Synthesis
by: Van Hoorick, Basile, et al.
Published: (2024)
by: Van Hoorick, Basile, et al.
Published: (2024)
Learning 3D Robotics Perception using Inductive Priors
by: Irshad, Muhammad Zubair
Published: (2024)
by: Irshad, Muhammad Zubair
Published: (2024)
NeRF-MAE: Masked AutoEncoders for Self-Supervised 3D Representation Learning for Neural Radiance Fields
by: Irshad, Muhammad Zubair, et al.
Published: (2024)
by: Irshad, Muhammad Zubair, et al.
Published: (2024)
DreamPlan: Efficient Reinforcement Fine-Tuning of Vision-Language Planners via Video World Models
by: Jia, Emily Yue-Ting, et al.
Published: (2026)
by: Jia, Emily Yue-Ting, et al.
Published: (2026)
RoVi-Aug: Robot and Viewpoint Augmentation for Cross-Embodiment Robot Learning
by: Chen, Lawrence Yunliang, et al.
Published: (2024)
by: Chen, Lawrence Yunliang, et al.
Published: (2024)
Towards Realistic Scene Generation with LiDAR Diffusion Models
by: Ran, Haoxi, et al.
Published: (2024)
by: Ran, Haoxi, et al.
Published: (2024)
FastMap: Revisiting Structure from Motion through First-Order Optimization
by: Li, Jiahao, et al.
Published: (2025)
by: Li, Jiahao, et al.
Published: (2025)
SplArt: Articulation Estimation and Part-Level Reconstruction with 3D Gaussian Splatting
by: Lin, Shengjie, et al.
Published: (2025)
by: Lin, Shengjie, et al.
Published: (2025)
ZeroGrasp: Zero-Shot Shape Reconstruction Enabled Robotic Grasping
by: Iwase, Shun, et al.
Published: (2025)
by: Iwase, Shun, et al.
Published: (2025)
Video Generators are Robot Policies
by: Liang, Junbang, et al.
Published: (2025)
by: Liang, Junbang, et al.
Published: (2025)
Controlling the World by Sleight of Hand
by: Sudhakar, Sruthi, et al.
Published: (2024)
by: Sudhakar, Sruthi, et al.
Published: (2024)
Learning Temporally Consistent Video Depth from Video Diffusion Priors
by: Shao, Jiahao, et al.
Published: (2024)
by: Shao, Jiahao, et al.
Published: (2024)
EmbodiedSplat: Personalized Real-to-Sim-to-Real Navigation with Gaussian Splats from a Mobile Device
by: Chhablani, Gunjan, et al.
Published: (2025)
by: Chhablani, Gunjan, et al.
Published: (2025)
Tar for Mortar: "The Library of Babel" and the Dream of Totality
by: Basile, Jonathan
Published: (2019)
by: Basile, Jonathan
Published: (2019)
Learning from Massive Human Videos for Universal Humanoid Pose Control
by: Mao, Jiageng, et al.
Published: (2024)
by: Mao, Jiageng, et al.
Published: (2024)
Humanoid Everyday: A Comprehensive Robotic Dataset for Open-World Humanoid Manipulation
by: Zhao, Zhenyu, et al.
Published: (2025)
by: Zhao, Zhenyu, et al.
Published: (2025)
Large Reward Models: Generalizable Online Robot Reward Generation with Vision-Language Models
by: Wu, Yanru, et al.
Published: (2026)
by: Wu, Yanru, et al.
Published: (2026)
Bridging the Embodiment Gap: Disentangled Cross-Embodiment Video Editing
by: Li, Zhiyuan, et al.
Published: (2026)
by: Li, Zhiyuan, et al.
Published: (2026)
DreamShot: Personalized Storyboard Synthesis with Video Diffusion Prior
by: Huang, Junjia, et al.
Published: (2026)
by: Huang, Junjia, et al.
Published: (2026)
ReferEverything: Towards Segmenting Everything We Can Speak of in Videos
by: Bagchi, Anurag, et al.
Published: (2024)
by: Bagchi, Anurag, et al.
Published: (2024)
EscherNet++: Simultaneous Amodal Completion and Scalable View Synthesis through Masked Fine-Tuning and Enhanced Feed-Forward 3D Reconstruction
by: Zhang, Xinan, et al.
Published: (2025)
by: Zhang, Xinan, et al.
Published: (2025)
Understanding Video Transformers via Universal Concept Discovery
by: Kowal, Matthew, et al.
Published: (2024)
by: Kowal, Matthew, et al.
Published: (2024)
Tenma: Robust Cross-Embodiment Robot Manipulation with Diffusion Transformer
by: Davies, Travis, et al.
Published: (2025)
by: Davies, Travis, et al.
Published: (2025)
Repurposing Video Diffusion Transformers for Robust Point Tracking
by: Son, Soowon, et al.
Published: (2025)
by: Son, Soowon, et al.
Published: (2025)
Robot Learning from Any Images
by: Zhao, Siheng, et al.
Published: (2025)
by: Zhao, Siheng, et al.
Published: (2025)
Diffusion Knows Transparency: Repurposing Video Diffusion for Transparent Object Depth and Normal Estimation
by: Xu, Shaocong, et al.
Published: (2025)
by: Xu, Shaocong, et al.
Published: (2025)
Pleural subxyphoid drain confers better pulmonary function and clinical outcomes in chronic obstructive pulmonary disease after off-pump coronary artery bypass grafting: a randomized controlled trial
by: Solange Guizilini
Published: (2014)
by: Solange Guizilini
Published: (2014)
Generative 4D Scene Gaussian Splatting with Object View-Synthesis Priors
by: Chu, Wen-Hsuan, et al.
Published: (2025)
by: Chu, Wen-Hsuan, et al.
Published: (2025)
Self-Supervised Geometry-Guided Initialization for Robust Monocular Visual Odometry
by: Kanai, Takayuki, et al.
Published: (2024)
by: Kanai, Takayuki, et al.
Published: (2024)
Neural Fields in Robotics: A Survey
by: Irshad, Muhammad Zubair, et al.
Published: (2024)
by: Irshad, Muhammad Zubair, et al.
Published: (2024)
Repurposing Pre-trained Video Diffusion Models for Event-based Video Interpolation
by: Chen, Jingxi, et al.
Published: (2024)
by: Chen, Jingxi, et al.
Published: (2024)
UMI-on-Air: Embodiment-Aware Guidance for Embodiment-Agnostic Visuomotor Policies
by: Gupta, Harsh, et al.
Published: (2025)
by: Gupta, Harsh, et al.
Published: (2025)
Real2Render2Real: Scaling Robot Data Without Dynamics Simulation or Robot Hardware
by: Yu, Justin, et al.
Published: (2025)
by: Yu, Justin, et al.
Published: (2025)
Capturing Visual Environment Structure Correlates with Control Performance
by: Dong, Jiahua, et al.
Published: (2026)
by: Dong, Jiahua, et al.
Published: (2026)
Similar Items
-
RoboDream: Compositional World Models for Scalable Robot Data Synthesis
by: Ye, Junjie, et al.
Published: (2026) -
Fiducial Exoskeletons: Image-Centric Robot State Estimation
by: Smith, Cameron, et al.
Published: (2026) -
AnyView: Synthesizing Any Novel View in Dynamic Scenes
by: Van Hoorick, Basile, et al.
Published: (2026) -
SIRE: SE(3) Intrinsic Rigidity Embeddings
by: Smith, Cameron, et al.
Published: (2025) -
GRIN: Zero-Shot Metric Depth with Pixel-Level Diffusion
by: Guizilini, Vitor, et al.
Published: (2024)