RoboDream: Compositional World Models for Scalable Robot Data Synthesis
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Ye, Junjie, Xue, Rong, Van Hoorick, Basile, Li, Runhao, Rajaprakash, Harshitha, Tokmakov, Pavel, Irshad, Muhammad Zubair, Guizilini, Vitor, Wang, Yue |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
AnchorDream: Repurposing Video Diffusion for Embodiment-Aware Robot Data Synthesis
von: Ye, Junjie, et al.
Veröffentlicht: (2025)
von: Ye, Junjie, et al.
Veröffentlicht: (2025)
Fiducial Exoskeletons: Image-Centric Robot State Estimation
von: Smith, Cameron, et al.
Veröffentlicht: (2026)
von: Smith, Cameron, et al.
Veröffentlicht: (2026)
AnyView: Synthesizing Any Novel View in Dynamic Scenes
von: Van Hoorick, Basile, et al.
Veröffentlicht: (2026)
von: Van Hoorick, Basile, et al.
Veröffentlicht: (2026)
SIRE: SE(3) Intrinsic Rigidity Embeddings
von: Smith, Cameron, et al.
Veröffentlicht: (2025)
von: Smith, Cameron, et al.
Veröffentlicht: (2025)
RoboDreamer: Learning Compositional World Models for Robot Imagination
von: Zhou, Siyuan, et al.
Veröffentlicht: (2024)
von: Zhou, Siyuan, et al.
Veröffentlicht: (2024)
GRIN: Zero-Shot Metric Depth with Pixel-Level Diffusion
von: Guizilini, Vitor, et al.
Veröffentlicht: (2024)
von: Guizilini, Vitor, et al.
Veröffentlicht: (2024)
Zero-Shot Novel View and Depth Synthesis with Multi-View Geometric Diffusion
von: Guizilini, Vitor, et al.
Veröffentlicht: (2025)
von: Guizilini, Vitor, et al.
Veröffentlicht: (2025)
Generative Camera Dolly: Extreme Monocular Dynamic Novel View Synthesis
von: Van Hoorick, Basile, et al.
Veröffentlicht: (2024)
von: Van Hoorick, Basile, et al.
Veröffentlicht: (2024)
Learning 3D Robotics Perception using Inductive Priors
von: Irshad, Muhammad Zubair
Veröffentlicht: (2024)
von: Irshad, Muhammad Zubair
Veröffentlicht: (2024)
NeRF-MAE: Masked AutoEncoders for Self-Supervised 3D Representation Learning for Neural Radiance Fields
von: Irshad, Muhammad Zubair, et al.
Veröffentlicht: (2024)
von: Irshad, Muhammad Zubair, et al.
Veröffentlicht: (2024)
Controlling the World by Sleight of Hand
von: Sudhakar, Sruthi, et al.
Veröffentlicht: (2024)
von: Sudhakar, Sruthi, et al.
Veröffentlicht: (2024)
DreamPlan: Efficient Reinforcement Fine-Tuning of Vision-Language Planners via Video World Models
von: Jia, Emily Yue-Ting, et al.
Veröffentlicht: (2026)
von: Jia, Emily Yue-Ting, et al.
Veröffentlicht: (2026)
FastMap: Revisiting Structure from Motion through First-Order Optimization
von: Li, Jiahao, et al.
Veröffentlicht: (2025)
von: Li, Jiahao, et al.
Veröffentlicht: (2025)
Humanoid Everyday: A Comprehensive Robotic Dataset for Open-World Humanoid Manipulation
von: Zhao, Zhenyu, et al.
Veröffentlicht: (2025)
von: Zhao, Zhenyu, et al.
Veröffentlicht: (2025)
SplArt: Articulation Estimation and Part-Level Reconstruction with 3D Gaussian Splatting
von: Lin, Shengjie, et al.
Veröffentlicht: (2025)
von: Lin, Shengjie, et al.
Veröffentlicht: (2025)
ZeroGrasp: Zero-Shot Shape Reconstruction Enabled Robotic Grasping
von: Iwase, Shun, et al.
Veröffentlicht: (2025)
von: Iwase, Shun, et al.
Veröffentlicht: (2025)
Towards Realistic Scene Generation with LiDAR Diffusion Models
von: Ran, Haoxi, et al.
Veröffentlicht: (2024)
von: Ran, Haoxi, et al.
Veröffentlicht: (2024)
HAND Me the Data: Fast Robot Adaptation via Hand Path Retrieval
von: Hong, Matthew, et al.
Veröffentlicht: (2025)
von: Hong, Matthew, et al.
Veröffentlicht: (2025)
EscherNet++: Simultaneous Amodal Completion and Scalable View Synthesis through Masked Fine-Tuning and Enhanced Feed-Forward 3D Reconstruction
von: Zhang, Xinan, et al.
Veröffentlicht: (2025)
von: Zhang, Xinan, et al.
Veröffentlicht: (2025)
Tar for Mortar: "The Library of Babel" and the Dream of Totality
von: Basile, Jonathan
Veröffentlicht: (2019)
von: Basile, Jonathan
Veröffentlicht: (2019)
PhysBench: Benchmarking and Enhancing Vision-Language Models for Physical World Understanding
von: Chow, Wei, et al.
Veröffentlicht: (2025)
von: Chow, Wei, et al.
Veröffentlicht: (2025)
Large Reward Models: Generalizable Online Robot Reward Generation with Vision-Language Models
von: Wu, Yanru, et al.
Veröffentlicht: (2026)
von: Wu, Yanru, et al.
Veröffentlicht: (2026)
PolaRiS: Scalable Real-to-Sim Evaluations for Generalist Robot Policies
von: Jain, Arhan, et al.
Veröffentlicht: (2025)
von: Jain, Arhan, et al.
Veröffentlicht: (2025)
Robot Learning from a Physical World Model
von: Mao, Jiageng, et al.
Veröffentlicht: (2025)
von: Mao, Jiageng, et al.
Veröffentlicht: (2025)
Dream to Manipulate: Compositional World Models Empowering Robot Imitation Learning with Imagination
von: Barcellona, Leonardo, et al.
Veröffentlicht: (2024)
von: Barcellona, Leonardo, et al.
Veröffentlicht: (2024)
RoboPocket: Improve Robot Policies Instantly with Your Phone
von: Fang, Junjie, et al.
Veröffentlicht: (2026)
von: Fang, Junjie, et al.
Veröffentlicht: (2026)
EmbodiedSplat: Personalized Real-to-Sim-to-Real Navigation with Gaussian Splats from a Mobile Device
von: Chhablani, Gunjan, et al.
Veröffentlicht: (2025)
von: Chhablani, Gunjan, et al.
Veröffentlicht: (2025)
RoboMatrix: A Skill-centric Hierarchical Framework for Scalable Robot Task Planning and Execution in Open-World
von: Mao, Weixin, et al.
Veröffentlicht: (2024)
von: Mao, Weixin, et al.
Veröffentlicht: (2024)
Video Generators are Robot Policies
von: Liang, Junbang, et al.
Veröffentlicht: (2025)
von: Liang, Junbang, et al.
Veröffentlicht: (2025)
RoboEval: Where Robotic Manipulation Meets Structured and Scalable Evaluation
von: Wang, Yi Ru, et al.
Veröffentlicht: (2025)
von: Wang, Yi Ru, et al.
Veröffentlicht: (2025)
RoboClaw: An Agentic Framework for Scalable Long-Horizon Robotic Tasks
von: Li, Ruiying, et al.
Veröffentlicht: (2026)
von: Li, Ruiying, et al.
Veröffentlicht: (2026)
Walk through Paintings: Egocentric World Models from Internet Priors
von: Bagchi, Anurag, et al.
Veröffentlicht: (2026)
von: Bagchi, Anurag, et al.
Veröffentlicht: (2026)
Robot Learning from Any Images
von: Zhao, Siheng, et al.
Veröffentlicht: (2025)
von: Zhao, Siheng, et al.
Veröffentlicht: (2025)
Pleural subxyphoid drain confers better pulmonary function and clinical outcomes in chronic obstructive pulmonary disease after off-pump coronary artery bypass grafting: a randomized controlled trial
von: Solange Guizilini
Veröffentlicht: (2014)
von: Solange Guizilini
Veröffentlicht: (2014)
Generative 4D Scene Gaussian Splatting with Object View-Synthesis Priors
von: Chu, Wen-Hsuan, et al.
Veröffentlicht: (2025)
von: Chu, Wen-Hsuan, et al.
Veröffentlicht: (2025)
Self-Supervised Geometry-Guided Initialization for Robust Monocular Visual Odometry
von: Kanai, Takayuki, et al.
Veröffentlicht: (2024)
von: Kanai, Takayuki, et al.
Veröffentlicht: (2024)
Neural Fields in Robotics: A Survey
von: Irshad, Muhammad Zubair, et al.
Veröffentlicht: (2024)
von: Irshad, Muhammad Zubair, et al.
Veröffentlicht: (2024)
RoboArena: Distributed Real-World Evaluation of Generalist Robot Policies
von: Atreya, Pranav, et al.
Veröffentlicht: (2025)
von: Atreya, Pranav, et al.
Veröffentlicht: (2025)
Real2Render2Real: Scaling Robot Data Without Dynamics Simulation or Robot Hardware
von: Yu, Justin, et al.
Veröffentlicht: (2025)
von: Yu, Justin, et al.
Veröffentlicht: (2025)
Capturing Visual Environment Structure Correlates with Control Performance
von: Dong, Jiahua, et al.
Veröffentlicht: (2026)
von: Dong, Jiahua, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
AnchorDream: Repurposing Video Diffusion for Embodiment-Aware Robot Data Synthesis
von: Ye, Junjie, et al.
Veröffentlicht: (2025) -
Fiducial Exoskeletons: Image-Centric Robot State Estimation
von: Smith, Cameron, et al.
Veröffentlicht: (2026) -
AnyView: Synthesizing Any Novel View in Dynamic Scenes
von: Van Hoorick, Basile, et al.
Veröffentlicht: (2026) -
SIRE: SE(3) Intrinsic Rigidity Embeddings
von: Smith, Cameron, et al.
Veröffentlicht: (2025) -
RoboDreamer: Learning Compositional World Models for Robot Imagination
von: Zhou, Siyuan, et al.
Veröffentlicht: (2024)