Robotic Scene Cloning:Advancing Zero-Shot Robotic Scene Adaptation in Manipulation via Visual Prompt Editing
Fuente:
arXiv
Saved in:
| Main Authors: | Huang, Binyuan, Wen, Yuqing, Zhao, Yucheng, Hu, Yaosi, Wang, Tiancai, Chen, Chang Wen, Fan, Haoqiang, Chen, Zhenzhong |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
SubjectDrive: Scaling Generative Data in Autonomous Driving via Subject Control
by: Huang, Binyuan, et al.
Published: (2024)
by: Huang, Binyuan, et al.
Published: (2024)
ROSA: Harnessing Robot States for Vision-Language and Action Alignment
by: Wen, Yuqing, et al.
Published: (2025)
by: Wen, Yuqing, et al.
Published: (2025)
ManiAgent: An Agentic Framework for General Robotic Manipulation
by: Yang, Yi, et al.
Published: (2025)
by: Yang, Yi, et al.
Published: (2025)
IntentionVLA: Generalizable and Efficient Embodied Intention Reasoning for Human-Robot Interaction
by: Chen, Yandu, et al.
Published: (2025)
by: Chen, Yandu, et al.
Published: (2025)
Behavior Cloning of MPC for 3-DOF Robotic Manipulators
by: Guegan, Theo, et al.
Published: (2026)
by: Guegan, Theo, et al.
Published: (2026)
Zero-Shot Visual Generalization in Robot Manipulation
by: Batra, Sumeet, et al.
Published: (2025)
by: Batra, Sumeet, et al.
Published: (2025)
SpatialActor: Exploring Disentangled Spatial Representations for Robust Robotic Manipulation
by: Shi, Hao, et al.
Published: (2025)
by: Shi, Hao, et al.
Published: (2025)
Integration of Robot and Scene Kinematics for Sequential Mobile Manipulation Planning
by: Jiao, Ziyuan, et al.
Published: (2025)
by: Jiao, Ziyuan, et al.
Published: (2025)
SpatialAnt: Autonomous Zero-Shot Robot Navigation via Active Scene Reconstruction and Visual Anticipation
by: Zhang, Jiwen, et al.
Published: (2026)
by: Zhang, Jiwen, et al.
Published: (2026)
SegGrasp: Zero-Shot Task-Oriented Grasping via Semantic and Geometric Guided Segmentation
by: Li, Haosheng, et al.
Published: (2024)
by: Li, Haosheng, et al.
Published: (2024)
Ross3D: Reconstructive Visual Instruction Tuning with 3D-Awareness
by: Wang, Haochen, et al.
Published: (2025)
by: Wang, Haochen, et al.
Published: (2025)
MemoryVLA: Perceptual-Cognitive Memory in Vision-Language-Action Models for Robotic Manipulation
by: Shi, Hao, et al.
Published: (2025)
by: Shi, Hao, et al.
Published: (2025)
Veo-Act: How Far Can Frontier Video Models Advance Generalizable Robot Manipulation?
by: Zhang, Zhongru, et al.
Published: (2026)
by: Zhang, Zhongru, et al.
Published: (2026)
AeroScene: Progressive Scene Synthesis for Aerial Robotics
by: Vu, Nghia, et al.
Published: (2026)
by: Vu, Nghia, et al.
Published: (2026)
SceneComplete: Open-World 3D Scene Completion in Cluttered Real World Environments for Robot Manipulation
by: Agarwal, Aditya, et al.
Published: (2024)
by: Agarwal, Aditya, et al.
Published: (2024)
Improving Robotic Manipulation Robustness via NICE Scene Surgery
by: Pakdamansavoji, Sajjad, et al.
Published: (2025)
by: Pakdamansavoji, Sajjad, et al.
Published: (2025)
Semantic Scene Segmentation for Robotics
by: Hurtado, Juana Valeria, et al.
Published: (2024)
by: Hurtado, Juana Valeria, et al.
Published: (2024)
RoboAug: One Annotation to Hundreds of Scenes via Region-Contrastive Data Augmentation for Robotic Manipulation
by: Wang, Xinhua, et al.
Published: (2026)
by: Wang, Xinhua, et al.
Published: (2026)
AgentWorld: An Interactive Simulation Platform for Scene Construction and Mobile Robotic Manipulation
by: Zhang, Yizheng, et al.
Published: (2025)
by: Zhang, Yizheng, et al.
Published: (2025)
Semantically Safe Robot Manipulation: From Semantic Scene Understanding to Motion Safeguards
by: Brunke, Lukas, et al.
Published: (2024)
by: Brunke, Lukas, et al.
Published: (2024)
ClutterGen: A Cluttered Scene Generator for Robot Learning
by: Jia, Yinsen, et al.
Published: (2024)
by: Jia, Yinsen, et al.
Published: (2024)
RoboOcc: Enhancing the Geometric and Semantic Scene Understanding for Robots
by: Zhang, Zhang, et al.
Published: (2025)
by: Zhang, Zhang, et al.
Published: (2025)
Never-Ending Behavior-Cloning Agent for Robotic Manipulation
by: Liang, Wenqi, et al.
Published: (2024)
by: Liang, Wenqi, et al.
Published: (2024)
Rethinking Latent Redundancy in Behavior Cloning: An Information Bottleneck Approach for Robot Manipulation
by: Bai, Shuanghao, et al.
Published: (2025)
by: Bai, Shuanghao, et al.
Published: (2025)
MSGField: A Unified Scene Representation Integrating Motion, Semantics, and Geometry for Robotic Manipulation
by: Sheng, Yu, et al.
Published: (2024)
by: Sheng, Yu, et al.
Published: (2024)
Egocentric Visual Self-Modeling for Autonomous Robot Dynamics Prediction and Adaptation
by: Hu, Yuhang, et al.
Published: (2022)
by: Hu, Yuhang, et al.
Published: (2022)
Efficient Hybrid SE(3)-Equivariant Visuomotor Flow Policy via Spherical Harmonics for Robot Manipulation
by: Zhang, Qinglun, et al.
Published: (2026)
by: Zhang, Qinglun, et al.
Published: (2026)
SceneGraphGrounder: Zero-Shot 3D Visual Grounding via Structured Scene Graph Matching
by: Sun, Xuefei, et al.
Published: (2026)
by: Sun, Xuefei, et al.
Published: (2026)
SeedPolicy: Horizon Scaling via Self-Evolving Diffusion Policy for Robot Manipulation
by: Gui, Youqiang, et al.
Published: (2026)
by: Gui, Youqiang, et al.
Published: (2026)
LLaDA-VLA: Vision Language Diffusion Action Models
by: Wen, Yuqing, et al.
Published: (2025)
by: Wen, Yuqing, et al.
Published: (2025)
From the Laboratory to Real-World Application: Evaluating Zero-Shot Scene Interpretation on Edge Devices for Mobile Robotics
by: Schuler, Nicolas, et al.
Published: (2025)
by: Schuler, Nicolas, et al.
Published: (2025)
Running VLAs at Real-time Speed
by: Ma, Yunchao, et al.
Published: (2025)
by: Ma, Yunchao, et al.
Published: (2025)
UniManip: General-Purpose Zero-Shot Robotic Manipulation with Agentic Operational Graph
by: Liu, Haichao, et al.
Published: (2026)
by: Liu, Haichao, et al.
Published: (2026)
Maestro: Orchestrating Robotics Modules with Vision-Language Models for Zero-Shot Generalist Robots
by: Shi, Junyao, et al.
Published: (2025)
by: Shi, Junyao, et al.
Published: (2025)
Glad: A Streaming Scene Generator for Autonomous Driving
by: Xie, Bin, et al.
Published: (2025)
by: Xie, Bin, et al.
Published: (2025)
MALMM: Multi-Agent Large Language Models for Zero-Shot Robotics Manipulation
by: Singh, Harsh, et al.
Published: (2024)
by: Singh, Harsh, et al.
Published: (2024)
ZeroSCD: Zero-Shot Street Scene Change Detection
by: Kannan, Shyam Sundar, et al.
Published: (2024)
by: Kannan, Shyam Sundar, et al.
Published: (2024)
FlowPolicy: Enabling Fast and Robust 3D Flow-based Policy via Consistency Flow Matching for Robot Manipulation
by: Zhang, Qinglun, et al.
Published: (2024)
by: Zhang, Qinglun, et al.
Published: (2024)
Green Screen Augmentation Enables Scene Generalisation in Robotic Manipulation
by: Teoh, Eugene, et al.
Published: (2024)
by: Teoh, Eugene, et al.
Published: (2024)
Data Scaling Laws in Imitation Learning for Robotic Manipulation
by: Hu, Yingdong, et al.
Published: (2024)
by: Hu, Yingdong, et al.
Published: (2024)
Similar Items
-
SubjectDrive: Scaling Generative Data in Autonomous Driving via Subject Control
by: Huang, Binyuan, et al.
Published: (2024) -
ROSA: Harnessing Robot States for Vision-Language and Action Alignment
by: Wen, Yuqing, et al.
Published: (2025) -
ManiAgent: An Agentic Framework for General Robotic Manipulation
by: Yang, Yi, et al.
Published: (2025) -
IntentionVLA: Generalizable and Efficient Embodied Intention Reasoning for Human-Robot Interaction
by: Chen, Yandu, et al.
Published: (2025) -
Behavior Cloning of MPC for 3-DOF Robotic Manipulators
by: Guegan, Theo, et al.
Published: (2026)