Cross-Modal Instructions for Robot Motion Generation
Fuente:
arXiv
Saved in:
| Main Authors: | Barron, William, Dong, Xiaoxiang, Johnson-Roberson, Matthew, Zhi, Weiming |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Joint Flow Trajectory Optimization For Feasible Robot Motion Generation from Video Demonstrations
by: Dong, Xiaoxiang, et al.
Published: (2025)
by: Dong, Xiaoxiang, et al.
Published: (2025)
SplaTraj: Camera Trajectory Generation with Semantic Gaussian Splatting
by: Liu, Xinyi, et al.
Published: (2024)
by: Liu, Xinyi, et al.
Published: (2024)
PhotoReg: Photometrically Registering 3D Gaussian Splatting Models
by: Yuan, Ziwen, et al.
Published: (2024)
by: Yuan, Ziwen, et al.
Published: (2024)
Infinite Leagues Under the Sea: Photorealistic 3D Underwater Terrain Generation by Latent Fractal Diffusion Models
by: Zhang, Tianyi, et al.
Published: (2025)
by: Zhang, Tianyi, et al.
Published: (2025)
Efficient Construction of Implicit Surface Models From a Single Image for Motion Generation
by: Chu, Wei-Teng, et al.
Published: (2025)
by: Chu, Wei-Teng, et al.
Published: (2025)
From Single Images to Motion Policies via Video-Generation Environment Representations
by: Zhi, Weiming, et al.
Published: (2025)
by: Zhi, Weiming, et al.
Published: (2025)
GraphSeg: Segmented 3D Representations via Graph Edge Addition and Contraction
by: Tang, Haozhan, et al.
Published: (2025)
by: Tang, Haozhan, et al.
Published: (2025)
Rewind-IL: Online Failure Detection and State Respawning for Imitation Learning
by: Zheng, Gehan, et al.
Published: (2026)
by: Zheng, Gehan, et al.
Published: (2026)
Bi-Manual Joint Camera Calibration and Scene Representation
by: Tang, Haozhan, et al.
Published: (2025)
by: Tang, Haozhan, et al.
Published: (2025)
DOSE3 : Diffusion-based Out-of-distribution detection on SE(3) trajectories
by: Cheng, Hongzhe, et al.
Published: (2025)
by: Cheng, Hongzhe, et al.
Published: (2025)
Toward General-Purpose Robots via Foundation Models: A Survey and Meta-Analysis
by: Hu, Yafei, et al.
Published: (2023)
by: Hu, Yafei, et al.
Published: (2023)
Towards Ambiguity-Free Spatial Foundation Model: Rethinking and Decoupling Depth Ambiguity
by: Xu, Xiaohao, et al.
Published: (2025)
by: Xu, Xiaohao, et al.
Published: (2025)
DarkGS: Learning Neural Illumination and 3D Gaussians Relighting for Robotic Exploration in the Dark
by: Zhang, Tianyi, et al.
Published: (2024)
by: Zhang, Tianyi, et al.
Published: (2024)
Preemptive Motion Planning for Human-to-Robot Indirect Placement Handovers
by: Choi, Andrew, et al.
Published: (2022)
by: Choi, Andrew, et al.
Published: (2022)
CoCMT: Communication-Efficient Cross-Modal Transformer for Collaborative Perception
by: Wang, Rujia, et al.
Published: (2025)
by: Wang, Rujia, et al.
Published: (2025)
Visual Sync: Multi-Camera Synchronization via Cross-View Object Motion
by: Liu, Shaowei, et al.
Published: (2025)
by: Liu, Shaowei, et al.
Published: (2025)
LiDAR-BIND-T: Improved and Temporally Consistent Sensor Modality Translation and Fusion for Robotic Applications
by: Balemans, Niels, et al.
Published: (2025)
by: Balemans, Niels, et al.
Published: (2025)
Natural Language Instructions for Scene-Responsive Human-in-the-Loop Motion Planning in Autonomous Driving using Vision-Language-Action Models
by: Martinez-Sanchez, Angel, et al.
Published: (2026)
by: Martinez-Sanchez, Angel, et al.
Published: (2026)
ContraMap: Contrastive Uncertainty Mapping for Robot Environment Representation
by: Le, Chi Cuong, et al.
Published: (2026)
by: Le, Chi Cuong, et al.
Published: (2026)
On the Evaluation of Generative Robotic Simulations
by: Chen, Feng, et al.
Published: (2024)
by: Chen, Feng, et al.
Published: (2024)
TAX-Pose: Task-Specific Cross-Pose Estimation for Robot Manipulation
by: Pan, Chuer, et al.
Published: (2022)
by: Pan, Chuer, et al.
Published: (2022)
Adaptive Prediction Ensemble: Improving Out-of-Distribution Generalization of Motion Forecasting
by: Li, Jinning, et al.
Published: (2024)
by: Li, Jinning, et al.
Published: (2024)
Unlocking Generalization for Robotics via Modularity and Scale
by: Dalal, Murtaza
Published: (2025)
by: Dalal, Murtaza
Published: (2025)
VLM See, Robot Do: Human Demo Video to Robot Action Plan via Vision Language Model
by: Wang, Beichen, et al.
Published: (2024)
by: Wang, Beichen, et al.
Published: (2024)
CRAFT: Video Diffusion for Bimanual Robot Data Generation
by: Chen, Jason, et al.
Published: (2026)
by: Chen, Jason, et al.
Published: (2026)
Geometry-aware 4D Video Generation for Robot Manipulation
by: Liu, Zeyi, et al.
Published: (2025)
by: Liu, Zeyi, et al.
Published: (2025)
Instruction-Guided Visual Masking
by: Zheng, Jinliang, et al.
Published: (2024)
by: Zheng, Jinliang, et al.
Published: (2024)
MALLVI: A Multi-Agent Framework for Integrated Generalized Robotics Manipulation
by: Taji, Mehrshad, et al.
Published: (2026)
by: Taji, Mehrshad, et al.
Published: (2026)
Interaction-Merged Motion Planning: Effectively Leveraging Diverse Motion Datasets for Robust Planning
by: Lee, Giwon, et al.
Published: (2025)
by: Lee, Giwon, et al.
Published: (2025)
M3Bench: Benchmarking Whole-body Motion Generation for Mobile Manipulation in 3D Scenes
by: Zhang, Zeyu, et al.
Published: (2024)
by: Zhang, Zeyu, et al.
Published: (2024)
Unifying Scene Representation and Hand-Eye Calibration with 3D Foundation Models
by: Zhi, Weiming, et al.
Published: (2024)
by: Zhi, Weiming, et al.
Published: (2024)
ROPA: Synthetic Robot Pose Generation for RGB-D Bimanual Data Augmentation
by: Chen, Jason, et al.
Published: (2025)
by: Chen, Jason, et al.
Published: (2025)
ConEQsA: Concurrent and Asynchronous Embodied Questions Scheduling and Answering
by: Wang, Haisheng, et al.
Published: (2025)
by: Wang, Haisheng, et al.
Published: (2025)
Moto: Latent Motion Token as the Bridging Language for Learning Robot Manipulation from Videos
by: Chen, Yi, et al.
Published: (2024)
by: Chen, Yi, et al.
Published: (2024)
P3-PO: Prescriptive Point Priors for Visuo-Spatial Generalization of Robot Policies
by: Levy, Mara, et al.
Published: (2024)
by: Levy, Mara, et al.
Published: (2024)
GenSim2: Scaling Robot Data Generation with Multi-modal and Reasoning LLMs
by: Hua, Pu, et al.
Published: (2024)
by: Hua, Pu, et al.
Published: (2024)
TidyBot++: An Open-Source Holonomic Mobile Manipulator for Robot Learning
by: Wu, Jimmy, et al.
Published: (2024)
by: Wu, Jimmy, et al.
Published: (2024)
Dynamic Cross-Modal Prompt Generation for Multimodal Continual Instruction Tuning
by: Hu, Tao, et al.
Published: (2026)
by: Hu, Tao, et al.
Published: (2026)
RoboGen: Towards Unleashing Infinite Data for Automated Robot Learning via Generative Simulation
by: Wang, Yufei, et al.
Published: (2023)
by: Wang, Yufei, et al.
Published: (2023)
Neural MP: A Generalist Neural Motion Planner
by: Dalal, Murtaza, et al.
Published: (2024)
by: Dalal, Murtaza, et al.
Published: (2024)
Similar Items
-
Joint Flow Trajectory Optimization For Feasible Robot Motion Generation from Video Demonstrations
by: Dong, Xiaoxiang, et al.
Published: (2025) -
SplaTraj: Camera Trajectory Generation with Semantic Gaussian Splatting
by: Liu, Xinyi, et al.
Published: (2024) -
PhotoReg: Photometrically Registering 3D Gaussian Splatting Models
by: Yuan, Ziwen, et al.
Published: (2024) -
Infinite Leagues Under the Sea: Photorealistic 3D Underwater Terrain Generation by Latent Fractal Diffusion Models
by: Zhang, Tianyi, et al.
Published: (2025) -
Efficient Construction of Implicit Surface Models From a Single Image for Motion Generation
by: Chu, Wei-Teng, et al.
Published: (2025)