CRAFT: Video Diffusion for Bimanual Robot Data Generation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Chen, Jason, Liu, I-Chun Arthur, Sukhatme, Gaurav, Seita, Daniel |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
ROPA: Synthetic Robot Pose Generation for RGB-D Bimanual Data Augmentation
von: Chen, Jason, et al.
Veröffentlicht: (2025)
von: Chen, Jason, et al.
Veröffentlicht: (2025)
D-CODA: Diffusion for Coordinated Dual-Arm Data Augmentation
von: Liu, I-Chun Arthur, et al.
Veröffentlicht: (2025)
von: Liu, I-Chun Arthur, et al.
Veröffentlicht: (2025)
VoxAct-B: Voxel-Based Acting and Stabilizing Policy for Bimanual Manipulation
von: Liu, I-Chun Arthur, et al.
Veröffentlicht: (2024)
von: Liu, I-Chun Arthur, et al.
Veröffentlicht: (2024)
Zero-Shot Generalization of Vision-Based RL Without Data Augmentation
von: Batra, Sumeet, et al.
Veröffentlicht: (2024)
von: Batra, Sumeet, et al.
Veröffentlicht: (2024)
RDT-1B: a Diffusion Foundation Model for Bimanual Manipulation
von: Liu, Songming, et al.
Veröffentlicht: (2024)
von: Liu, Songming, et al.
Veröffentlicht: (2024)
CLAMP: Contrastive Learning for 3D Multi-View Action-Conditioned Robotic Manipulation Pretraining
von: Liu, I-Chun Arthur, et al.
Veröffentlicht: (2026)
von: Liu, I-Chun Arthur, et al.
Veröffentlicht: (2026)
PerAct2: Benchmarking and Learning for Robotic Bimanual Manipulation Tasks
von: Grotz, Markus, et al.
Veröffentlicht: (2024)
von: Grotz, Markus, et al.
Veröffentlicht: (2024)
Bimanual Dexterity for Complex Tasks
von: Shaw, Kenneth, et al.
Veröffentlicht: (2024)
von: Shaw, Kenneth, et al.
Veröffentlicht: (2024)
DexMimicGen: Automated Data Generation for Bimanual Dexterous Manipulation via Imitation Learning
von: Jiang, Zhenyu, et al.
Veröffentlicht: (2024)
von: Jiang, Zhenyu, et al.
Veröffentlicht: (2024)
Zero-Shot Visual Generalization in Robot Manipulation
von: Batra, Sumeet, et al.
Veröffentlicht: (2025)
von: Batra, Sumeet, et al.
Veröffentlicht: (2025)
Geometry-aware 4D Video Generation for Robot Manipulation
von: Liu, Zeyi, et al.
Veröffentlicht: (2025)
von: Liu, Zeyi, et al.
Veröffentlicht: (2025)
The Ingredients for Robotic Diffusion Transformers
von: Dasari, Sudeep, et al.
Veröffentlicht: (2024)
von: Dasari, Sudeep, et al.
Veröffentlicht: (2024)
GenSim2: Scaling Robot Data Generation with Multi-modal and Reasoning LLMs
von: Hua, Pu, et al.
Veröffentlicht: (2024)
von: Hua, Pu, et al.
Veröffentlicht: (2024)
VLM See, Robot Do: Human Demo Video to Robot Action Plan via Vision Language Model
von: Wang, Beichen, et al.
Veröffentlicht: (2024)
von: Wang, Beichen, et al.
Veröffentlicht: (2024)
Vidar: Embodied Video Diffusion Model for Generalist Manipulation
von: Feng, Yao, et al.
Veröffentlicht: (2025)
von: Feng, Yao, et al.
Veröffentlicht: (2025)
On the Evaluation of Generative Robotic Simulations
von: Chen, Feng, et al.
Veröffentlicht: (2024)
von: Chen, Feng, et al.
Veröffentlicht: (2024)
RoboGen: Towards Unleashing Infinite Data for Automated Robot Learning via Generative Simulation
von: Wang, Yufei, et al.
Veröffentlicht: (2023)
von: Wang, Yufei, et al.
Veröffentlicht: (2023)
Turning Video Models into Generalist Robot Policies
von: Li, Sizhe Lester, et al.
Veröffentlicht: (2026)
von: Li, Sizhe Lester, et al.
Veröffentlicht: (2026)
Video Diffusion Alignment via Reward Gradients
von: Prabhudesai, Mihir, et al.
Veröffentlicht: (2024)
von: Prabhudesai, Mihir, et al.
Veröffentlicht: (2024)
Scalable Vision-Language-Action Model Pretraining for Robotic Manipulation with Real-Life Human Activity Videos
von: Li, Qixiu, et al.
Veröffentlicht: (2025)
von: Li, Qixiu, et al.
Veröffentlicht: (2025)
SceneFoundry: Generating Interactive Infinite 3D Worlds
von: Chen, ChunTeng, et al.
Veröffentlicht: (2026)
von: Chen, ChunTeng, et al.
Veröffentlicht: (2026)
PhysBench: Benchmarking and Enhancing Vision-Language Models for Physical World Understanding
von: Chow, Wei, et al.
Veröffentlicht: (2025)
von: Chow, Wei, et al.
Veröffentlicht: (2025)
DeeAD: Dynamic Early Exit of Vision-Language Action for Efficient Autonomous Driving
von: HU, Haibo, et al.
Veröffentlicht: (2025)
von: HU, Haibo, et al.
Veröffentlicht: (2025)
DreamDojo: A Generalist Robot World Model from Large-Scale Human Videos
von: Gao, Shenyuan, et al.
Veröffentlicht: (2026)
von: Gao, Shenyuan, et al.
Veröffentlicht: (2026)
Hierarchical Diffusion Policy for Kinematics-Aware Multi-Task Robotic Manipulation
von: Ma, Xiao, et al.
Veröffentlicht: (2024)
von: Ma, Xiao, et al.
Veröffentlicht: (2024)
Learning by Watching: A Review of Video-based Learning Approaches for Robot Manipulation
von: Eze, Chrisantus, et al.
Veröffentlicht: (2024)
von: Eze, Chrisantus, et al.
Veröffentlicht: (2024)
mimic-video: Video-Action Models for Generalizable Robot Control Beyond VLAs
von: Pai, Jonas, et al.
Veröffentlicht: (2025)
von: Pai, Jonas, et al.
Veröffentlicht: (2025)
OKAMI: Teaching Humanoid Robots Manipulation Skills through Single Video Imitation
von: Li, Jinhan, et al.
Veröffentlicht: (2024)
von: Li, Jinhan, et al.
Veröffentlicht: (2024)
Composing Pre-Trained Object-Centric Representations for Robotics From "What" and "Where" Foundation Models
von: Shi, Junyao, et al.
Veröffentlicht: (2024)
von: Shi, Junyao, et al.
Veröffentlicht: (2024)
Bimanual Grasp Synthesis for Dexterous Robot Hands
von: Shao, Yanming, et al.
Veröffentlicht: (2024)
von: Shao, Yanming, et al.
Veröffentlicht: (2024)
Cross-Modal Instructions for Robot Motion Generation
von: Barron, William, et al.
Veröffentlicht: (2025)
von: Barron, William, et al.
Veröffentlicht: (2025)
Unlocking Generalization for Robotics via Modularity and Scale
von: Dalal, Murtaza
Veröffentlicht: (2025)
von: Dalal, Murtaza
Veröffentlicht: (2025)
Diffusion Beats Autoregressive in Data-Constrained Settings
von: Prabhudesai, Mihir, et al.
Veröffentlicht: (2025)
von: Prabhudesai, Mihir, et al.
Veröffentlicht: (2025)
NIL: No-data Imitation Learning by Leveraging Pre-trained Video Diffusion Models
von: Albaba, Mert, et al.
Veröffentlicht: (2025)
von: Albaba, Mert, et al.
Veröffentlicht: (2025)
E0: Enhancing Generalization and Fine-Grained Control in VLA Models via Tweedie Discrete Diffusion
von: Zhan, Zhihao, et al.
Veröffentlicht: (2025)
von: Zhan, Zhihao, et al.
Veröffentlicht: (2025)
RoCoDA: Counterfactual Data Augmentation for Data-Efficient Robot Learning from Demonstrations
von: Ameperosa, Ezra, et al.
Veröffentlicht: (2024)
von: Ameperosa, Ezra, et al.
Veröffentlicht: (2024)
Information-driven Affordance Discovery for Efficient Robotic Manipulation
von: Mazzaglia, Pietro, et al.
Veröffentlicht: (2024)
von: Mazzaglia, Pietro, et al.
Veröffentlicht: (2024)
VidCRAFT3: Camera, Object, and Lighting Control for Image-to-Video Generation
von: Zheng, Sixiao, et al.
Veröffentlicht: (2025)
von: Zheng, Sixiao, et al.
Veröffentlicht: (2025)
MapDiffusion: Generative Diffusion for Vectorized Online HD Map Construction and Uncertainty Estimation in Autonomous Driving
von: Monninger, Thomas, et al.
Veröffentlicht: (2025)
von: Monninger, Thomas, et al.
Veröffentlicht: (2025)
MALLVI: A Multi-Agent Framework for Integrated Generalized Robotics Manipulation
von: Taji, Mehrshad, et al.
Veröffentlicht: (2026)
von: Taji, Mehrshad, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
ROPA: Synthetic Robot Pose Generation for RGB-D Bimanual Data Augmentation
von: Chen, Jason, et al.
Veröffentlicht: (2025) -
D-CODA: Diffusion for Coordinated Dual-Arm Data Augmentation
von: Liu, I-Chun Arthur, et al.
Veröffentlicht: (2025) -
VoxAct-B: Voxel-Based Acting and Stabilizing Policy for Bimanual Manipulation
von: Liu, I-Chun Arthur, et al.
Veröffentlicht: (2024) -
Zero-Shot Generalization of Vision-Based RL Without Data Augmentation
von: Batra, Sumeet, et al.
Veröffentlicht: (2024) -
RDT-1B: a Diffusion Foundation Model for Bimanual Manipulation
von: Liu, Songming, et al.
Veröffentlicht: (2024)