Unlocking Generalization for Robotics via Modularity and Scale
Fuente:
arXiv
Saved in:
| Main Author: | Dalal, Murtaza |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Plan-Seq-Learn: Language Model Guided RL for Solving Long Horizon Robotics Tasks
by: Dalal, Murtaza, et al.
Published: (2024)
by: Dalal, Murtaza, et al.
Published: (2024)
Neural MP: A Generalist Neural Motion Planner
by: Dalal, Murtaza, et al.
Published: (2024)
by: Dalal, Murtaza, et al.
Published: (2024)
Scaling Robot Policy Learning via Zero-Shot Labeling with Foundation Models
by: Blank, Nils, et al.
Published: (2024)
by: Blank, Nils, et al.
Published: (2024)
GenSim2: Scaling Robot Data Generation with Multi-modal and Reasoning LLMs
by: Hua, Pu, et al.
Published: (2024)
by: Hua, Pu, et al.
Published: (2024)
Toward General-Purpose Robots via Foundation Models: A Survey and Meta-Analysis
by: Hu, Yafei, et al.
Published: (2023)
by: Hu, Yafei, et al.
Published: (2023)
On the Evaluation of Generative Robotic Simulations
by: Chen, Feng, et al.
Published: (2024)
by: Chen, Feng, et al.
Published: (2024)
RoboGen: Towards Unleashing Infinite Data for Automated Robot Learning via Generative Simulation
by: Wang, Yufei, et al.
Published: (2023)
by: Wang, Yufei, et al.
Published: (2023)
RobotArena $\infty$: Scalable Robot Benchmarking via Real-to-Sim Translation
by: Jangir, Yash, et al.
Published: (2025)
by: Jangir, Yash, et al.
Published: (2025)
Cross-Modal Instructions for Robot Motion Generation
by: Barron, William, et al.
Published: (2025)
by: Barron, William, et al.
Published: (2025)
DiffGen: Robot Demonstration Generation via Differentiable Physics Simulation, Differentiable Rendering, and Vision-Language Model
by: Jin, Yang, et al.
Published: (2024)
by: Jin, Yang, et al.
Published: (2024)
CRAFT: Video Diffusion for Bimanual Robot Data Generation
by: Chen, Jason, et al.
Published: (2026)
by: Chen, Jason, et al.
Published: (2026)
DreamDojo: A Generalist Robot World Model from Large-Scale Human Videos
by: Gao, Shenyuan, et al.
Published: (2026)
by: Gao, Shenyuan, et al.
Published: (2026)
Geometry-aware 4D Video Generation for Robot Manipulation
by: Liu, Zeyi, et al.
Published: (2025)
by: Liu, Zeyi, et al.
Published: (2025)
VLM See, Robot Do: Human Demo Video to Robot Action Plan via Vision Language Model
by: Wang, Beichen, et al.
Published: (2024)
by: Wang, Beichen, et al.
Published: (2024)
MALLVI: A Multi-Agent Framework for Integrated Generalized Robotics Manipulation
by: Taji, Mehrshad, et al.
Published: (2026)
by: Taji, Mehrshad, et al.
Published: (2026)
ROPA: Synthetic Robot Pose Generation for RGB-D Bimanual Data Augmentation
by: Chen, Jason, et al.
Published: (2025)
by: Chen, Jason, et al.
Published: (2025)
P3-PO: Prescriptive Point Priors for Visuo-Spatial Generalization of Robot Policies
by: Levy, Mara, et al.
Published: (2024)
by: Levy, Mara, et al.
Published: (2024)
RoboEXP: Action-Conditioned Scene Graph via Interactive Exploration for Robotic Manipulation
by: Jiang, Hanxiao, et al.
Published: (2024)
by: Jiang, Hanxiao, et al.
Published: (2024)
UAV-VLA: Vision-Language-Action System for Large Scale Aerial Mission Generation
by: Sautenkov, Oleg, et al.
Published: (2025)
by: Sautenkov, Oleg, et al.
Published: (2025)
OK-Robot: What Really Matters in Integrating Open-Knowledge Models for Robotics
by: Liu, Peiqi, et al.
Published: (2024)
by: Liu, Peiqi, et al.
Published: (2024)
RDT2: Exploring the Scaling Limit of UMI Data Towards Zero-Shot Cross-Embodiment Generalization
by: Liu, Songming, et al.
Published: (2026)
by: Liu, Songming, et al.
Published: (2026)
The Ingredients for Robotic Diffusion Transformers
by: Dasari, Sudeep, et al.
Published: (2024)
by: Dasari, Sudeep, et al.
Published: (2024)
AutoRT: Embodied Foundation Models for Large Scale Orchestration of Robotic Agents
by: Ahn, Michael, et al.
Published: (2024)
by: Ahn, Michael, et al.
Published: (2024)
Neural Fields in Robotics: A Survey
by: Irshad, Muhammad Zubair, et al.
Published: (2024)
by: Irshad, Muhammad Zubair, et al.
Published: (2024)
Bifurcation Identification for Ultrasound-driven Robotic Cannulation
by: Morales, Cecilia G., et al.
Published: (2024)
by: Morales, Cecilia G., et al.
Published: (2024)
Redundancy-aware Action Spaces for Robot Learning
by: Mazzaglia, Pietro, et al.
Published: (2024)
by: Mazzaglia, Pietro, et al.
Published: (2024)
Rapid Motor Adaptation for Robotic Manipulator Arms
by: Liang, Yichao, et al.
Published: (2023)
by: Liang, Yichao, et al.
Published: (2023)
Turning Video Models into Generalist Robot Policies
by: Li, Sizhe Lester, et al.
Published: (2026)
by: Li, Sizhe Lester, et al.
Published: (2026)
Semantically Controllable Augmentations for Generalizable Robot Learning
by: Chen, Zoey, et al.
Published: (2024)
by: Chen, Zoey, et al.
Published: (2024)
IFG: Internet-Scale Guidance for Functional Grasping Generation
by: Liu, Ray Muxin, et al.
Published: (2025)
by: Liu, Ray Muxin, et al.
Published: (2025)
A Synthetic Dataset for Manometry Recognition in Robotic Applications
by: Saraiva, Pedro Antonio Rabelo, et al.
Published: (2025)
by: Saraiva, Pedro Antonio Rabelo, et al.
Published: (2025)
EmbodiSwap for Zero-Shot Robot Imitation Learning
by: Dessalene, Eadom, et al.
Published: (2025)
by: Dessalene, Eadom, et al.
Published: (2025)
Learned Visual Navigation for Under-Canopy Agricultural Robots
by: Sivakumar, Arun Narenthiran, et al.
Published: (2021)
by: Sivakumar, Arun Narenthiran, et al.
Published: (2021)
Surfer: Progressive Reasoning with World Models for Robotic Manipulation
by: Ren, Pengzhen, et al.
Published: (2023)
by: Ren, Pengzhen, et al.
Published: (2023)
Leveraging Locality to Boost Sample Efficiency in Robotic Manipulation
by: Zhang, Tong, et al.
Published: (2024)
by: Zhang, Tong, et al.
Published: (2024)
Knolling Bot: Teaching Robots the Human Notion of Tidiness
by: Hu, Yuhang, et al.
Published: (2023)
by: Hu, Yuhang, et al.
Published: (2023)
Information-driven Affordance Discovery for Efficient Robotic Manipulation
by: Mazzaglia, Pietro, et al.
Published: (2024)
by: Mazzaglia, Pietro, et al.
Published: (2024)
SegDAC: Visual Generalization in Reinforcement Learning via Dynamic Object Tokens
by: Brown, Alexandre, et al.
Published: (2025)
by: Brown, Alexandre, et al.
Published: (2025)
CHOrD: Generation of Collision-Free, House-Scale, and Organized Digital Twins for 3D Indoor Scenes with Controllable Floor Plans and Optimal Layouts
by: Su, Chong, et al.
Published: (2025)
by: Su, Chong, et al.
Published: (2025)
GWM: Towards Scalable Gaussian World Models for Robotic Manipulation
by: Lu, Guanxing, et al.
Published: (2025)
by: Lu, Guanxing, et al.
Published: (2025)
Similar Items
-
Plan-Seq-Learn: Language Model Guided RL for Solving Long Horizon Robotics Tasks
by: Dalal, Murtaza, et al.
Published: (2024) -
Neural MP: A Generalist Neural Motion Planner
by: Dalal, Murtaza, et al.
Published: (2024) -
Scaling Robot Policy Learning via Zero-Shot Labeling with Foundation Models
by: Blank, Nils, et al.
Published: (2024) -
GenSim2: Scaling Robot Data Generation with Multi-modal and Reasoning LLMs
by: Hua, Pu, et al.
Published: (2024) -
Toward General-Purpose Robots via Foundation Models: A Survey and Meta-Analysis
by: Hu, Yafei, et al.
Published: (2023)