RobotDesignGPT: Automated Robot Design Synthesis using Vision Language Models
Fuente:
arXiv
Saved in:
| Main Authors: | Sontakke, Nitish, Kumar, K. Niranjan, Ha, Sehoon |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
A Design Co-Pilot for Task-Tailored Manipulators
by: Külz, Jonathan, et al.
Published: (2025)
by: Külz, Jonathan, et al.
Published: (2025)
LineRides: Line-Guided Reinforcement Learning for Bicycle Robot Stunts
by: Rho, Seungeun, et al.
Published: (2026)
by: Rho, Seungeun, et al.
Published: (2026)
Transforming a Quadruped into a Guide Robot for the Visually Impaired: Formalizing Wayfinding, Interaction Modeling, and Safety Mechanism
by: Kim, J. Taery, et al.
Published: (2023)
by: Kim, J. Taery, et al.
Published: (2023)
Dynamics-Guided Diffusion Model for Sensor-less Robot Manipulator Design
by: Xu, Xiaomeng, et al.
Published: (2024)
by: Xu, Xiaomeng, et al.
Published: (2024)
Co-NavGPT: Multi-Robot Cooperative Visual Semantic Navigation Using Vision Language Models
by: Yu, Bangguo, et al.
Published: (2023)
by: Yu, Bangguo, et al.
Published: (2023)
3D-Grounded Vision-Language Framework for Robotic Task Planning: Automated Prompt Synthesis and Supervised Reasoning
by: Tang, Guoqin, et al.
Published: (2025)
by: Tang, Guoqin, et al.
Published: (2025)
INGRID: Intelligent Generative Robotic Design Using Large Language Models
by: Jia, Guanglu, et al.
Published: (2025)
by: Jia, Guanglu, et al.
Published: (2025)
Maestro: Orchestrating Robotics Modules with Vision-Language Models for Zero-Shot Generalist Robots
by: Shi, Junyao, et al.
Published: (2025)
by: Shi, Junyao, et al.
Published: (2025)
Creation of Novel Soft Robot Designs using Generative AI
by: Chan, Wee Kiat, et al.
Published: (2024)
by: Chan, Wee Kiat, et al.
Published: (2024)
Adversarial Attacks on Robotic Vision Language Action Models
by: Jones, Eliot Krzysztof, et al.
Published: (2025)
by: Jones, Eliot Krzysztof, et al.
Published: (2025)
EvoNav: Evolutionary Reward Function Design for Robot Navigation with Large Language Models
by: Zhao, Zhikai, et al.
Published: (2026)
by: Zhao, Zhikai, et al.
Published: (2026)
Vision-Language Models on the Edge for Real-Time Robotic Perception
by: Ahmad, Sarat, et al.
Published: (2026)
by: Ahmad, Sarat, et al.
Published: (2026)
Exploring the Adversarial Vulnerabilities of Vision-Language-Action Models in Robotics
by: Wang, Taowen, et al.
Published: (2024)
by: Wang, Taowen, et al.
Published: (2024)
Vision-Language-Policy Model for Dynamic Robot Task Planning
by: Wang, Jin, et al.
Published: (2025)
by: Wang, Jin, et al.
Published: (2025)
Emergence of Human to Robot Transfer in Vision-Language-Action Models
by: Kareer, Simar, et al.
Published: (2025)
by: Kareer, Simar, et al.
Published: (2025)
Commonsense Reasoning for Legged Robot Adaptation with Vision-Language Models
by: Chen, Annie S., et al.
Published: (2024)
by: Chen, Annie S., et al.
Published: (2024)
Experiences from Benchmarking Vision-Language-Action Models for Robotic Manipulation
by: Zhang, Yihao, et al.
Published: (2025)
by: Zhang, Yihao, et al.
Published: (2025)
ROSGPT_Vision: Commanding Robots Using Only Language Models' Prompts
by: Benjdira, Bilel, et al.
Published: (2023)
by: Benjdira, Bilel, et al.
Published: (2023)
VLMgineer: Vision Language Models as Robotic Toolsmiths
by: Gao, George Jiayuan, et al.
Published: (2025)
by: Gao, George Jiayuan, et al.
Published: (2025)
CubeRobot: Grounding Language in Rubik's Cube Manipulation via Vision-Language Model
by: Wang, Feiyang, et al.
Published: (2025)
by: Wang, Feiyang, et al.
Published: (2025)
Design Optimizer for Planar Soft-Growing Robot Manipulators
by: Stroppa, Fabio
Published: (2023)
by: Stroppa, Fabio
Published: (2023)
Large Reward Models: Generalizable Online Robot Reward Generation with Vision-Language Models
by: Wu, Yanru, et al.
Published: (2026)
by: Wu, Yanru, et al.
Published: (2026)
Neural Operators for Design-Space Surrogate Modeling of Tendon-Actuated Continuum Robots
by: Frieden, Branden, et al.
Published: (2026)
by: Frieden, Branden, et al.
Published: (2026)
Automated Hybrid Reward Scheduling via Large Language Models for Robotic Skill Learning
by: Huang, Changxin, et al.
Published: (2025)
by: Huang, Changxin, et al.
Published: (2025)
ManipBench: Benchmarking Vision-Language Models for Low-Level Robot Manipulation
by: Zhao, Enyu, et al.
Published: (2025)
by: Zhao, Enyu, et al.
Published: (2025)
Planning with Vision-Language Models and a Use Case in Robot-Assisted Teaching
by: Dang, Xuzhe, et al.
Published: (2025)
by: Dang, Xuzhe, et al.
Published: (2025)
Deep Reinforcement Learning for Multi-Agent Coordination
by: Aina, Kehinde O., et al.
Published: (2025)
by: Aina, Kehinde O., et al.
Published: (2025)
Vision-Language Foundation Models as Effective Robot Imitators
by: Li, Xinghang, et al.
Published: (2023)
by: Li, Xinghang, et al.
Published: (2023)
Optimizing Structured Data Processing through Robotic Process Automation
by: Bhardwaj, Vivek, et al.
Published: (2024)
by: Bhardwaj, Vivek, et al.
Published: (2024)
Embodied Human Simulation for Quantitative Design and Analysis of Interactive Robotics
by: Zuo, Chenhui, et al.
Published: (2026)
by: Zuo, Chenhui, et al.
Published: (2026)
COMRES-VLM: Coordinated Multi-Robot Exploration and Search using Vision Language Models
by: Wang, Ruiyang, et al.
Published: (2025)
by: Wang, Ruiyang, et al.
Published: (2025)
Characterizing Vision-Language-Action Models across XPUs: Constraints and Acceleration for On-Robot Deployment
by: Zhou, Kaijun, et al.
Published: (2026)
by: Zhou, Kaijun, et al.
Published: (2026)
GraspCorrect: Robotic Grasp Correction via Vision-Language Model-Guided Feedback
by: Lee, Sungjae, et al.
Published: (2025)
by: Lee, Sungjae, et al.
Published: (2025)
A Vision-Language-Action-Critic Model for Robotic Real-World Reinforcement Learning
by: Zhai, Shaopeng, et al.
Published: (2025)
by: Zhai, Shaopeng, et al.
Published: (2025)
Text2Robot: Evolutionary Robot Design from Text Descriptions
by: Ringel, Ryan P., et al.
Published: (2024)
by: Ringel, Ryan P., et al.
Published: (2024)
HM3D-OVON: A Dataset and Benchmark for Open-Vocabulary Object Goal Navigation
by: Yokoyama, Naoki, et al.
Published: (2024)
by: Yokoyama, Naoki, et al.
Published: (2024)
LADEV: A Language-Driven Testing and Evaluation Platform for Vision-Language-Action Models in Robotic Manipulation
by: Wang, Zhijie, et al.
Published: (2024)
by: Wang, Zhijie, et al.
Published: (2024)
LACY: A Vision-Language Model-based Language-Action Cycle for Self-Improving Robotic Manipulation
by: Hong, Youngjin, et al.
Published: (2025)
by: Hong, Youngjin, et al.
Published: (2025)
ManiFoundation Model for General-Purpose Robotic Manipulation of Contact Synthesis with Arbitrary Objects and Robots
by: Xu, Zhixuan, et al.
Published: (2024)
by: Xu, Zhixuan, et al.
Published: (2024)
Shape Your Body: Value Gradients for Multi-Embodiment Robot Design
by: Bohlinger, Nico, et al.
Published: (2026)
by: Bohlinger, Nico, et al.
Published: (2026)
Similar Items
-
A Design Co-Pilot for Task-Tailored Manipulators
by: Külz, Jonathan, et al.
Published: (2025) -
LineRides: Line-Guided Reinforcement Learning for Bicycle Robot Stunts
by: Rho, Seungeun, et al.
Published: (2026) -
Transforming a Quadruped into a Guide Robot for the Visually Impaired: Formalizing Wayfinding, Interaction Modeling, and Safety Mechanism
by: Kim, J. Taery, et al.
Published: (2023) -
Dynamics-Guided Diffusion Model for Sensor-less Robot Manipulator Design
by: Xu, Xiaomeng, et al.
Published: (2024) -
Co-NavGPT: Multi-Robot Cooperative Visual Semantic Navigation Using Vision Language Models
by: Yu, Bangguo, et al.
Published: (2023)