ROSGPT_Vision: Commanding Robots Using Only Language Models' Prompts
Fuente:
arXiv
Salvato in:
| Autori principali: | Benjdira, Bilel, Koubaa, Anis, Ali, Anas M. |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2023
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Streamlined Global and Local Features Combinator (SGLC) for High Resolution Image Dehazing
di: Benjdira, Bilel, et al.
Pubblicazione: (2023)
di: Benjdira, Bilel, et al.
Pubblicazione: (2023)
Dual-Stage Global and Local Feature Framework for Image Dehazing
di: Ali, Anas M., et al.
Pubblicazione: (2025)
di: Ali, Anas M., et al.
Pubblicazione: (2025)
License Plate Super-Resolution Using Diffusion Models
di: AlHalawani, Sawsan, et al.
Pubblicazione: (2023)
di: AlHalawani, Sawsan, et al.
Pubblicazione: (2023)
Reducing Latency in LLM-Based Natural Language Commands Processing for Robot Navigation
di: Pollini, Diego, et al.
Pubblicazione: (2025)
di: Pollini, Diego, et al.
Pubblicazione: (2025)
Vision-Language Models on the Edge for Real-Time Robotic Perception
di: Ahmad, Sarat, et al.
Pubblicazione: (2026)
di: Ahmad, Sarat, et al.
Pubblicazione: (2026)
Parking Analytics Framework using Deep Learning
di: Benjdira, Bilel, et al.
Pubblicazione: (2022)
di: Benjdira, Bilel, et al.
Pubblicazione: (2022)
CLARA: Classifying and Disambiguating User Commands for Reliable Interactive Robotic Agents
di: Park, Jeongeun, et al.
Pubblicazione: (2023)
di: Park, Jeongeun, et al.
Pubblicazione: (2023)
Agentic UAVs: LLM-Driven Autonomy with Integrated Tool-Calling and Cognitive Reasoning
di: Koubaa, Anis, et al.
Pubblicazione: (2025)
di: Koubaa, Anis, et al.
Pubblicazione: (2025)
Commonsense Reasoning for Legged Robot Adaptation with Vision-Language Models
di: Chen, Annie S., et al.
Pubblicazione: (2024)
di: Chen, Annie S., et al.
Pubblicazione: (2024)
Commanding Humanoid by Free-form Language: A Large Language Action Model with Unified Motion Vocabulary
di: Liu, Zhirui, et al.
Pubblicazione: (2025)
di: Liu, Zhirui, et al.
Pubblicazione: (2025)
Adversarial Attacks on Robotic Vision Language Action Models
di: Jones, Eliot Krzysztof, et al.
Pubblicazione: (2025)
di: Jones, Eliot Krzysztof, et al.
Pubblicazione: (2025)
Co-NavGPT: Multi-Robot Cooperative Visual Semantic Navigation Using Vision Language Models
di: Yu, Bangguo, et al.
Pubblicazione: (2023)
di: Yu, Bangguo, et al.
Pubblicazione: (2023)
3D-Grounded Vision-Language Framework for Robotic Task Planning: Automated Prompt Synthesis and Supervised Reasoning
di: Tang, Guoqin, et al.
Pubblicazione: (2025)
di: Tang, Guoqin, et al.
Pubblicazione: (2025)
Exploring the Adversarial Vulnerabilities of Vision-Language-Action Models in Robotics
di: Wang, Taowen, et al.
Pubblicazione: (2024)
di: Wang, Taowen, et al.
Pubblicazione: (2024)
Vision-Language-Policy Model for Dynamic Robot Task Planning
di: Wang, Jin, et al.
Pubblicazione: (2025)
di: Wang, Jin, et al.
Pubblicazione: (2025)
Emergence of Human to Robot Transfer in Vision-Language-Action Models
di: Kareer, Simar, et al.
Pubblicazione: (2025)
di: Kareer, Simar, et al.
Pubblicazione: (2025)
Maestro: Orchestrating Robotics Modules with Vision-Language Models for Zero-Shot Generalist Robots
di: Shi, Junyao, et al.
Pubblicazione: (2025)
di: Shi, Junyao, et al.
Pubblicazione: (2025)
RobotDesignGPT: Automated Robot Design Synthesis using Vision Language Models
di: Sontakke, Nitish, et al.
Pubblicazione: (2026)
di: Sontakke, Nitish, et al.
Pubblicazione: (2026)
DKPROMPT: Domain Knowledge Prompting Vision-Language Models for Open-World Planning
di: Zhang, Xiaohan, et al.
Pubblicazione: (2024)
di: Zhang, Xiaohan, et al.
Pubblicazione: (2024)
HiFi-CS: Towards Open Vocabulary Visual Grounding For Robotic Grasping Using Vision-Language Models
di: Bhat, Vineet, et al.
Pubblicazione: (2024)
di: Bhat, Vineet, et al.
Pubblicazione: (2024)
Experiences from Benchmarking Vision-Language-Action Models for Robotic Manipulation
di: Zhang, Yihao, et al.
Pubblicazione: (2025)
di: Zhang, Yihao, et al.
Pubblicazione: (2025)
CubeRobot: Grounding Language in Rubik's Cube Manipulation via Vision-Language Model
di: Wang, Feiyang, et al.
Pubblicazione: (2025)
di: Wang, Feiyang, et al.
Pubblicazione: (2025)
Bring My Cup! Personalizing Vision-Language-Action Models with Visual Attentive Prompting
di: Lee, Sangoh, et al.
Pubblicazione: (2025)
di: Lee, Sangoh, et al.
Pubblicazione: (2025)
Manual2Skill: Learning to Read Manuals and Acquire Robotic Skills for Furniture Assembly Using Vision-Language Models
di: Tie, Chenrui, et al.
Pubblicazione: (2025)
di: Tie, Chenrui, et al.
Pubblicazione: (2025)
Red-Teaming Vision-Language-Action Models via Quality Diversity Prompt Generation for Robust Robot Policies
di: Srikanth, Siddharth, et al.
Pubblicazione: (2026)
di: Srikanth, Siddharth, et al.
Pubblicazione: (2026)
Large Reward Models: Generalizable Online Robot Reward Generation with Vision-Language Models
di: Wu, Yanru, et al.
Pubblicazione: (2026)
di: Wu, Yanru, et al.
Pubblicazione: (2026)
ManipBench: Benchmarking Vision-Language Models for Low-Level Robot Manipulation
di: Zhao, Enyu, et al.
Pubblicazione: (2025)
di: Zhao, Enyu, et al.
Pubblicazione: (2025)
Planning with Vision-Language Models and a Use Case in Robot-Assisted Teaching
di: Dang, Xuzhe, et al.
Pubblicazione: (2025)
di: Dang, Xuzhe, et al.
Pubblicazione: (2025)
A Unified Framework for Real-Time Failure Handling in Robotics Using Vision-Language Models, Reactive Planner and Behavior Trees
di: Ahmad, Faseeh, et al.
Pubblicazione: (2025)
di: Ahmad, Faseeh, et al.
Pubblicazione: (2025)
VLMgineer: Vision Language Models as Robotic Toolsmiths
di: Gao, George Jiayuan, et al.
Pubblicazione: (2025)
di: Gao, George Jiayuan, et al.
Pubblicazione: (2025)
LADEV: A Language-Driven Testing and Evaluation Platform for Vision-Language-Action Models in Robotic Manipulation
di: Wang, Zhijie, et al.
Pubblicazione: (2024)
di: Wang, Zhijie, et al.
Pubblicazione: (2024)
LACY: A Vision-Language Model-based Language-Action Cycle for Self-Improving Robotic Manipulation
di: Hong, Youngjin, et al.
Pubblicazione: (2025)
di: Hong, Youngjin, et al.
Pubblicazione: (2025)
Characterizing Vision-Language-Action Models across XPUs: Constraints and Acceleration for On-Robot Deployment
di: Zhou, Kaijun, et al.
Pubblicazione: (2026)
di: Zhou, Kaijun, et al.
Pubblicazione: (2026)
GraspCorrect: Robotic Grasp Correction via Vision-Language Model-Guided Feedback
di: Lee, Sungjae, et al.
Pubblicazione: (2025)
di: Lee, Sungjae, et al.
Pubblicazione: (2025)
A Vision-Language-Action-Critic Model for Robotic Real-World Reinforcement Learning
di: Zhai, Shaopeng, et al.
Pubblicazione: (2025)
di: Zhai, Shaopeng, et al.
Pubblicazione: (2025)
Safety Aware Task Planning via Large Language Models in Robotics
di: Khan, Azal Ahmad, et al.
Pubblicazione: (2025)
di: Khan, Azal Ahmad, et al.
Pubblicazione: (2025)
Visual Prompting for Robotic Manipulation with Annotation-Guided Pick-and-Place Using ACT
di: Muttaqien, Muhammad A., et al.
Pubblicazione: (2025)
di: Muttaqien, Muhammad A., et al.
Pubblicazione: (2025)
V-VLAPS: Value-Guided Planning for Vision-Language-Action Models
di: Ren, Ke, et al.
Pubblicazione: (2026)
di: Ren, Ke, et al.
Pubblicazione: (2026)
INGRID: Intelligent Generative Robotic Design Using Large Language Models
di: Jia, Guanglu, et al.
Pubblicazione: (2025)
di: Jia, Guanglu, et al.
Pubblicazione: (2025)
Interactive Navigation in Environments with Traversable Obstacles Using Large Language and Vision-Language Models
di: Zhang, Zhen, et al.
Pubblicazione: (2023)
di: Zhang, Zhen, et al.
Pubblicazione: (2023)
Documenti analoghi
-
Streamlined Global and Local Features Combinator (SGLC) for High Resolution Image Dehazing
di: Benjdira, Bilel, et al.
Pubblicazione: (2023) -
Dual-Stage Global and Local Feature Framework for Image Dehazing
di: Ali, Anas M., et al.
Pubblicazione: (2025) -
License Plate Super-Resolution Using Diffusion Models
di: AlHalawani, Sawsan, et al.
Pubblicazione: (2023) -
Reducing Latency in LLM-Based Natural Language Commands Processing for Robot Navigation
di: Pollini, Diego, et al.
Pubblicazione: (2025) -
Vision-Language Models on the Edge for Real-Time Robotic Perception
di: Ahmad, Sarat, et al.
Pubblicazione: (2026)