Leveraging Foundation Models for Enhancing Robot Perception and Action
Fuente:
arXiv
Saved in:
| Main Author: | Mirjalili, Reihaneh |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Logically Constrained Robotics Transformers for Enhanced Perception-Action Planning
by: Kapoor, Parv, et al.
Published: (2024)
by: Kapoor, Parv, et al.
Published: (2024)
AttenA+: Rectifying Action Inequality in Robotic Foundation Models
by: Peng, Daojie, et al.
Published: (2026)
by: Peng, Daojie, et al.
Published: (2026)
Leveraging Large Language Models for Enhancing Autonomous Vehicle Perception
by: Karagounis, Athanasios
Published: (2024)
by: Karagounis, Athanasios
Published: (2024)
Understanding Physical Properties of Unseen Deformable Objects by Leveraging Large Language Models and Robot Actions
by: Park, Changmin, et al.
Published: (2025)
by: Park, Changmin, et al.
Published: (2025)
A Scalable Multi-Robot Framework for Decentralized and Asynchronous Perception-Action-Communication Loops
by: Agarwal, Saurav, et al.
Published: (2023)
by: Agarwal, Saurav, et al.
Published: (2023)
Universal Actions for Enhanced Embodied Foundation Models
by: Zheng, Jinliang, et al.
Published: (2025)
by: Zheng, Jinliang, et al.
Published: (2025)
Transferring Foundation Models for Generalizable Robotic Manipulation
by: Yang, Jiange, et al.
Published: (2023)
by: Yang, Jiange, et al.
Published: (2023)
Adversarial Attacks on Robotic Vision Language Action Models
by: Jones, Eliot Krzysztof, et al.
Published: (2025)
by: Jones, Eliot Krzysztof, et al.
Published: (2025)
Verifiably Following Complex Robot Instructions with Foundation Models
by: Quartey, Benedict, et al.
Published: (2024)
by: Quartey, Benedict, et al.
Published: (2024)
Specification-Aware Distribution Shaping for Robotics Foundation Models
by: Yüksel, Sadık Bera, et al.
Published: (2026)
by: Yüksel, Sadık Bera, et al.
Published: (2026)
Humanoid World Models: Open World Foundation Models for Humanoid Robotics
by: Ali, Muhammad Qasim, et al.
Published: (2025)
by: Ali, Muhammad Qasim, et al.
Published: (2025)
Toward Accurate Long-Horizon Robotic Manipulation: Language-to-Action with Foundation Models via Scene Graphs
by: Dinesh, Sushil Samuel, et al.
Published: (2025)
by: Dinesh, Sushil Samuel, et al.
Published: (2025)
Emergence of Human to Robot Transfer in Vision-Language-Action Models
by: Kareer, Simar, et al.
Published: (2025)
by: Kareer, Simar, et al.
Published: (2025)
Exploring the Adversarial Vulnerabilities of Vision-Language-Action Models in Robotics
by: Wang, Taowen, et al.
Published: (2024)
by: Wang, Taowen, et al.
Published: (2024)
Vision-Language Models on the Edge for Real-Time Robotic Perception
by: Ahmad, Sarat, et al.
Published: (2026)
by: Ahmad, Sarat, et al.
Published: (2026)
Robots Can Multitask Too: Integrating a Memory Architecture and LLMs for Enhanced Cross-Task Robot Action Generation
by: Ali, Hassan, et al.
Published: (2024)
by: Ali, Hassan, et al.
Published: (2024)
Are Transformers Truly Foundational for Robotics?
by: Marshall, James A. R., et al.
Published: (2024)
by: Marshall, James A. R., et al.
Published: (2024)
Natural Selection via Foundation Models for Soft Robot Evolution
by: Chen, Changhe, et al.
Published: (2025)
by: Chen, Changhe, et al.
Published: (2025)
Robotics-Inspired Guardrails for Foundation Models in Socially Sensitive Domains
by: Ramnauth, Rebecca, et al.
Published: (2026)
by: Ramnauth, Rebecca, et al.
Published: (2026)
A Survey on Robotics with Foundation Models: toward Embodied AI
by: Xu, Zhiyuan, et al.
Published: (2024)
by: Xu, Zhiyuan, et al.
Published: (2024)
ManiFoundation Model for General-Purpose Robotic Manipulation of Contact Synthesis with Arbitrary Objects and Robots
by: Xu, Zhixuan, et al.
Published: (2024)
by: Xu, Zhixuan, et al.
Published: (2024)
Experiences from Benchmarking Vision-Language-Action Models for Robotic Manipulation
by: Zhang, Yihao, et al.
Published: (2025)
by: Zhang, Yihao, et al.
Published: (2025)
CoPAL: Corrective Planning of Robot Actions with Large Language Models
by: Joublin, Frank, et al.
Published: (2023)
by: Joublin, Frank, et al.
Published: (2023)
Temporal Binding Foundation Model for Material Property Recognition via Tactile Sequence Perception
by: You, Hengxu, et al.
Published: (2025)
by: You, Hengxu, et al.
Published: (2025)
VGGT-DP: Generalizable Robot Control via Vision Foundation Models
by: Ge, Shijia, et al.
Published: (2025)
by: Ge, Shijia, et al.
Published: (2025)
Robot-Powered Data Flywheels: Deploying Robots in the Wild for Continual Data Collection and Foundation Model Adaptation
by: Grannen, Jennifer, et al.
Published: (2025)
by: Grannen, Jennifer, et al.
Published: (2025)
Instruct2Act: From Human Instruction to Actions Sequencing and Execution via Robot Action Network for Robotic Manipulation
by: Sharma, Archit, et al.
Published: (2026)
by: Sharma, Archit, et al.
Published: (2026)
Online Foundation Model Selection in Robotics
by: Li, Po-han, et al.
Published: (2024)
by: Li, Po-han, et al.
Published: (2024)
ChronoDreamer: Action-Conditioned World Model as an Online Simulator for Robotic Planning
by: Zhou, Zhenhao, et al.
Published: (2025)
by: Zhou, Zhenhao, et al.
Published: (2025)
An Ontology for Unified Modeling of Tasks, Actions, Environments, and Capabilities in Personal Service Robotics
by: Martorana, Margherita, et al.
Published: (2025)
by: Martorana, Margherita, et al.
Published: (2025)
VLM-Vac: Enhancing Smart Vacuums through VLM Knowledge Distillation and Language-Guided Experience Replay
by: Mirjalili, Reihaneh, et al.
Published: (2024)
by: Mirjalili, Reihaneh, et al.
Published: (2024)
DeMaVLA: A Vision-Language-Action Foundation Model for Generalizable Deformable Manipulation
by: Su, Taiyi, et al.
Published: (2026)
by: Su, Taiyi, et al.
Published: (2026)
Action Flow Matching for Continual Robot Learning
by: Murillo-Gonzalez, Alejandro, et al.
Published: (2025)
by: Murillo-Gonzalez, Alejandro, et al.
Published: (2025)
Visuo-Haptic Object Perception for Robots: An Overview
by: Navarro-Guerrero, Nicolás, et al.
Published: (2022)
by: Navarro-Guerrero, Nicolás, et al.
Published: (2022)
Deploying Foundation Model-Enabled Air and Ground Robots in the Field: Challenges and Opportunities
by: Ravichandran, Zachary, et al.
Published: (2025)
by: Ravichandran, Zachary, et al.
Published: (2025)
EO-1: An Open Unified Embodied Foundation Model for General Robot Control
by: Qu, Delin, et al.
Published: (2025)
by: Qu, Delin, et al.
Published: (2025)
ManipDreamer: Boosting Robotic Manipulation World Model with Action Tree and Visual Guidance
by: Li, Ying, et al.
Published: (2025)
by: Li, Ying, et al.
Published: (2025)
A Vision-Language-Action-Critic Model for Robotic Real-World Reinforcement Learning
by: Zhai, Shaopeng, et al.
Published: (2025)
by: Zhai, Shaopeng, et al.
Published: (2025)
Characterizing Vision-Language-Action Models across XPUs: Constraints and Acceleration for On-Robot Deployment
by: Zhou, Kaijun, et al.
Published: (2026)
by: Zhou, Kaijun, et al.
Published: (2026)
Robotic Foundation Models for Industrial Control: A Comprehensive Survey and Readiness Assessment Framework
by: Kube, David, et al.
Published: (2026)
by: Kube, David, et al.
Published: (2026)
Similar Items
-
Logically Constrained Robotics Transformers for Enhanced Perception-Action Planning
by: Kapoor, Parv, et al.
Published: (2024) -
AttenA+: Rectifying Action Inequality in Robotic Foundation Models
by: Peng, Daojie, et al.
Published: (2026) -
Leveraging Large Language Models for Enhancing Autonomous Vehicle Perception
by: Karagounis, Athanasios
Published: (2024) -
Understanding Physical Properties of Unseen Deformable Objects by Leveraging Large Language Models and Robot Actions
by: Park, Changmin, et al.
Published: (2025) -
A Scalable Multi-Robot Framework for Decentralized and Asynchronous Perception-Action-Communication Loops
by: Agarwal, Saurav, et al.
Published: (2023)