Vision Language Action Models in Robotic Manipulation: A Systematic Review
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Din, Muhayy Ud, Akram, Waseem, Saoud, Lyes Saad, Rosell, Jan, Hussain, Irfan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
LLM-VLM Fusion Framework for Autonomous Maritime Port Inspection using a Heterogeneous UAV-USV System
von: Din, Muhayy Ud, et al.
Veröffentlicht: (2026)
von: Din, Muhayy Ud, et al.
Veröffentlicht: (2026)
Benchmarking Vision-Based Object Tracking for USVs in Complex Maritime Environments
von: Din, Muhayy Ud, et al.
Veröffentlicht: (2024)
von: Din, Muhayy Ud, et al.
Veröffentlicht: (2024)
EBA-AI: Ethics-Guided Bias-Aware AI for Efficient Underwater Image Enhancement and Coral Reef Monitoring
von: Saoud, Lyes Saad, et al.
Veröffentlicht: (2025)
von: Saoud, Lyes Saad, et al.
Veröffentlicht: (2025)
FishDet-M: A Unified Large-Scale Benchmark for Robust Fish Detection and CLIP-Guided Model Selection in Diverse Aquatic Visual Domains
von: Abujabal, Muayad, et al.
Veröffentlicht: (2025)
von: Abujabal, Muayad, et al.
Veröffentlicht: (2025)
A Review of Generative AI in Aquaculture: Foundations, Applications, and Future Directions for Smart and Sustainable Farming
von: Akram, Waseem, et al.
Veröffentlicht: (2025)
von: Akram, Waseem, et al.
Veröffentlicht: (2025)
Lang2Manip: A Tool for LLM-Based Symbolic-to-Geometric Planning for Manipulation
von: Din, Muhayy Ud, et al.
Veröffentlicht: (2025)
von: Din, Muhayy Ud, et al.
Veröffentlicht: (2025)
Real-Time Threaded Houbara Detection and Segmentation for Wildlife Conservation using Mobile Platforms
von: Saoud, Lyes Saad, et al.
Veröffentlicht: (2025)
von: Saoud, Lyes Saad, et al.
Veröffentlicht: (2025)
AquaChat: An LLM-Guided ROV Framework for Adaptive Inspection of Aquaculture Net Pens
von: Akram, Waseem, et al.
Veröffentlicht: (2025)
von: Akram, Waseem, et al.
Veröffentlicht: (2025)
MVTD: A Benchmark Dataset for Maritime Visual Object Tracking
von: Bakht, Ahsan Baidar, et al.
Veröffentlicht: (2025)
von: Bakht, Ahsan Baidar, et al.
Veröffentlicht: (2025)
LLM-guided Task and Motion Planning using Knowledge-based Reasoning
von: Din, Muhayy Ud, et al.
Veröffentlicht: (2024)
von: Din, Muhayy Ud, et al.
Veröffentlicht: (2024)
Maritime Mission Planning for Unmanned Surface Vessel using Large Language Model
von: Din, Muhayy Ud, et al.
Veröffentlicht: (2025)
von: Din, Muhayy Ud, et al.
Veröffentlicht: (2025)
Bio-Inspired Robotic Houbara: From Development to Field Deployment for Behavioral Studies
von: Saoud, Lyes Saad, et al.
Veröffentlicht: (2025)
von: Saoud, Lyes Saad, et al.
Veröffentlicht: (2025)
Object Manipulation in Marine Environments using Reinforcement Learning
von: Nader, Ahmed, et al.
Veröffentlicht: (2024)
von: Nader, Ahmed, et al.
Veröffentlicht: (2024)
Prioritized Real‐Time UAV‐Based Vessel Detection for Efficient Maritime Search
von: Lyes Saad Saoud, et al.
Veröffentlicht: (2025)
von: Lyes Saad Saoud, et al.
Veröffentlicht: (2025)
Robust Collision Detection for Robots with Variable Stiffness Actuation by Using MAD-CNN: Modularized-Attention-Dilated Convolutional Neural Network
von: Niu, Zhenwei, et al.
Veröffentlicht: (2023)
von: Niu, Zhenwei, et al.
Veröffentlicht: (2023)
BitVLA: 1-bit Vision-Language-Action Models for Robotics Manipulation
von: Wang, Hongyu, et al.
Veröffentlicht: (2025)
von: Wang, Hongyu, et al.
Veröffentlicht: (2025)
MemoryVLA: Perceptual-Cognitive Memory in Vision-Language-Action Models for Robotic Manipulation
von: Shi, Hao, et al.
Veröffentlicht: (2025)
von: Shi, Hao, et al.
Veröffentlicht: (2025)
MAP-VLA: Memory-Augmented Prompting for Vision-Language-Action Model in Robotic Manipulation
von: Li, Runhao, et al.
Veröffentlicht: (2025)
von: Li, Runhao, et al.
Veröffentlicht: (2025)
Large VLM-based Vision-Language-Action Models for Robotic Manipulation: A Survey
von: Shao, Rui, et al.
Veröffentlicht: (2025)
von: Shao, Rui, et al.
Veröffentlicht: (2025)
EveryDayVLA: A Vision-Language-Action Model for Affordable Robotic Manipulation
von: Chopra, Samarth, et al.
Veröffentlicht: (2025)
von: Chopra, Samarth, et al.
Veröffentlicht: (2025)
See, Plan, Rewind: Progress-Aware Vision-Language-Action Models for Robust Robotic Manipulation
von: Dai, Tingjun, et al.
Veröffentlicht: (2026)
von: Dai, Tingjun, et al.
Veröffentlicht: (2026)
SaPaVe: Towards Active Perception and Manipulation in Vision-Language-Action Models for Robotics
von: Liu, Mengzhen, et al.
Veröffentlicht: (2026)
von: Liu, Mengzhen, et al.
Veröffentlicht: (2026)
TinyVLA: Towards Fast, Data-Efficient Vision-Language-Action Models for Robotic Manipulation
von: Wen, Junjie, et al.
Veröffentlicht: (2024)
von: Wen, Junjie, et al.
Veröffentlicht: (2024)
Observe Then Act: Asynchronous Active Vision-Action Model for Robotic Manipulation
von: Wang, Guokang, et al.
Veröffentlicht: (2024)
von: Wang, Guokang, et al.
Veröffentlicht: (2024)
Explainable Adversarial-Robust Vision-Language-Action Model for Robotic Manipulation
von: Kim, Ju-Young, et al.
Veröffentlicht: (2025)
von: Kim, Ju-Young, et al.
Veröffentlicht: (2025)
Conditional Variational Auto Encoder Based Dynamic Motion for Multi-task Imitation Learning
von: Xu, Binzhao, et al.
Veröffentlicht: (2024)
von: Xu, Binzhao, et al.
Veröffentlicht: (2024)
Autonomous Underwater Robotic System for Aquaculture Applications
von: Akram, Waseem, et al.
Veröffentlicht: (2023)
von: Akram, Waseem, et al.
Veröffentlicht: (2023)
NaVILA: Legged Robot Vision-Language-Action Model for Navigation
von: Cheng, An-Chieh, et al.
Veröffentlicht: (2024)
von: Cheng, An-Chieh, et al.
Veröffentlicht: (2024)
QUAR-VLA: Vision-Language-Action Model for Quadruped Robots
von: Ding, Pengxiang, et al.
Veröffentlicht: (2023)
von: Ding, Pengxiang, et al.
Veröffentlicht: (2023)
DynamicVLA: A Vision-Language-Action Model for Dynamic Object Manipulation
von: Xie, Haozhe, et al.
Veröffentlicht: (2026)
von: Xie, Haozhe, et al.
Veröffentlicht: (2026)
What Matters in Building Vision-Language-Action Models for Generalist Robots
von: Li, Xinghang, et al.
Veröffentlicht: (2024)
von: Li, Xinghang, et al.
Veröffentlicht: (2024)
FD-VLA: Force-Distilled Vision-Language-Action Model for Contact-Rich Manipulation
von: Zhao, Ruiteng, et al.
Veröffentlicht: (2026)
von: Zhao, Ruiteng, et al.
Veröffentlicht: (2026)
Manipulate-Anything: Automating Real-World Robots using Vision-Language Models
von: Duan, Jiafei, et al.
Veröffentlicht: (2024)
von: Duan, Jiafei, et al.
Veröffentlicht: (2024)
From Pixels to Tokens: A Systematic Study of Latent Action Supervision for Vision-Language-Action Models
von: Lin, Yihan, et al.
Veröffentlicht: (2026)
von: Lin, Yihan, et al.
Veröffentlicht: (2026)
Exploring the Limits of Vision-Language-Action Manipulations in Cross-task Generalization
von: Zhou, Jiaming, et al.
Veröffentlicht: (2025)
von: Zhou, Jiaming, et al.
Veröffentlicht: (2025)
ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver
von: Song, Wenxuan, et al.
Veröffentlicht: (2025)
von: Song, Wenxuan, et al.
Veröffentlicht: (2025)
RoboGround: Robotic Manipulation with Grounded Vision-Language Priors
von: Huang, Haifeng, et al.
Veröffentlicht: (2025)
von: Huang, Haifeng, et al.
Veröffentlicht: (2025)
MUOT_3M: A 3 Million Frame Multimodal Underwater Benchmark and the MUTrack Tracking Method
von: Bakht, Ahsan Baidar, et al.
Veröffentlicht: (2026)
von: Bakht, Ahsan Baidar, et al.
Veröffentlicht: (2026)
InstructVLA: Vision-Language-Action Instruction Tuning from Understanding to Manipulation
von: Yang, Shuai, et al.
Veröffentlicht: (2025)
von: Yang, Shuai, et al.
Veröffentlicht: (2025)
Physically Grounded Vision-Language Models for Robotic Manipulation
von: Gao, Jensen, et al.
Veröffentlicht: (2023)
von: Gao, Jensen, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
LLM-VLM Fusion Framework for Autonomous Maritime Port Inspection using a Heterogeneous UAV-USV System
von: Din, Muhayy Ud, et al.
Veröffentlicht: (2026) -
Benchmarking Vision-Based Object Tracking for USVs in Complex Maritime Environments
von: Din, Muhayy Ud, et al.
Veröffentlicht: (2024) -
EBA-AI: Ethics-Guided Bias-Aware AI for Efficient Underwater Image Enhancement and Coral Reef Monitoring
von: Saoud, Lyes Saad, et al.
Veröffentlicht: (2025) -
FishDet-M: A Unified Large-Scale Benchmark for Robust Fish Detection and CLIP-Guided Model Selection in Diverse Aquatic Visual Domains
von: Abujabal, Muayad, et al.
Veröffentlicht: (2025) -
A Review of Generative AI in Aquaculture: Foundations, Applications, and Future Directions for Smart and Sustainable Farming
von: Akram, Waseem, et al.
Veröffentlicht: (2025)