From Decision to Action in Surgical Autonomy: Multi-Modal Large Language Models for Robot-Assisted Blood Suction
Fuente:
arXiv
Salvato in:
| Autori principali: | Zargarzadeh, Sadra, Mirzaei, Maryam, Ou, Yafei, Tavakoli, Mahdi |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
A Realistic Surgical Simulator for Non-Rigid and Contact-Rich Manipulation in Surgeries with the da Vinci Research Kit
di: Ou, Yafei, et al.
Pubblicazione: (2024)
di: Ou, Yafei, et al.
Pubblicazione: (2024)
Learning Autonomous Surgical Irrigation and Suction with the da Vinci Research Kit Using Reinforcement Learning
di: Ou, Yafei, et al.
Pubblicazione: (2024)
di: Ou, Yafei, et al.
Pubblicazione: (2024)
CRESSim-MPM: A Material Point Method Library for Surgical Soft Body Simulation with Cutting and Suturing
di: Ou, Yafei, et al.
Pubblicazione: (2025)
di: Ou, Yafei, et al.
Pubblicazione: (2025)
SuctionPrompt: Visual-assisted Robotic Picking with a Suction Cup Using Vision-Language Models and Facile Hardware Design
di: Motoda, Tomohiro, et al.
Pubblicazione: (2024)
di: Motoda, Tomohiro, et al.
Pubblicazione: (2024)
Evaluation of (Shared) Autonomy in Robot‐Assisted Vitreoretinal Surgery Using a Surgical Model
di: Murilo M. Marinho, et al.
Pubblicazione: (2025)
di: Murilo M. Marinho, et al.
Pubblicazione: (2025)
Evaluating Gait Symmetry with a Smart Robotic Walker: A Novel Approach to Mobility Assessment
di: Chalaki, Mahdi, et al.
Pubblicazione: (2024)
di: Chalaki, Mahdi, et al.
Pubblicazione: (2024)
Demonstrating Multi-Suction Item Picking at Scale via Multi-Modal Learning of Pick Success
di: Wang, Che, et al.
Pubblicazione: (2025)
di: Wang, Che, et al.
Pubblicazione: (2025)
MMCD: Multi-Modal Collaborative Decision-Making for Connected Autonomy with Knowledge Distillation
di: Liu, Rui, et al.
Pubblicazione: (2025)
di: Liu, Rui, et al.
Pubblicazione: (2025)
Explaining Autonomy: Enhancing Human-Robot Interaction through Explanation Generation with Large Language Models
di: Sobrín-Hidalgo, David, et al.
Pubblicazione: (2024)
di: Sobrín-Hidalgo, David, et al.
Pubblicazione: (2024)
Suction Leap-Hand: Suction Cups on a Multi-fingered Hand Enable Embodied Dexterity and In-Hand Teleoperation
di: Zhaole, Sun, et al.
Pubblicazione: (2025)
di: Zhaole, Sun, et al.
Pubblicazione: (2025)
ConceptBot: Enhancing Robot's Autonomy through Task Decomposition with Large Language Models and Knowledge Graph
di: Leanza, Alessandro, et al.
Pubblicazione: (2025)
di: Leanza, Alessandro, et al.
Pubblicazione: (2025)
SuFIA: Language-Guided Augmented Dexterity for Robotic Surgical Assistants
di: Moghani, Masoud, et al.
Pubblicazione: (2024)
di: Moghani, Masoud, et al.
Pubblicazione: (2024)
MMaDA-VLA: Large Diffusion Vision-Language-Action Model with Unified Multi-Modal Instruction and Generation
di: Liu, Yang, et al.
Pubblicazione: (2026)
di: Liu, Yang, et al.
Pubblicazione: (2026)
LaMI: Large Language Models for Multi-Modal Human-Robot Interaction
di: Wang, Chao, et al.
Pubblicazione: (2024)
di: Wang, Chao, et al.
Pubblicazione: (2024)
OmniVLA: An Omni-Modal Vision-Language-Action Model for Robot Navigation
di: Hirose, Noriaki, et al.
Pubblicazione: (2025)
di: Hirose, Noriaki, et al.
Pubblicazione: (2025)
ManipVQA: Injecting Robotic Affordance and Physically Grounded Information into Multi-Modal Large Language Models
di: Huang, Siyuan, et al.
Pubblicazione: (2024)
di: Huang, Siyuan, et al.
Pubblicazione: (2024)
ThermoAct:Thermal-Aware Vision-Language-Action Models for Robotic Perception and Decision-Making
di: Son, Young-Chae, et al.
Pubblicazione: (2026)
di: Son, Young-Chae, et al.
Pubblicazione: (2026)
Sign Language: Towards Sign Understanding for Robot Autonomy
di: Agrawal, Ayush, et al.
Pubblicazione: (2025)
di: Agrawal, Ayush, et al.
Pubblicazione: (2025)
SurgSync: Time-Synchronized Multi-Modal Data Collection Framework and Dataset for Surgical Robotics
di: Zhou, Haoying, et al.
Pubblicazione: (2026)
di: Zhou, Haoying, et al.
Pubblicazione: (2026)
Exploring Large Language Models to Facilitate Variable Autonomy for Human-Robot Teaming
di: Lakhnati, Younes, et al.
Pubblicazione: (2023)
di: Lakhnati, Younes, et al.
Pubblicazione: (2023)
Multi-Agent Systems for Robotic Autonomy with LLMs
di: Chen, Junhong, et al.
Pubblicazione: (2025)
di: Chen, Junhong, et al.
Pubblicazione: (2025)
CompSLAM: Complementary Hierarchical Multi-Modal Localization and Mapping for Robot Autonomy in Underground Environments
di: Khattak, Shehryar, et al.
Pubblicazione: (2025)
di: Khattak, Shehryar, et al.
Pubblicazione: (2025)
Decision-Aware Uncertainty Evaluation of Vision-Language Model-Based Early Action Anticipation for Human-Robot Interaction
di: Du, Zhaoda, et al.
Pubblicazione: (2026)
di: Du, Zhaoda, et al.
Pubblicazione: (2026)
The Unified Autonomy Stack: Toward a Blueprint for Generalizable Robot Autonomy
di: Dharmadhikari, Mihir, et al.
Pubblicazione: (2026)
di: Dharmadhikari, Mihir, et al.
Pubblicazione: (2026)
Trace-Focused Diffusion Policy for Multi-Modal Action Disambiguation in Long-Horizon Robotic Manipulation
di: Hu, Yuxuan, et al.
Pubblicazione: (2026)
di: Hu, Yuxuan, et al.
Pubblicazione: (2026)
Force-Constrained Visual Policy: Safe Robot-Assisted Dressing via Multi-Modal Sensing
di: Sun, Zhanyi, et al.
Pubblicazione: (2023)
di: Sun, Zhanyi, et al.
Pubblicazione: (2023)
CoPAL: Corrective Planning of Robot Actions with Large Language Models
di: Joublin, Frank, et al.
Pubblicazione: (2023)
di: Joublin, Frank, et al.
Pubblicazione: (2023)
Incremental Learning for Robot Shared Autonomy
di: Tao, Yiran, et al.
Pubblicazione: (2024)
di: Tao, Yiran, et al.
Pubblicazione: (2024)
Multi Layered Autonomy and AI Ecologies in Robotic Art Installations
di: Chen, Baoyang, et al.
Pubblicazione: (2025)
di: Chen, Baoyang, et al.
Pubblicazione: (2025)
M2R2: MultiModal Robotic Representation for Temporal Action Segmentation
di: Sliwowski, Daniel, et al.
Pubblicazione: (2025)
di: Sliwowski, Daniel, et al.
Pubblicazione: (2025)
FAM-HRI: Foundation-Model Assisted Multi-Modal Human-Robot Interaction Combining Gaze and Speech
di: Lai, Yuzhi, et al.
Pubblicazione: (2025)
di: Lai, Yuzhi, et al.
Pubblicazione: (2025)
Vision-Based Fuzzy Control System for Smart Walkers: Enhancing Usability for Stroke Survivors with Unilateral Upper Limb Impairments
di: Chalaki, Mahdi, et al.
Pubblicazione: (2025)
di: Chalaki, Mahdi, et al.
Pubblicazione: (2025)
Toward RAPS: the Robot Autonomy Perception Scale
di: Silva, Rafael Sousa, et al.
Pubblicazione: (2024)
di: Silva, Rafael Sousa, et al.
Pubblicazione: (2024)
Pre-Surgical Planner for Robot-Assisted Vitreoretinal Surgery: Integrating Eye Posture, Robot Position and Insertion Point
di: Inagaki, Satoshi, et al.
Pubblicazione: (2025)
di: Inagaki, Satoshi, et al.
Pubblicazione: (2025)
Multi-objective Cross-task Learning via Goal-conditioned GPT-based Decision Transformers for Surgical Robot Task Automation
di: Fu, Jiawei, et al.
Pubblicazione: (2024)
di: Fu, Jiawei, et al.
Pubblicazione: (2024)
CognitiveDog: Large Multimodal Model Based System to Translate Vision and Language into Action of Quadruped Robot
di: Lykov, Artem, et al.
Pubblicazione: (2024)
di: Lykov, Artem, et al.
Pubblicazione: (2024)
RAIDER: Tool-Equipped Large Language Model Agent for Robotic Action Issue Detection, Explanation and Recovery
di: Izquierdo-Badiola, Silvia, et al.
Pubblicazione: (2025)
di: Izquierdo-Badiola, Silvia, et al.
Pubblicazione: (2025)
Symbolic Runtime Verification and Adaptive Decision-Making for Robot-Assisted Dressing
di: Rafiq, Yasmin, et al.
Pubblicazione: (2025)
di: Rafiq, Yasmin, et al.
Pubblicazione: (2025)
On Robustness of Vision-Language-Action Model against Multi-Modal Perturbations
di: Guo, Jianing, et al.
Pubblicazione: (2025)
di: Guo, Jianing, et al.
Pubblicazione: (2025)
Haptic search with the Smart Suction Cup on adversarial objects
di: Lee, Jungpyo, et al.
Pubblicazione: (2023)
di: Lee, Jungpyo, et al.
Pubblicazione: (2023)
Documenti analoghi
-
A Realistic Surgical Simulator for Non-Rigid and Contact-Rich Manipulation in Surgeries with the da Vinci Research Kit
di: Ou, Yafei, et al.
Pubblicazione: (2024) -
Learning Autonomous Surgical Irrigation and Suction with the da Vinci Research Kit Using Reinforcement Learning
di: Ou, Yafei, et al.
Pubblicazione: (2024) -
CRESSim-MPM: A Material Point Method Library for Surgical Soft Body Simulation with Cutting and Suturing
di: Ou, Yafei, et al.
Pubblicazione: (2025) -
SuctionPrompt: Visual-assisted Robotic Picking with a Suction Cup Using Vision-Language Models and Facile Hardware Design
di: Motoda, Tomohiro, et al.
Pubblicazione: (2024) -
Evaluation of (Shared) Autonomy in Robot‐Assisted Vitreoretinal Surgery Using a Surgical Model
di: Murilo M. Marinho, et al.
Pubblicazione: (2025)