DINO-CVA: A Multimodal Goal-Conditioned Vision-to-Action Model for Autonomous Catheter Navigation
Fuente:
arXiv
Saved in:
| Main Authors: | Fekri, Pedram, Roshanfar, Majid, Barbeau, Samuel, Famouri, Seyedfarzad, Looi, Thomas, Podolsky, Dale, Zadeh, Mehrdad, Dargahi, Javad |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
H-Net: A Multitask Architecture for Simultaneous 3D Force Estimation and Stereo Semantic Segmentation in Intracardiac Catheters
by: Fekri, Pedram, et al.
Published: (2024)
by: Fekri, Pedram, et al.
Published: (2024)
TransForSeg: A Multitask Stereo ViT for Joint Stereo Segmentation and 3D Force Estimation in Catheterization
by: Fekri, Pedram, et al.
Published: (2025)
by: Fekri, Pedram, et al.
Published: (2025)
Learning-Based Modeling of a Magnetically Steerable Soft Suction Device for Endoscopic Endonasal Interventions
by: Roshanfar, Majid, et al.
Published: (2025)
by: Roshanfar, Majid, et al.
Published: (2025)
Learning-based Force Sensing and Impedance Matching for Safe Haptic Feedback in Robot-assisted Laparoscopic Surgery
by: Aiden, et al.
Published: (2026)
by: Aiden, et al.
Published: (2026)
CTA: Cross-Task Alignment for Better Test Time Training
by: Barbeau, Samuel, et al.
Published: (2025)
by: Barbeau, Samuel, et al.
Published: (2025)
LatentHDR: Decoupling Exposure from Diffusion via Conditional Latent-to-Latent Mapping for Text/Image-to-Panoramic HDR
by: Fekri, Pedram, et al.
Published: (2026)
by: Fekri, Pedram, et al.
Published: (2026)
Goal-Oriented Semantic Communication for Logical Decision Making
by: Saz, Ahmet Faruk, et al.
Published: (2026)
by: Saz, Ahmet Faruk, et al.
Published: (2026)
Neural Collision Detection for Multi-arm Laparoscopy Surgical Robots Through Learning-from-Simulation
by: Ghiasi, Sarvin, et al.
Published: (2026)
by: Ghiasi, Sarvin, et al.
Published: (2026)
DINO Pre-training for Vision-based End-to-end Autonomous Driving
by: Juneja, Shubham, et al.
Published: (2024)
by: Juneja, Shubham, et al.
Published: (2024)
Autonomous Wheel Loader Navigation Using Goal-Conditioned Actor-Critic MPC
by: Mäki-Penttilä, Aleksi, et al.
Published: (2024)
by: Mäki-Penttilä, Aleksi, et al.
Published: (2024)
Meet Corina Sadler, CVA
by: Megan Venzin
Published: (2024)
by: Megan Venzin
Published: (2024)
CVA Sensitivities, Hedging and Risk
by: Crépey, Stéphane, et al.
Published: (2024)
by: Crépey, Stéphane, et al.
Published: (2024)
Optimus-2: Multimodal Minecraft Agent with Goal-Observation-Action Conditioned Policy
by: Li, Zaijing, et al.
Published: (2025)
by: Li, Zaijing, et al.
Published: (2025)
Improved Actions using The Renormalization Group
by: Segall, Guy, et al.
Published: (2024)
by: Segall, Guy, et al.
Published: (2024)
Revisiting the Learning Objectives of Vision-Language Reward Models
by: Roy, Simon, et al.
Published: (2025)
by: Roy, Simon, et al.
Published: (2025)
Navigating Intelligence: A Survey of Google OR-Tools and Machine Learning for Global Path Planning in Autonomous Vehicles
by: Benoit, Alexandre, et al.
Published: (2025)
by: Benoit, Alexandre, et al.
Published: (2025)
Fast and Stable Credit Gamma of CVA
by: Daluiso, Roberto
Published: (2023)
by: Daluiso, Roberto
Published: (2023)
GoViG: Goal-Conditioned Visual Navigation Instruction Generation via Multimodal Reasoning
by: Wu, Fengyi, et al.
Published: (2025)
by: Wu, Fengyi, et al.
Published: (2025)
Goal2Pixel: Grounding Goals to Pixels for Vision-Language Navigation
by: Bao, Muyi, et al.
Published: (2026)
by: Bao, Muyi, et al.
Published: (2026)
AutoFly: Vision-Language-Action Model for UAV Autonomous Navigation in the Wild
by: Sun, Xiaolou, et al.
Published: (2026)
by: Sun, Xiaolou, et al.
Published: (2026)
Unbounded symbols, heat flow, and Toeplitz operators
by: Looi, Sam
Published: (2026)
by: Looi, Sam
Published: (2026)
A counterexample to the Berger--Coburn conjecture
by: Looi, Sam
Published: (2026)
by: Looi, Sam
Published: (2026)
Positive Berezin liminf does not imply essential positivity for radial Toeplitz operators on Bergman and Fock spaces
by: Looi, Sam
Published: (2026)
by: Looi, Sam
Published: (2026)
AI Tools in Software Development: Developer Perceptions and Usage Patterns
by: Looi, Mark
Published: (2026)
by: Looi, Mark
Published: (2026)
Towards Real-Time Autonomous Navigation: Transformer-Based Catheter Tip Tracking in Fluoroscopy
by: Robertshaw, Harry, et al.
Published: (2026)
by: Robertshaw, Harry, et al.
Published: (2026)
Driving with DINO: Vision Foundation Features as a Unified Bridge for Sim-to-Real Generation in Autonomous Driving
by: Chen, Xuyang, et al.
Published: (2026)
by: Chen, Xuyang, et al.
Published: (2026)
Various metric forms of all type D black holes and their application
by: Podolsky, Jiri
Published: (2025)
by: Podolsky, Jiri
Published: (2025)
Public Libraries in Nineteen States: 1987. Number of Libraries, Library Outlets, and Staff. E.D.TABS.
by: Podolsky, Arthur
Published: (1989)
by: Podolsky, Arthur
Published: (1989)
Academic Libraries: 1988. E.D. TABS.
by: Podolsky, Arthur
Published: (1990)
by: Podolsky, Arthur
Published: (1990)
Public Libraries in 50 States and the District of Columbia: 1989. E.D. TABS.
by: Podolsky, Arthur
Published: (1991)
by: Podolsky, Arthur
Published: (1991)
CULL-MT: Compression Using Language and Layer pruning for Machine Translation
by: Rostami, Pedram, et al.
Published: (2024)
by: Rostami, Pedram, et al.
Published: (2024)
Self-Predictive Representation for Autonomous UAV Object-Goal Navigation
by: Ayala, Angel, et al.
Published: (2026)
by: Ayala, Angel, et al.
Published: (2026)
UnderwaterVLA: Dual-brain Vision-Language-Action architecture for Autonomous Underwater Navigation
by: Wang, Zhangyuan, et al.
Published: (2025)
by: Wang, Zhangyuan, et al.
Published: (2025)
From HNSW to Information-Theoretic Binarization: Rethinking the Architecture of Scalable Vector Search
by: Abtahi, Seyed Moein, et al.
Published: (2025)
by: Abtahi, Seyed Moein, et al.
Published: (2025)
DINO-Foresight: Looking into the Future with DINO
by: Karypidis, Efstathios, et al.
Published: (2024)
by: Karypidis, Efstathios, et al.
Published: (2024)
Multimodal Perception for Goal-oriented Navigation: A Survey
by: Ieong, I-Tak, et al.
Published: (2025)
by: Ieong, I-Tak, et al.
Published: (2025)
Sample-Efficient Learning with Online Expert Correction for Autonomous Catheter Steering in Endovascular Bifurcation Navigation
by: Wang, Hao, et al.
Published: (2026)
by: Wang, Hao, et al.
Published: (2026)
TabulaX: Leveraging Large Language Models for Multi-Class Table Transformations
by: Nobari, Arash Dargahi, et al.
Published: (2024)
by: Nobari, Arash Dargahi, et al.
Published: (2024)
AnySlot: Goal-Conditioned Vision-Language-Action Policies for Zero-Shot Slot-Level Placement
by: Hu, Zhaofeng, et al.
Published: (2026)
by: Hu, Zhaofeng, et al.
Published: (2026)
Goal-Conditioned Reinforcement Learning for Data-Driven Maritime Navigation
by: Vaidheeswaran, Vaishnav, et al.
Published: (2025)
by: Vaidheeswaran, Vaishnav, et al.
Published: (2025)
Similar Items
-
H-Net: A Multitask Architecture for Simultaneous 3D Force Estimation and Stereo Semantic Segmentation in Intracardiac Catheters
by: Fekri, Pedram, et al.
Published: (2024) -
TransForSeg: A Multitask Stereo ViT for Joint Stereo Segmentation and 3D Force Estimation in Catheterization
by: Fekri, Pedram, et al.
Published: (2025) -
Learning-Based Modeling of a Magnetically Steerable Soft Suction Device for Endoscopic Endonasal Interventions
by: Roshanfar, Majid, et al.
Published: (2025) -
Learning-based Force Sensing and Impedance Matching for Safe Haptic Feedback in Robot-assisted Laparoscopic Surgery
by: Aiden, et al.
Published: (2026) -
CTA: Cross-Task Alignment for Better Test Time Training
by: Barbeau, Samuel, et al.
Published: (2025)