TransForSeg: A Multitask Stereo ViT for Joint Stereo Segmentation and 3D Force Estimation in Catheterization
Fuente:
arXiv
Saved in:
| Main Authors: | Fekri, Pedram, Zadeh, Mehrdad, Dargahi, Javad |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
H-Net: A Multitask Architecture for Simultaneous 3D Force Estimation and Stereo Semantic Segmentation in Intracardiac Catheters
by: Fekri, Pedram, et al.
Published: (2024)
by: Fekri, Pedram, et al.
Published: (2024)
DINO-CVA: A Multimodal Goal-Conditioned Vision-to-Action Model for Autonomous Catheter Navigation
by: Fekri, Pedram, et al.
Published: (2025)
by: Fekri, Pedram, et al.
Published: (2025)
S$^3$M-Net: Joint Learning of Semantic Segmentation and Stereo Matching for Autonomous Driving
by: Wu, Zhiyuan, et al.
Published: (2024)
by: Wu, Zhiyuan, et al.
Published: (2024)
StereoAdapter: Adapting Stereo Depth Estimation to Underwater Scenes
by: Wu, Zhengri, et al.
Published: (2025)
by: Wu, Zhengri, et al.
Published: (2025)
FIReStereo: Forest InfraRed Stereo Dataset for UAS Depth Perception in Visually Degraded Environments
by: Dhrafani, Devansh, et al.
Published: (2024)
by: Dhrafani, Devansh, et al.
Published: (2024)
Helvipad: A Real-World Dataset for Omnidirectional Stereo Depth Estimation
by: Zayene, Mehdi, et al.
Published: (2024)
by: Zayene, Mehdi, et al.
Published: (2024)
MLG-Stereo: ViT Based Stereo Matching with Multi-Stage Local-Global Enhancement
by: Zhang, Haoyu, et al.
Published: (2026)
by: Zhang, Haoyu, et al.
Published: (2026)
SENSE: Stereo OpEN Vocabulary SEmantic Segmentation
by: Campagnolo, Thomas, et al.
Published: (2026)
by: Campagnolo, Thomas, et al.
Published: (2026)
Playing to Vision Foundation Model's Strengths in Stereo Matching
by: Liu, Chuang-Wei, et al.
Published: (2024)
by: Liu, Chuang-Wei, et al.
Published: (2024)
{S\textsuperscript{2}M\textsuperscript{2}}: Scalable Stereo Matching Model for Reliable Depth Estimation
by: Min, Junhong, et al.
Published: (2025)
by: Min, Junhong, et al.
Published: (2025)
StereoPolicy: Improving Robotic Manipulation Policies via Stereo Perception
by: Han, Evans, et al.
Published: (2026)
by: Han, Evans, et al.
Published: (2026)
Fast-FoundationStereo: Real-Time Zero-Shot Stereo Matching
by: Wen, Bowen, et al.
Published: (2025)
by: Wen, Bowen, et al.
Published: (2025)
Event-based Stereo Depth Estimation: A Survey
by: Ghosh, Suman, et al.
Published: (2024)
by: Ghosh, Suman, et al.
Published: (2024)
Learning Robust Stereo Matching in the Wild with Selective Mixture-of-Experts
by: Wang, Yun, et al.
Published: (2025)
by: Wang, Yun, et al.
Published: (2025)
FoundationStereo: Zero-Shot Stereo Matching
by: Wen, Bowen, et al.
Published: (2025)
by: Wen, Bowen, et al.
Published: (2025)
ViTA-Seg: Vision Transformer for Amodal Segmentation in Robotics
by: Caramia, Donato, et al.
Published: (2025)
by: Caramia, Donato, et al.
Published: (2025)
TiCoSS: Tightening the Coupling between Semantic Segmentation and Stereo Matching within A Joint Learning Framework
by: Tang, Guanfeng, et al.
Published: (2024)
by: Tang, Guanfeng, et al.
Published: (2024)
Object Depth and Size Estimation using Stereo-vision and Integration with SLAM
by: Hamad, Layth, et al.
Published: (2024)
by: Hamad, Layth, et al.
Published: (2024)
KVN: Keypoints Voting Network with Differentiable RANSAC for Stereo Pose Estimation
by: Donadi, Ivano, et al.
Published: (2023)
by: Donadi, Ivano, et al.
Published: (2023)
SurgPose: Generalisable Surgical Instrument Pose Estimation using Zero-Shot Learning and Stereo Vision
by: Rai, Utsav, et al.
Published: (2025)
by: Rai, Utsav, et al.
Published: (2025)
StereoVAE: A lightweight stereo-matching system using embedded GPUs
by: Chang, Qiong, et al.
Published: (2023)
by: Chang, Qiong, et al.
Published: (2023)
Mobile Robotic Multi-View Photometric Stereo
by: Kumar, Suryansh
Published: (2025)
by: Kumar, Suryansh
Published: (2025)
Icy Moon Surface Simulation and Stereo Depth Estimation for Sampling Autonomy
by: Bhaskara, Ramchander, et al.
Published: (2024)
by: Bhaskara, Ramchander, et al.
Published: (2024)
SToRe3D: Sparse Token Relevance in ViTs for Efficient Multi-View 3D Object Detection
by: Papais, Sandro, et al.
Published: (2026)
by: Papais, Sandro, et al.
Published: (2026)
Stereo Hand-Object Reconstruction for Human-to-Robot Handover
by: Pang, Yik Lung, et al.
Published: (2024)
by: Pang, Yik Lung, et al.
Published: (2024)
IMU-Aided Event-based Stereo Visual Odometry
by: Niu, Junkai, et al.
Published: (2024)
by: Niu, Junkai, et al.
Published: (2024)
VAT: Vision Action Transformer by Unlocking Full Representation of ViT
by: Li, Wenhao, et al.
Published: (2025)
by: Li, Wenhao, et al.
Published: (2025)
Boosting Omnidirectional Stereo Matching with a Pre-trained Depth Foundation Model
by: Endres, Jannik, et al.
Published: (2025)
by: Endres, Jannik, et al.
Published: (2025)
Dusk Till Dawn: Self-supervised Nighttime Stereo Depth Estimation using Visual Foundation Models
by: Vankadari, Madhu, et al.
Published: (2024)
by: Vankadari, Madhu, et al.
Published: (2024)
CMD: Constraining Multimodal Distribution for Domain Adaptation in Stereo Matching
by: Shen, Zhelun, et al.
Published: (2025)
by: Shen, Zhelun, et al.
Published: (2025)
BridgeDepth: Bridging Monocular and Stereo Reasoning with Latent Alignment
by: Guan, Tongfan, et al.
Published: (2025)
by: Guan, Tongfan, et al.
Published: (2025)
High-Definition 5MP Stereo Vision Sensing for Robotics
by: Jiang, Leaf, et al.
Published: (2026)
by: Jiang, Leaf, et al.
Published: (2026)
EV-MGDispNet: Motion-Guided Event-Based Stereo Disparity Estimation Network with Left-Right Consistency
by: Jiang, Junjie, et al.
Published: (2024)
by: Jiang, Junjie, et al.
Published: (2024)
ODTFormer: Efficient Obstacle Detection and Tracking with Stereo Cameras Based on Transformer
by: Ding, Tianye, et al.
Published: (2024)
by: Ding, Tianye, et al.
Published: (2024)
ESVO2: Direct Visual-Inertial Odometry with Stereo Event Cameras
by: Niu, Junkai, et al.
Published: (2024)
by: Niu, Junkai, et al.
Published: (2024)
ClearDepth: Enhanced Stereo Perception of Transparent Objects for Robotic Manipulation
by: Bai, Kaixin, et al.
Published: (2024)
by: Bai, Kaixin, et al.
Published: (2024)
ActMVS: Active Scene Reconstruction with Monocular Multi-View Stereo
by: Pu, Guo, et al.
Published: (2026)
by: Pu, Guo, et al.
Published: (2026)
High-Speed Stereo Visual SLAM for Low-Powered Computing Devices
by: Kumar, Ashish, et al.
Published: (2024)
by: Kumar, Ashish, et al.
Published: (2024)
Dive Deeper into Rectifying Homography for Stereo Camera Online Self-Calibration
by: Zhao, Hongbo, et al.
Published: (2023)
by: Zhao, Hongbo, et al.
Published: (2023)
ViT-VS: On the Applicability of Pretrained Vision Transformer Features for Generalizable Visual Servoing
by: Scherl, Alessandro, et al.
Published: (2025)
by: Scherl, Alessandro, et al.
Published: (2025)
Similar Items
-
H-Net: A Multitask Architecture for Simultaneous 3D Force Estimation and Stereo Semantic Segmentation in Intracardiac Catheters
by: Fekri, Pedram, et al.
Published: (2024) -
DINO-CVA: A Multimodal Goal-Conditioned Vision-to-Action Model for Autonomous Catheter Navigation
by: Fekri, Pedram, et al.
Published: (2025) -
S$^3$M-Net: Joint Learning of Semantic Segmentation and Stereo Matching for Autonomous Driving
by: Wu, Zhiyuan, et al.
Published: (2024) -
StereoAdapter: Adapting Stereo Depth Estimation to Underwater Scenes
by: Wu, Zhengri, et al.
Published: (2025) -
FIReStereo: Forest InfraRed Stereo Dataset for UAS Depth Perception in Visually Degraded Environments
by: Dhrafani, Devansh, et al.
Published: (2024)