Data-Efficient Learning for Generalizable Surgical Video Understanding
Fuente:
arXiv
Saved in:
| Main Author: | Nasirihaghighi, Sahar |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Dual Invariance Self-training for Reliable Semi-supervised Surgical Phase Recognition
by: Nasirihaghighi, Sahar, et al.
Published: (2025)
by: Nasirihaghighi, Sahar, et al.
Published: (2025)
SASVi -- Segment Any Surgical Video
by: Sivakumar, Ssharvien Kumar, et al.
Published: (2025)
by: Sivakumar, Ssharvien Kumar, et al.
Published: (2025)
VidFuncta: Towards Generalizable Neural Representations for Ultrasound Videos
by: Wolleb, Julia, et al.
Published: (2025)
by: Wolleb, Julia, et al.
Published: (2025)
Ophora: A Large-Scale Data-Driven Text-Guided Ophthalmic Surgical Video Generation Model
by: Li, Wei, et al.
Published: (2025)
by: Li, Wei, et al.
Published: (2025)
Slot-BERT: Self-supervised Object Discovery in Surgical Video
by: Liao, Guiqiu, et al.
Published: (2025)
by: Liao, Guiqiu, et al.
Published: (2025)
Weakly Supervised YOLO Network for Surgical Instrument Localization in Endoscopic Videos
by: Wei, Rongfeng, et al.
Published: (2023)
by: Wei, Rongfeng, et al.
Published: (2023)
Spatio-Temporal Representation Decoupling and Enhancement for Federated Instrument Segmentation in Surgical Videos
by: Fang, Zheng, et al.
Published: (2025)
by: Fang, Zheng, et al.
Published: (2025)
Towards Robust and Generalizable Lensless Imaging with Modular Learned Reconstruction
by: Bezzam, Eric, et al.
Published: (2025)
by: Bezzam, Eric, et al.
Published: (2025)
SurgTPGS: Semantic 3D Surgical Scene Understanding with Text Promptable Gaussian Splatting
by: Huang, Yiming, et al.
Published: (2025)
by: Huang, Yiming, et al.
Published: (2025)
Style Content Decomposition-based Data Augmentation for Domain Generalizable Medical Image Segmentation
by: Shen, Zhiqiang, et al.
Published: (2025)
by: Shen, Zhiqiang, et al.
Published: (2025)
Zero-Shot Surgical Tool Segmentation in Monocular Video Using Segment Anything Model 2
by: Lou, Ange, et al.
Published: (2024)
by: Lou, Ange, et al.
Published: (2024)
Towards Generalizable Tumor Synthesis
by: Chen, Qi, et al.
Published: (2024)
by: Chen, Qi, et al.
Published: (2024)
GBT-SAM: A Parameter-Efficient Depth-Aware Model for Generalizable Brain tumour Segmentation on mp-MRI
by: Diana-Albelda, Cecilia, et al.
Published: (2025)
by: Diana-Albelda, Cecilia, et al.
Published: (2025)
GENRE-CMR: Generalizable Deep Learning for Diverse Multi-Domain Cardiac MRI Reconstruction
by: Hamedani, Kian Anvari, et al.
Published: (2025)
by: Hamedani, Kian Anvari, et al.
Published: (2025)
Transfer CLIP for Generalizable Image Denoising
by: Cheng, Jun, et al.
Published: (2024)
by: Cheng, Jun, et al.
Published: (2024)
Toward Zero-Shot Learning for Visual Dehazing of Urological Surgical Robots
by: Wu, Renkai, et al.
Published: (2024)
by: Wu, Renkai, et al.
Published: (2024)
Towards Robust and Generalizable Continuous Space-Time Video Super-Resolution with Events
by: Wei, Shuoyan, et al.
Published: (2025)
by: Wei, Shuoyan, et al.
Published: (2025)
ICME 2025 Generalizable HDR and SDR Video Quality Measurement Grand Challenge
by: Chen, Yixu, et al.
Published: (2025)
by: Chen, Yixu, et al.
Published: (2025)
A Coding Framework and Benchmark towards Low-Bitrate Video Understanding
by: Tian, Yuan, et al.
Published: (2022)
by: Tian, Yuan, et al.
Published: (2022)
GANESH: Generalizable NeRF for Lensless Imaging
by: Madavan, Rakesh Raj, et al.
Published: (2024)
by: Madavan, Rakesh Raj, et al.
Published: (2024)
pyMEAL: A Multi-Encoder Augmentation-Aware Learning for Robust and Generalizable Medical Image Translation
by: Ilyas, Abdul-mojeed Olabisi, et al.
Published: (2025)
by: Ilyas, Abdul-mojeed Olabisi, et al.
Published: (2025)
CMTA: Leveraging Cross-Modal Temporal Artifacts for Generalizable AI-Generated Video Detection
by: Wang, Hang, et al.
Published: (2026)
by: Wang, Hang, et al.
Published: (2026)
Toward Generalizable Multiple Sclerosis Lesion Segmentation Models
by: Badea, Liviu, et al.
Published: (2024)
by: Badea, Liviu, et al.
Published: (2024)
Surgical SAM 2: Real-time Segment Anything in Surgical Video by Efficient Frame Pruning
by: Liu, Haofeng, et al.
Published: (2024)
by: Liu, Haofeng, et al.
Published: (2024)
ReSW-VL: Representation Learning for Surgical Workflow Analysis Using Vision-Language Model
by: Kondo, Satoshi
Published: (2025)
by: Kondo, Satoshi
Published: (2025)
Segmentation-Guided Spatial Indexing for Generalizable and Explainable Deepfake Detection
by: Al-Zyoud, Izaldein, et al.
Published: (2026)
by: Al-Zyoud, Izaldein, et al.
Published: (2026)
When Eye-Tracking Meets Machine Learning: A Systematic Review on Applications in Medical Image Analysis
by: Moradizeyveh, Sahar, et al.
Published: (2024)
by: Moradizeyveh, Sahar, et al.
Published: (2024)
An Efficient and Generalizable Transfer Learning Method for Weather Condition Detection on Ground Terminals
by: Zhang, Wenxuan, et al.
Published: (2025)
by: Zhang, Wenxuan, et al.
Published: (2025)
Taming Diffusion Transformer for Efficient Mobile Video Generation in Seconds
by: Wu, Yushu, et al.
Published: (2025)
by: Wu, Yushu, et al.
Published: (2025)
A Continual Learning-driven Model for Accurate and Generalizable Segmentation of Clinically Comprehensive and Fine-grained Whole-body Anatomies in CT
by: Guo, Dazhou, et al.
Published: (2025)
by: Guo, Dazhou, et al.
Published: (2025)
SAM 2 in Robotic Surgery: An Empirical Evaluation for Robustness and Generalization in Surgical Video Segmentation
by: Yu, Jieming, et al.
Published: (2024)
by: Yu, Jieming, et al.
Published: (2024)
Noise-Inspired Diffusion Model for Generalizable Low-Dose CT Reconstruction
by: Gao, Qi, et al.
Published: (2025)
by: Gao, Qi, et al.
Published: (2025)
GS-Marker: Generalizable and Robust Watermarking for 3D Gaussian Splatting
by: Li, Lijiang, et al.
Published: (2025)
by: Li, Lijiang, et al.
Published: (2025)
Radiologist-in-the-Loop Self-Training for Generalizable CT Metal Artifact Reduction
by: Ma, Chenglong, et al.
Published: (2025)
by: Ma, Chenglong, et al.
Published: (2025)
Generalizable CT-Free PET Attenuation and Scatter Correction for Pediatric Patients
by: Wu, Jia-Mian, et al.
Published: (2026)
by: Wu, Jia-Mian, et al.
Published: (2026)
Multi-Aperture Fusion of Transformer-Convolutional Network (MFTC-Net) for 3D Medical Image Segmentation and Visualization
by: Shabani, Siyavash, et al.
Published: (2024)
by: Shabani, Siyavash, et al.
Published: (2024)
LeanVAE: An Ultra-Efficient Reconstruction VAE for Video Diffusion Models
by: Cheng, Yu, et al.
Published: (2025)
by: Cheng, Yu, et al.
Published: (2025)
Video Generation Models as World Models: Efficient Paradigms, Architectures and Algorithms
by: He, Muyang, et al.
Published: (2026)
by: He, Muyang, et al.
Published: (2026)
Understanding Benefits and Pitfalls of Current Methods for the Segmentation of Undersampled MRI Data
by: Morshuis, Jan Nikolas, et al.
Published: (2025)
by: Morshuis, Jan Nikolas, et al.
Published: (2025)
ConStyX: Content Style Augmentation for Generalizable Medical Image Segmentation
by: Chen, Xi, et al.
Published: (2025)
by: Chen, Xi, et al.
Published: (2025)
Similar Items
-
Dual Invariance Self-training for Reliable Semi-supervised Surgical Phase Recognition
by: Nasirihaghighi, Sahar, et al.
Published: (2025) -
SASVi -- Segment Any Surgical Video
by: Sivakumar, Ssharvien Kumar, et al.
Published: (2025) -
VidFuncta: Towards Generalizable Neural Representations for Ultrasound Videos
by: Wolleb, Julia, et al.
Published: (2025) -
Ophora: A Large-Scale Data-Driven Text-Guided Ophthalmic Surgical Video Generation Model
by: Li, Wei, et al.
Published: (2025) -
Slot-BERT: Self-supervised Object Discovery in Surgical Video
by: Liao, Guiqiu, et al.
Published: (2025)