Post-surgical Endometriosis Segmentation in Laparoscopic Videos
Fuente:
arXiv
Saved in:
| Main Authors: | Leibetseder, Andreas, Schoeffmann, Klaus, Keckstein, Jörg, Keckstein, Simon |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
GLENDA: Gynecologic Laparoscopy Endometriosis Dataset
by: Leibetseder, Andreas, et al.
Published: (2025)
by: Leibetseder, Andreas, et al.
Published: (2025)
diveXplore 6.0: ITEC's Interactive Video Exploration System at VBS 2022
by: Leibetseder, Andreas, et al.
Published: (2025)
by: Leibetseder, Andreas, et al.
Published: (2025)
Identifying Surgical Instruments in Laparoscopy Using Deep Learning Instance Segmentation
by: Kletz, Sabrina, et al.
Published: (2025)
by: Kletz, Sabrina, et al.
Published: (2025)
lifeXplore at the Lifelog Search Challenge 2020
by: Leibetseder, Andreas, et al.
Published: (2025)
by: Leibetseder, Andreas, et al.
Published: (2025)
Less is More - diveXplore 5.0 at VBS 2021
by: Leibetseder, Andreas, et al.
Published: (2025)
by: Leibetseder, Andreas, et al.
Published: (2025)
lifeXplore at the Lifelog Search Challenge 2021
by: Leibetseder, Andreas, et al.
Published: (2025)
by: Leibetseder, Andreas, et al.
Published: (2025)
Omnidirectional Video Super-Resolution using Deep Learning
by: Baniya, Arbind Agrahari, et al.
Published: (2025)
by: Baniya, Arbind Agrahari, et al.
Published: (2025)
MVP: Winning Solution to SMP Challenge 2025 Video Track
by: Ye, Liliang, et al.
Published: (2025)
by: Ye, Liliang, et al.
Published: (2025)
Learning from Mistakes: Self-Regularizing Hierarchical Representations in Point Cloud Semantic Segmentation
by: Camuffo, Elena, et al.
Published: (2023)
by: Camuffo, Elena, et al.
Published: (2023)
Competitive Learning for Achieving Content-specific Filters in Video Coding for Machines
by: Zhang, Honglei, et al.
Published: (2024)
by: Zhang, Honglei, et al.
Published: (2024)
CinePile: A Long Video Question Answering Dataset and Benchmark
by: Rawal, Ruchit, et al.
Published: (2024)
by: Rawal, Ruchit, et al.
Published: (2024)
360VFI: A Dataset and Benchmark for Omnidirectional Video Frame Interpolation
by: Lu, Wenxuan, et al.
Published: (2024)
by: Lu, Wenxuan, et al.
Published: (2024)
Catalogue Grounded Multimodal Attribution for Museum Video under Resource and Regulatory Constraints
by: Nanang, Minsak, et al.
Published: (2026)
by: Nanang, Minsak, et al.
Published: (2026)
DeCo-VAE: Learning Compact Latents for Video Reconstruction via Decoupled Representation
by: Yin, Xiangchen, et al.
Published: (2025)
by: Yin, Xiangchen, et al.
Published: (2025)
LinVT: Empower Your Image-level Large Language Model to Understand Videos
by: Gao, Lishuai, et al.
Published: (2024)
by: Gao, Lishuai, et al.
Published: (2024)
CreativeVR: Diffusion-Prior-Guided Approach for Structure and Motion Restoration in Generative and Real Videos
by: Panambur, Tejas, et al.
Published: (2025)
by: Panambur, Tejas, et al.
Published: (2025)
Video DataFlywheel: Resolving the Impossible Data Trinity in Video-Language Understanding
by: Wang, Xiao, et al.
Published: (2024)
by: Wang, Xiao, et al.
Published: (2024)
CLIP as RNN: Segment Countless Visual Concepts without Training Endeavor
by: Sun, Shuyang, et al.
Published: (2023)
by: Sun, Shuyang, et al.
Published: (2023)
EditYourself: Audio-Driven Generation and Manipulation of Talking Head Videos with Diffusion Transformers
by: Flynn, John, et al.
Published: (2026)
by: Flynn, John, et al.
Published: (2026)
VidMuse: A Simple Video-to-Music Generation Framework with Long-Short-Term Modeling
by: Tian, Zeyue, et al.
Published: (2024)
by: Tian, Zeyue, et al.
Published: (2024)
InteractiveVideo: User-Centric Controllable Video Generation with Synergistic Multimodal Instructions
by: Zhang, Yiyuan, et al.
Published: (2024)
by: Zhang, Yiyuan, et al.
Published: (2024)
Control-A-Video: Controllable Text-to-Video Diffusion Models with Motion Prior and Reward Feedback Learning
by: Chen, Weifeng, et al.
Published: (2023)
by: Chen, Weifeng, et al.
Published: (2023)
RAVEN: Query-Guided Representation Alignment for Question Answering over Audio, Video, Embedded Sensors, and Natural Language
by: Biswas, Subrata, et al.
Published: (2025)
by: Biswas, Subrata, et al.
Published: (2025)
Evaluating the Impact of Point Cloud Colorization on Semantic Segmentation Accuracy
by: Zhu, Qinfeng, et al.
Published: (2024)
by: Zhu, Qinfeng, et al.
Published: (2024)
Progressive Confident Masking Attention Network for Audio-Visual Segmentation
by: Wang, Yuxuan, et al.
Published: (2024)
by: Wang, Yuxuan, et al.
Published: (2024)
LongVALE: Vision-Audio-Language-Event Benchmark Towards Time-Aware Omni-Modal Perception of Long Videos
by: Geng, Tiantian, et al.
Published: (2024)
by: Geng, Tiantian, et al.
Published: (2024)
David and Goliath: Small One-step Model Beats Large Diffusion with Score Post-training
by: Luo, Weijian, et al.
Published: (2024)
by: Luo, Weijian, et al.
Published: (2024)
Diffusion Model-Based Video Editing: A Survey
by: Sun, Wenhao, et al.
Published: (2024)
by: Sun, Wenhao, et al.
Published: (2024)
STIV: Scalable Text and Image Conditioned Video Generation
by: Lin, Zongyu, et al.
Published: (2024)
by: Lin, Zongyu, et al.
Published: (2024)
Storybooth: Training-free Multi-Subject Consistency for Improved Visual Storytelling
by: Singh, Jaskirat, et al.
Published: (2025)
by: Singh, Jaskirat, et al.
Published: (2025)
Zero-shot image privacy classification with Vision-Language Models
by: Baia, Alina Elena, et al.
Published: (2025)
by: Baia, Alina Elena, et al.
Published: (2025)
PMPGuard: Catching Pseudo-Matched Pairs in Remote Sensing Image-Text Retrieval
by: Ouyang, Pengxiang, et al.
Published: (2025)
by: Ouyang, Pengxiang, et al.
Published: (2025)
MCE: Towards a General Framework for Handling Missing Modalities under Imbalanced Missing Rates
by: Zhao, Binyu, et al.
Published: (2025)
by: Zhao, Binyu, et al.
Published: (2025)
InvZW: Invariant Feature Learning via Noise-Adversarial Training for Robust Image Zero-Watermarking
by: Tanvir, Abdullah All, et al.
Published: (2025)
by: Tanvir, Abdullah All, et al.
Published: (2025)
Principled Multimodal Representation Learning
by: Liu, Xiaohao, et al.
Published: (2025)
by: Liu, Xiaohao, et al.
Published: (2025)
Detecting Content Rating Violations in Android Applications: A Vision-Language Approach
by: Denipitiyage, D., et al.
Published: (2025)
by: Denipitiyage, D., et al.
Published: (2025)
Dual-Granularity Cross-Modal Identity Association for Weakly-Supervised Text-to-Person Image Matching
by: Zhang, Yafei, et al.
Published: (2025)
by: Zhang, Yafei, et al.
Published: (2025)
GroMo: Plant Growth Modeling with Multiview Images
by: Bhatt, Ruchi, et al.
Published: (2025)
by: Bhatt, Ruchi, et al.
Published: (2025)
From Pixels to Feelings: Aligning MLLMs with Human Cognitive Perception of Images
by: Chen, Yiming, et al.
Published: (2025)
by: Chen, Yiming, et al.
Published: (2025)
VIVAT: Virtuous Improving VAE Training through Artifact Mitigation
by: Novitskiy, Lev, et al.
Published: (2025)
by: Novitskiy, Lev, et al.
Published: (2025)
Similar Items
-
GLENDA: Gynecologic Laparoscopy Endometriosis Dataset
by: Leibetseder, Andreas, et al.
Published: (2025) -
diveXplore 6.0: ITEC's Interactive Video Exploration System at VBS 2022
by: Leibetseder, Andreas, et al.
Published: (2025) -
Identifying Surgical Instruments in Laparoscopy Using Deep Learning Instance Segmentation
by: Kletz, Sabrina, et al.
Published: (2025) -
lifeXplore at the Lifelog Search Challenge 2020
by: Leibetseder, Andreas, et al.
Published: (2025) -
Less is More - diveXplore 5.0 at VBS 2021
by: Leibetseder, Andreas, et al.
Published: (2025)