SimGen: A Diffusion-Based Framework for Simultaneous Surgical Image and Segmentation Mask Generation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Bhat, Aditya, Bose, Rupak, Nwoye, Chinedu Innocent, Padoy, Nicolas |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
CoSimGen: Controllable Diffusion Model for Simultaneous Image and Mask Generation
von: Bose, Rupak, et al.
Veröffentlicht: (2025)
von: Bose, Rupak, et al.
Veröffentlicht: (2025)
SurgiTrack: Fine-Grained Multi-Class Multi-Tool Tracking in Surgical Videos
von: Nwoye, Chinedu Innocent, et al.
Veröffentlicht: (2024)
von: Nwoye, Chinedu Innocent, et al.
Veröffentlicht: (2024)
Surgical Text-to-Image Generation
von: Nwoye, Chinedu Innocent, et al.
Veröffentlicht: (2024)
von: Nwoye, Chinedu Innocent, et al.
Veröffentlicht: (2024)
Feature Mixing Approach for Detecting Intraoperative Adverse Events in Laparoscopic Roux-en-Y Gastric Bypass Surgery
von: Bose, Rupak, et al.
Veröffentlicht: (2025)
von: Bose, Rupak, et al.
Veröffentlicht: (2025)
State-Change Learning for Prediction of Future Events in Endoscopic Videos
von: Sharma, Saurav, et al.
Veröffentlicht: (2025)
von: Sharma, Saurav, et al.
Veröffentlicht: (2025)
CholecTrack20: A Multi-Perspective Tracking Dataset for Surgical Tools
von: Nwoye, Chinedu Innocent, et al.
Veröffentlicht: (2023)
von: Nwoye, Chinedu Innocent, et al.
Veröffentlicht: (2023)
SimGen: Simulator-conditioned Driving Scene Generation
von: Zhou, Yunsong, et al.
Veröffentlicht: (2024)
von: Zhou, Yunsong, et al.
Veröffentlicht: (2024)
CycleSAM: Few-Shot Surgical Scene Segmentation with Cycle- and Scene-Consistent Feature Matching
von: Murali, Aditya, et al.
Veröffentlicht: (2024)
von: Murali, Aditya, et al.
Veröffentlicht: (2024)
SurgLaVi: Large-Scale Hierarchical Dataset for Surgical Vision-Language Representation Learning
von: Perez, Alejandra, et al.
Veröffentlicht: (2025)
von: Perez, Alejandra, et al.
Veröffentlicht: (2025)
Optimizing Latent Graph Representations of Surgical Scenes for Zero-Shot Domain Transfer
von: Satyanaik, Siddhant, et al.
Veröffentlicht: (2024)
von: Satyanaik, Siddhant, et al.
Veröffentlicht: (2024)
Jumpstarting Surgical Computer Vision
von: Alapatt, Deepak, et al.
Veröffentlicht: (2023)
von: Alapatt, Deepak, et al.
Veröffentlicht: (2023)
fine-CLIP: Enhancing Zero-Shot Fine-Grained Surgical Action Recognition with Vision-Language Models
von: Sharma, Saurav, et al.
Veröffentlicht: (2025)
von: Sharma, Saurav, et al.
Veröffentlicht: (2025)
DiffAtlas: GenAI-fying Atlas Segmentation via Image-Mask Diffusion
von: Zhang, Hantao, et al.
Veröffentlicht: (2025)
von: Zhang, Hantao, et al.
Veröffentlicht: (2025)
Overcoming Dimensional Collapse in Self-supervised Contrastive Learning for Medical Image Segmentation
von: Hassanpour, Jamshid, et al.
Veröffentlicht: (2024)
von: Hassanpour, Jamshid, et al.
Veröffentlicht: (2024)
On the Role of Depth in Surgical Vision Foundation Models: An Empirical Study of RGB-D Pre-training
von: Han, John J., et al.
Veröffentlicht: (2026)
von: Han, John J., et al.
Veröffentlicht: (2026)
SUREON: A Benchmark and Vision-Language-Model for Surgical Reasoning
von: Perez, Alejandra, et al.
Veröffentlicht: (2026)
von: Perez, Alejandra, et al.
Veröffentlicht: (2026)
GenMask: Adapting DiT for Segmentation via Direct Mask Generation
von: Yang, Yuhuan, et al.
Veröffentlicht: (2026)
von: Yang, Yuhuan, et al.
Veröffentlicht: (2026)
MATIS: Masked-Attention Transformers for Surgical Instrument Segmentation
von: Ayobi, Nicolás, et al.
Veröffentlicht: (2023)
von: Ayobi, Nicolás, et al.
Veröffentlicht: (2023)
The Endoscapes Dataset for Surgical Scene Segmentation, Object Detection, and Critical View of Safety Assessment: Official Splits and Benchmark
von: Murali, Aditya, et al.
Veröffentlicht: (2023)
von: Murali, Aditya, et al.
Veröffentlicht: (2023)
SkinDualGen: Prompt-Driven Diffusion for Simultaneous Image-Mask Generation in Skin Lesions
von: Xu, Zhaobin
Veröffentlicht: (2025)
von: Xu, Zhaobin
Veröffentlicht: (2025)
DExTeR: Weakly Semi-Supervised Object Detection with Class and Instance Experts for Medical Imaging
von: Meyer, Adrien, et al.
Veröffentlicht: (2026)
von: Meyer, Adrien, et al.
Veröffentlicht: (2026)
Procedure-Aware Surgical Video-language Pretraining with Hierarchical Knowledge Augmentation
von: Yuan, Kun, et al.
Veröffentlicht: (2024)
von: Yuan, Kun, et al.
Veröffentlicht: (2024)
UltraSam: A Foundation Model for Ultrasound using Large Open-Access Segmentation Datasets
von: Meyer, Adrien, et al.
Veröffentlicht: (2024)
von: Meyer, Adrien, et al.
Veröffentlicht: (2024)
Adaptation of Multi-modal Representation Models for Multi-task Surgical Computer Vision
von: Walimbe, Soham, et al.
Veröffentlicht: (2025)
von: Walimbe, Soham, et al.
Veröffentlicht: (2025)
HecVL: Hierarchical Video-Language Pretraining for Zero-shot Surgical Phase Recognition
von: Yuan, Kun, et al.
Veröffentlicht: (2024)
von: Yuan, Kun, et al.
Veröffentlicht: (2024)
EpiMask: Leveraging Epipolar Distance Based Masks in Cross-Attention for Satellite Image Matching
von: Deshmukh, Rahul, et al.
Veröffentlicht: (2026)
von: Deshmukh, Rahul, et al.
Veröffentlicht: (2026)
CliPPER: Contextual Video-Language Pretraining on Long-form Intraoperative Surgical Procedures for Event Recognition
von: Stilz, Florian, et al.
Veröffentlicht: (2026)
von: Stilz, Florian, et al.
Veröffentlicht: (2026)
Multi-view Video-Pose Pretraining for Operating Room Surgical Activity Recognition
von: Hamoud, Idris, et al.
Veröffentlicht: (2025)
von: Hamoud, Idris, et al.
Veröffentlicht: (2025)
GenSelfDiff-HIS: Generative Self-Supervision Using Diffusion for Histopathological Image Segmentation
von: Purma, Vishnuvardhan, et al.
Veröffentlicht: (2023)
von: Purma, Vishnuvardhan, et al.
Veröffentlicht: (2023)
End-to-End Learning of Multi-Organ Implicit Surfaces from 3D Medical Imaging Data
von: Zarin, Farahdiba, et al.
Veröffentlicht: (2025)
von: Zarin, Farahdiba, et al.
Veröffentlicht: (2025)
Endoshare: A Publicly Available, Surgeons-Friendly Solution to De-Identify and Manage Surgical Videos
von: Arboit, Lorenzo, et al.
Veröffentlicht: (2025)
von: Arboit, Lorenzo, et al.
Veröffentlicht: (2025)
Text-driven Adaptation of Foundation Models for Few-shot Surgical Workflow Analysis
von: Chen, Tingxuan, et al.
Veröffentlicht: (2025)
von: Chen, Tingxuan, et al.
Veröffentlicht: (2025)
Accelerating Inference of Masked Image Generators via Reinforcement Learning
von: Subbaraman, Pranav, et al.
Veröffentlicht: (2025)
von: Subbaraman, Pranav, et al.
Veröffentlicht: (2025)
Where It Moves, It Matters: Referring Surgical Instrument Segmentation via Motion
von: Wei, Meng, et al.
Veröffentlicht: (2026)
von: Wei, Meng, et al.
Veröffentlicht: (2026)
Diffusion-Guided Mask-Consistent Paired Mixing for Endoscopic Image Segmentation
von: Jie, Pengyu, et al.
Veröffentlicht: (2025)
von: Jie, Pengyu, et al.
Veröffentlicht: (2025)
Soil Image Segmentation Based on Mask R-CNN
von: Chen, Yida, et al.
Veröffentlicht: (2023)
von: Chen, Yida, et al.
Veröffentlicht: (2023)
GenTron: Diffusion Transformers for Image and Video Generation
von: Chen, Shoufa, et al.
Veröffentlicht: (2023)
von: Chen, Shoufa, et al.
Veröffentlicht: (2023)
SelfPose3d: Self-Supervised Multi-Person Multi-View 3d Pose Estimation
von: Srivastav, Vinkle, et al.
Veröffentlicht: (2024)
von: Srivastav, Vinkle, et al.
Veröffentlicht: (2024)
S4M: 4-points to Segment Anything
von: Meyer, Adrien, et al.
Veröffentlicht: (2025)
von: Meyer, Adrien, et al.
Veröffentlicht: (2025)
Harnessing Diffusion-Generated Synthetic Images for Fair Image Classification
von: Basu, Abhipsa, et al.
Veröffentlicht: (2025)
von: Basu, Abhipsa, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
CoSimGen: Controllable Diffusion Model for Simultaneous Image and Mask Generation
von: Bose, Rupak, et al.
Veröffentlicht: (2025) -
SurgiTrack: Fine-Grained Multi-Class Multi-Tool Tracking in Surgical Videos
von: Nwoye, Chinedu Innocent, et al.
Veröffentlicht: (2024) -
Surgical Text-to-Image Generation
von: Nwoye, Chinedu Innocent, et al.
Veröffentlicht: (2024) -
Feature Mixing Approach for Detecting Intraoperative Adverse Events in Laparoscopic Roux-en-Y Gastric Bypass Surgery
von: Bose, Rupak, et al.
Veröffentlicht: (2025) -
State-Change Learning for Prediction of Future Events in Endoscopic Videos
von: Sharma, Saurav, et al.
Veröffentlicht: (2025)