CoStoDet-DDPM: Collaborative Training of Stochastic and Deterministic Models Improves Surgical Workflow Anticipation and Recognition
Fuente:
arXiv
Saved in:
| Main Authors: | Yang, Kaixiang, Li, Xin, Li, Qiang, Wang, Zhiwei |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
DACAT: Dual-stream Adaptive Clip-aware Time Modeling for Robust Online Surgical Phase Recognition
by: Yang, Kaixiang, et al.
Published: (2024)
by: Yang, Kaixiang, et al.
Published: (2024)
Bidirectional Mammogram View Translation with Column-Aware and Implicit 3D Conditional Diffusion
by: Li, Xin, et al.
Published: (2025)
by: Li, Xin, et al.
Published: (2025)
Joint Holistic and Lesion Controllable Mammogram Synthesis via Gated Conditional Diffusion Model
by: Li, Xin, et al.
Published: (2025)
by: Li, Xin, et al.
Published: (2025)
SWAG: Long-term Surgical Workflow Prediction with Generative-based Anticipation
by: Boels, Maxence, et al.
Published: (2024)
by: Boels, Maxence, et al.
Published: (2024)
Adaptive Graph Learning from Spatial Information for Surgical Workflow Anticipation
by: Zhang, Francis Xiatian, et al.
Published: (2024)
by: Zhang, Francis Xiatian, et al.
Published: (2024)
SuPRA: Surgical Phase Recognition and Anticipation for Intra-Operative Planning
by: Boels, Maxence, et al.
Published: (2024)
by: Boels, Maxence, et al.
Published: (2024)
GCD-DDPM: A Generative Change Detection Model Based on Difference-Feature Guided DDPM
by: Wen, Yihan, et al.
Published: (2023)
by: Wen, Yihan, et al.
Published: (2023)
From Recognition to Prediction: Leveraging Sequence Reasoning for Action Anticipation
by: Liu, Xin, et al.
Published: (2024)
by: Liu, Xin, et al.
Published: (2024)
LoopVLA: Learning Sufficiency in Recurrent Refinement for Vision-Language-Action Models
by: Shen, Boyang, et al.
Published: (2026)
by: Shen, Boyang, et al.
Published: (2026)
Surgical Workflow Recognition and Blocking Effectiveness Detection in Laparoscopic Liver Resections with Pringle Maneuver
by: Guo, Diandian, et al.
Published: (2024)
by: Guo, Diandian, et al.
Published: (2024)
DCCS-Det: Directional Context and Cross-Scale-Aware Detector for Infrared Small Target
by: Li, Shuying, et al.
Published: (2026)
by: Li, Shuying, et al.
Published: (2026)
FIA-Edit: Frequency-Interactive Attention for Efficient and High-Fidelity Inversion-Free Text-Guided Image Editing
by: Yang, Kaixiang, et al.
Published: (2025)
by: Yang, Kaixiang, et al.
Published: (2025)
Decoupling Feature Representations of Ego and Other Modalities for Incomplete Multi-modal Brain Tumor Segmentation
by: Yang, Kaixiang, et al.
Published: (2024)
by: Yang, Kaixiang, et al.
Published: (2024)
CoPESD: A Multi-Level Surgical Motion Dataset for Training Large Vision-Language Models to Co-Pilot Endoscopic Submucosal Dissection
by: Wang, Guankun, et al.
Published: (2024)
by: Wang, Guankun, et al.
Published: (2024)
MatchDet: A Collaborative Framework for Image Matching and Object Detection
by: Lai, Jinxiang, et al.
Published: (2023)
by: Lai, Jinxiang, et al.
Published: (2023)
DSTED: Decoupling Temporal Stabilization and Discriminative Enhancement for Surgical Workflow Recognition
by: Chen, Yueyao, et al.
Published: (2025)
by: Chen, Yueyao, et al.
Published: (2025)
DetDiffusion: Synergizing Generative and Perceptive Models for Enhanced Data Generation and Perception
by: Wang, Yibo, et al.
Published: (2024)
by: Wang, Yibo, et al.
Published: (2024)
Multimodal Graph Representation Learning for Robust Surgical Workflow Recognition with Adversarial Feature Disentanglement
by: Bai, Long, et al.
Published: (2025)
by: Bai, Long, et al.
Published: (2025)
SpirDet: Towards Efficient, Accurate and Lightweight Infrared Small Target Detector
by: Mao, Qianchen, et al.
Published: (2024)
by: Mao, Qianchen, et al.
Published: (2024)
DST-Det: Simple Dynamic Self-Training for Open-Vocabulary Object Detection
by: Xu, Shilin, et al.
Published: (2023)
by: Xu, Shilin, et al.
Published: (2023)
Multimodal Large Models Are Effective Action Anticipators
by: Wang, Binglu, et al.
Published: (2025)
by: Wang, Binglu, et al.
Published: (2025)
RemDet: Rethinking Efficient Model Design for UAV Object Detection
by: Li, Chen, et al.
Published: (2024)
by: Li, Chen, et al.
Published: (2024)
CRASH: Crash Recognition and Anticipation System Harnessing with Context-Aware and Temporal Focus Attentions
by: Liao, Haicheng, et al.
Published: (2024)
by: Liao, Haicheng, et al.
Published: (2024)
Multi-Conditioned Denoising Diffusion Probabilistic Model (mDDPM) for Medical Image Synthesis
by: Krishna, Arjun, et al.
Published: (2024)
by: Krishna, Arjun, et al.
Published: (2024)
Stochastic Layer-Wise Shuffle for Improving Vision Mamba Training
by: Huang, Zizheng, et al.
Published: (2024)
by: Huang, Zizheng, et al.
Published: (2024)
HazyDet: Open-Source Benchmark for Drone-View Object Detection with Depth-Cues in Hazy Scenes
by: Feng, Changfeng, et al.
Published: (2024)
by: Feng, Changfeng, et al.
Published: (2024)
CoLLM-NAS: Collaborative Large Language Models for Efficient Knowledge-Guided Neural Architecture Search
by: Li, Zhe, et al.
Published: (2025)
by: Li, Zhe, et al.
Published: (2025)
DDPM-MoCo: Advancing Industrial Surface Defect Generation and Detection with Generative and Contrastive Learning
by: He, Yangfan, et al.
Published: (2024)
by: He, Yangfan, et al.
Published: (2024)
FreeVPS: Repurposing Training-Free SAM2 for Generalizable Video Polyp Segmentation
by: Hu, Qiang, et al.
Published: (2025)
by: Hu, Qiang, et al.
Published: (2025)
IBoxCLA: Towards Robust Box-supervised Segmentation of Polyp via Improved Box-dice and Contrastive Latent-anchors
by: Wang, Zhiwei, et al.
Published: (2023)
by: Wang, Zhiwei, et al.
Published: (2023)
RemoteDet-Mamba: A Hybrid Mamba-CNN Network for Multi-modal Object Detection in Remote Sensing Images
by: Ren, Kejun, et al.
Published: (2024)
by: Ren, Kejun, et al.
Published: (2024)
Surgformer: Surgical Transformer with Hierarchical Temporal Attention for Surgical Phase Recognition
by: Yang, Shu, et al.
Published: (2024)
by: Yang, Shu, et al.
Published: (2024)
SCA: Improve Semantic Consistent in Unrestricted Adversarial Attacks via DDPM Inversion
by: Pan, Zihao, et al.
Published: (2024)
by: Pan, Zihao, et al.
Published: (2024)
Spatial Information Bottleneck for Interpretable Visual Recognition
by: Shu, Kaixiang, et al.
Published: (2025)
by: Shu, Kaixiang, et al.
Published: (2025)
Gated Temporal Diffusion for Stochastic Long-Term Dense Anticipation
by: Zatsarynna, Olga, et al.
Published: (2024)
by: Zatsarynna, Olga, et al.
Published: (2024)
ConsistencyDet: A Few-step Denoising Framework for Object Detection Using the Consistency Model
by: Jiang, Lifan, et al.
Published: (2024)
by: Jiang, Lifan, et al.
Published: (2024)
Deterministic-to-Stochastic Diverse Latent Feature Mapping for Human Motion Synthesis
by: Hua, Yu, et al.
Published: (2025)
by: Hua, Yu, et al.
Published: (2025)
LMM-Det: Make Large Multimodal Models Excel in Object Detection
by: Li, Jincheng, et al.
Published: (2025)
by: Li, Jincheng, et al.
Published: (2025)
PromptDet: A Lightweight 3D Object Detection Framework with LiDAR Prompts
by: Guo, Kun, et al.
Published: (2024)
by: Guo, Kun, et al.
Published: (2024)
A Training-Free Framework for Video License Plate Tracking and Recognition with Only One-Shot
by: Ding, Haoxuan, et al.
Published: (2024)
by: Ding, Haoxuan, et al.
Published: (2024)
Similar Items
-
DACAT: Dual-stream Adaptive Clip-aware Time Modeling for Robust Online Surgical Phase Recognition
by: Yang, Kaixiang, et al.
Published: (2024) -
Bidirectional Mammogram View Translation with Column-Aware and Implicit 3D Conditional Diffusion
by: Li, Xin, et al.
Published: (2025) -
Joint Holistic and Lesion Controllable Mammogram Synthesis via Gated Conditional Diffusion Model
by: Li, Xin, et al.
Published: (2025) -
SWAG: Long-term Surgical Workflow Prediction with Generative-based Anticipation
by: Boels, Maxence, et al.
Published: (2024) -
Adaptive Graph Learning from Spatial Information for Surgical Workflow Anticipation
by: Zhang, Francis Xiatian, et al.
Published: (2024)