Video-SwinUNet: Spatio-temporal Deep Learning Framework for VFSS Instance Segmentation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zeng, Chengxi, Yang, Xinyu, Smithard, David, Mirmehdi, Majid, Gambaruto, Alberto M, Burghardt, Tilo |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2023
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Tuning Vision Foundation Model via Test-Time Prompt-Guided Training for VFSS Segmentations
von: Zeng, Chengxi, et al.
Veröffentlicht: (2025)
von: Zeng, Chengxi, et al.
Veröffentlicht: (2025)
Agglomerating Large Vision Encoders via Distillation for VFSS Segmentation
von: Zeng, Chengxi, et al.
Veröffentlicht: (2025)
von: Zeng, Chengxi, et al.
Veröffentlicht: (2025)
RBF-PINN: Non-Fourier Positional Embedding in Physics-Informed Neural Networks
von: Zeng, Chengxi, et al.
Veröffentlicht: (2024)
von: Zeng, Chengxi, et al.
Veröffentlicht: (2024)
Feature Mapping in Physics-Informed Neural Networks (PINNs)
von: Zeng, Chengxi, et al.
Veröffentlicht: (2024)
von: Zeng, Chengxi, et al.
Veröffentlicht: (2024)
Visual-textual Dermatoglyphic Animal Biometrics: A First Case Study on Panthera tigris
von: Li, Wenshuo, et al.
Veröffentlicht: (2025)
von: Li, Wenshuo, et al.
Veröffentlicht: (2025)
ChimpVLM: Ethogram-Enhanced Chimpanzee Behaviour Recognition
von: Brookes, Otto, et al.
Veröffentlicht: (2024)
von: Brookes, Otto, et al.
Veröffentlicht: (2024)
Deep in the Jungle: Towards Automating Chimpanzee Population Estimation
von: Raynes, Tom, et al.
Veröffentlicht: (2026)
von: Raynes, Tom, et al.
Veröffentlicht: (2026)
Improving Weakly-supervised Video Instance Segmentation by Leveraging Spatio-temporal Consistency
von: Arefi, Farnoosh, et al.
Veröffentlicht: (2024)
von: Arefi, Farnoosh, et al.
Veröffentlicht: (2024)
Co-STAR: Collaborative Curriculum Self-Training with Adaptive Regularization for Source-Free Video Domain Adaptation
von: Dadashzadeh, Amirhossein, et al.
Veröffentlicht: (2025)
von: Dadashzadeh, Amirhossein, et al.
Veröffentlicht: (2025)
SwinTextUNet: Integrating CLIP-Based Text Guidance into Swin Transformer U-Nets for Medical Image Segmentation
von: Yeafi, Ashfak, et al.
Veröffentlicht: (2026)
von: Yeafi, Ashfak, et al.
Veröffentlicht: (2026)
DualSwinFusionSeg: Multimodal Martian Landslide Segmentation via Dual Swin Transformer with Multi-Scale Fusion and UNet++
von: Kabir, Shahriar, et al.
Veröffentlicht: (2026)
von: Kabir, Shahriar, et al.
Veröffentlicht: (2026)
Long-tailed Species Recognition in the NACTI Wildlife Dataset
von: Liu, Zehua, et al.
Veröffentlicht: (2025)
von: Liu, Zehua, et al.
Veröffentlicht: (2025)
Ground-based image deconvolution with Swin Transformer UNet
von: Akhaury, Utsav, et al.
Veröffentlicht: (2024)
von: Akhaury, Utsav, et al.
Veröffentlicht: (2024)
Unsupervised View-Invariant Human Posture Representation
von: Sardari, Faegheh, et al.
Veröffentlicht: (2021)
von: Sardari, Faegheh, et al.
Veröffentlicht: (2021)
WildLive: Near Real-time Visual Wildlife Tracking onboard UAVs
von: Dat, Nguyen Ngoc, et al.
Veröffentlicht: (2025)
von: Dat, Nguyen Ngoc, et al.
Veröffentlicht: (2025)
Hierarchical Spatio-temporal Segmentation Network for Ejection Fraction Estimation in Echocardiography Videos
von: Wang, Dongfang, et al.
Veröffentlicht: (2025)
von: Wang, Dongfang, et al.
Veröffentlicht: (2025)
Adenocarcinoma Segmentation Using Pre-trained Swin-UNet with Parallel Cross-Attention for Multi-Domain Imaging
von: Qayyum, Abdul, et al.
Veröffentlicht: (2024)
von: Qayyum, Abdul, et al.
Veröffentlicht: (2024)
EfficientSAM3: Progressive Hierarchical Distillation for Video Concept Segmentation from SAM1, 2, and 3
von: Zeng, Chengxi, et al.
Veröffentlicht: (2025)
von: Zeng, Chengxi, et al.
Veröffentlicht: (2025)
Swin-UMamba: Mamba-based UNet with ImageNet-based pretraining
von: Liu, Jiarun, et al.
Veröffentlicht: (2024)
von: Liu, Jiarun, et al.
Veröffentlicht: (2024)
Image Segmentation via Variational Model Based Tailored UNet: A Deep Variational Framework
von: Qi, Kaili, et al.
Veröffentlicht: (2025)
von: Qi, Kaili, et al.
Veröffentlicht: (2025)
SVAG-Bench: A Large-Scale Benchmark for Multi-Instance Spatio-temporal Video Action Grounding
von: Hannan, Tanveer, et al.
Veröffentlicht: (2025)
von: Hannan, Tanveer, et al.
Veröffentlicht: (2025)
Unsupervised Cross-Domain 3D Human Pose Estimation via Pseudo-Label-Guided Global Transforms
von: Liu, Jingjing, et al.
Veröffentlicht: (2025)
von: Liu, Jingjing, et al.
Veröffentlicht: (2025)
Prediction of Thrombectomy Functional Outcomes using Multimodal Data
von: Samak, Zeynel A., et al.
Veröffentlicht: (2020)
von: Samak, Zeynel A., et al.
Veröffentlicht: (2020)
TranSOP: Transformer-based Multimodal Classification for Stroke Treatment Outcome Prediction
von: Samak, Zeynel A., et al.
Veröffentlicht: (2023)
von: Samak, Zeynel A., et al.
Veröffentlicht: (2023)
Automatic Prediction of Stroke Treatment Outcomes: Latest Advances and Perspectives
von: Samak, Zeynel A., et al.
Veröffentlicht: (2024)
von: Samak, Zeynel A., et al.
Veröffentlicht: (2024)
Skeleton-Snippet Contrastive Learning with Multiscale Feature Fusion for Action Localization
von: Cheng, Qiushuo, et al.
Veröffentlicht: (2025)
von: Cheng, Qiushuo, et al.
Veröffentlicht: (2025)
Text Embedded Swin-UMamba for DeepLesion Segmentation
von: Cheng, Ruida, et al.
Veröffentlicht: (2025)
von: Cheng, Ruida, et al.
Veröffentlicht: (2025)
Trajectory-guided Motion Perception for Facial Expression Quality Assessment in Neurological Disorders
von: Duan, Shuchao, et al.
Veröffentlicht: (2025)
von: Duan, Shuchao, et al.
Veröffentlicht: (2025)
SAM3-UNet: Simplified Adaptation of Segment Anything Model 3
von: Xiong, Xinyu, et al.
Veröffentlicht: (2025)
von: Xiong, Xinyu, et al.
Veröffentlicht: (2025)
OmniTransfer: All-in-one Framework for Spatio-temporal Video Transfer
von: Zhang, Pengze, et al.
Veröffentlicht: (2026)
von: Zhang, Pengze, et al.
Veröffentlicht: (2026)
SIS-Challenge: Event-based Spatio-temporal Instance Segmentation Challenge at the CVPR 2025 Event-based Vision Workshop
von: Hamann, Friedhelm, et al.
Veröffentlicht: (2025)
von: Hamann, Friedhelm, et al.
Veröffentlicht: (2025)
Training-Free Spatio-temporal Decoupled Reasoning Video Segmentation with Adaptive Object Memory
von: Zhu, Zhengtong, et al.
Veröffentlicht: (2026)
von: Zhu, Zhengtong, et al.
Veröffentlicht: (2026)
GCA-SUNet: A Gated Context-Aware Swin-UNet for Exemplar-Free Counting
von: Wu, Yuzhe, et al.
Veröffentlicht: (2024)
von: Wu, Yuzhe, et al.
Veröffentlicht: (2024)
A Temporal Modeling Framework for Video Pre-Training on Video Instance Segmentation
von: Zhong, Qing, et al.
Veröffentlicht: (2025)
von: Zhong, Qing, et al.
Veröffentlicht: (2025)
Segmenting Medical Images: From UNet to Res-UNet and nnUNet
von: Huang, Lina, et al.
Veröffentlicht: (2024)
von: Huang, Lina, et al.
Veröffentlicht: (2024)
BEVPredFormer: Spatio-temporal Attention for BEV Instance Prediction in Autonomous Driving
von: Antunes-García, Miguel, et al.
Veröffentlicht: (2026)
von: Antunes-García, Miguel, et al.
Veröffentlicht: (2026)
UVIS: Unsupervised Video Instance Segmentation
von: Huang, Shuaiyi, et al.
Veröffentlicht: (2024)
von: Huang, Shuaiyi, et al.
Veröffentlicht: (2024)
Hierarchical Visual Prompt Learning for Continual Video Instance Segmentation
von: Dong, Jiahua, et al.
Veröffentlicht: (2025)
von: Dong, Jiahua, et al.
Veröffentlicht: (2025)
Complete Instances Mining for Weakly Supervised Instance Segmentation
von: Li, Zecheng, et al.
Veröffentlicht: (2024)
von: Li, Zecheng, et al.
Veröffentlicht: (2024)
SCUNet++: Swin-UNet and CNN Bottleneck Hybrid Architecture with Multi-Fusion Dense Skip Connection for Pulmonary Embolism CT Image Segmentation
von: Chen, Yifei, et al.
Veröffentlicht: (2023)
von: Chen, Yifei, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
Tuning Vision Foundation Model via Test-Time Prompt-Guided Training for VFSS Segmentations
von: Zeng, Chengxi, et al.
Veröffentlicht: (2025) -
Agglomerating Large Vision Encoders via Distillation for VFSS Segmentation
von: Zeng, Chengxi, et al.
Veröffentlicht: (2025) -
RBF-PINN: Non-Fourier Positional Embedding in Physics-Informed Neural Networks
von: Zeng, Chengxi, et al.
Veröffentlicht: (2024) -
Feature Mapping in Physics-Informed Neural Networks (PINNs)
von: Zeng, Chengxi, et al.
Veröffentlicht: (2024) -
Visual-textual Dermatoglyphic Animal Biometrics: A First Case Study on Panthera tigris
von: Li, Wenshuo, et al.
Veröffentlicht: (2025)