Mobile U-ViT: Revisiting large kernel and U-shaped ViT for efficient medical image segmentation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Tang, Fenghe, Nian, Bingkun, Ding, Jianrui, Ma, Wenxin, Quan, Quan, Dong, Chengqi, Yang, Jie, Liu, Wei, Zhou, S. Kevin |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
CLAMP-ViT: Contrastive Data-Free Learning for Adaptive Post-Training Quantization of ViTs
von: Ramachandran, Akshat, et al.
Veröffentlicht: (2024)
von: Ramachandran, Akshat, et al.
Veröffentlicht: (2024)
U-RWKV: Lightweight medical image segmentation with direction-adaptive RWKV
von: Ye, Hongbo, et al.
Veröffentlicht: (2025)
von: Ye, Hongbo, et al.
Veröffentlicht: (2025)
ME-ViT: A Single-Load Memory-Efficient FPGA Accelerator for Vision Transformers
von: Marino, Kyle, et al.
Veröffentlicht: (2024)
von: Marino, Kyle, et al.
Veröffentlicht: (2024)
NT-ViT: Neural Transcoding Vision Transformers for EEG-to-fMRI Synthesis
von: Lanzino, Romeo, et al.
Veröffentlicht: (2024)
von: Lanzino, Romeo, et al.
Veröffentlicht: (2024)
PEANO-ViT: Power-Efficient Approximations of Non-Linearities in Vision Transformers
von: Sadeghi, Mohammad Erfan, et al.
Veröffentlicht: (2024)
von: Sadeghi, Mohammad Erfan, et al.
Veröffentlicht: (2024)
Stylizing ViT: Anatomy-Preserving Instance Style Transfer for Domain Generalization
von: Doerrich, Sebastian, et al.
Veröffentlicht: (2026)
von: Doerrich, Sebastian, et al.
Veröffentlicht: (2026)
Equi-ViT: Rotational Equivariant Vision Transformer for Robust Histopathology Analysis
von: Chen, Fuyao, et al.
Veröffentlicht: (2026)
von: Chen, Fuyao, et al.
Veröffentlicht: (2026)
DIOR-ViT: Differential Ordinal Learning Vision Transformer for Cancer Classification in Pathology Images
von: Lee, Ju Cheon, et al.
Veröffentlicht: (2024)
von: Lee, Ju Cheon, et al.
Veröffentlicht: (2024)
CNN-ViT Hybrid for Pneumonia Detection: Theory and Empiric on Limited Data without Pretraining
von: Basnet, Prashant Singh, et al.
Veröffentlicht: (2025)
von: Basnet, Prashant Singh, et al.
Veröffentlicht: (2025)
Retinal Malady Classification using AI: A novel ViT-SVM combination architecture
von: Jha, Shashwat, et al.
Veröffentlicht: (2026)
von: Jha, Shashwat, et al.
Veröffentlicht: (2026)
Efficient Breast and Ovarian Cancer Classification via ViT-Based Preprocessing and Transfer Learning
von: Rawat, Richa, et al.
Veröffentlicht: (2025)
von: Rawat, Richa, et al.
Veröffentlicht: (2025)
Unleashing Vision Foundation Models for Coronary Artery Segmentation: Parallel ViT-CNN Encoding and Variational Fusion
von: Dong, Caixia, et al.
Veröffentlicht: (2025)
von: Dong, Caixia, et al.
Veröffentlicht: (2025)
Diagnosing COVID-19 Severity from Chest X-Ray Images Using ViT and CNN Architectures
von: Lara, Luis, et al.
Veröffentlicht: (2025)
von: Lara, Luis, et al.
Veröffentlicht: (2025)
SAFE: a SAR Feature Extractor based on self-supervised learning and masked Siamese ViTs
von: Muzeau, Max, et al.
Veröffentlicht: (2024)
von: Muzeau, Max, et al.
Veröffentlicht: (2024)
When CNN Meet with ViT: Towards Semi-Supervised Learning for Multi-Class Medical Image Semantic Segmentation
von: Wang, Ziyang, et al.
Veröffentlicht: (2022)
von: Wang, Ziyang, et al.
Veröffentlicht: (2022)
SW-ViT: A Spatio-Temporal Vision Transformer Network with Post Denoiser for Sequential Multi-Push Ultrasound Shear Wave Elastography
von: Akash, Ahsan Habib, et al.
Veröffentlicht: (2025)
von: Akash, Ahsan Habib, et al.
Veröffentlicht: (2025)
Weak-Mamba-UNet: Visual Mamba Makes CNN and ViT Work Better for Scribble-based Medical Image Segmentation
von: Wang, Ziyang, et al.
Veröffentlicht: (2024)
von: Wang, Ziyang, et al.
Veröffentlicht: (2024)
MedSAM-CA: A CNN-Augmented ViT with Attention-Enhanced Multi-Scale Fusion for Medical Image Segmentation
von: Tian, Peiting, et al.
Veröffentlicht: (2025)
von: Tian, Peiting, et al.
Veröffentlicht: (2025)
ViT3D Alignment of LLaMA3: 3D Medical Image Report Generation
von: Li, Siyou, et al.
Veröffentlicht: (2024)
von: Li, Siyou, et al.
Veröffentlicht: (2024)
LUM-ViT: Learnable Under-sampling Mask Vision Transformer for Bandwidth Limited Optical Signal Acquisition
von: Liu, Lingfeng, et al.
Veröffentlicht: (2024)
von: Liu, Lingfeng, et al.
Veröffentlicht: (2024)
Boosting ViT-based MRI Reconstruction from the Perspectives of Frequency Modulation, Spatial Purification, and Scale Diversification
von: Meng, Yucong, et al.
Veröffentlicht: (2024)
von: Meng, Yucong, et al.
Veröffentlicht: (2024)
TransResNet: Integrating the Strengths of ViTs and CNNs for High Resolution Medical Image Segmentation via Feature Grafting
von: Sharif, Muhammad Hamza, et al.
Veröffentlicht: (2024)
von: Sharif, Muhammad Hamza, et al.
Veröffentlicht: (2024)
CMU-Net: A Strong ConvMixer-based Medical Ultrasound Image Segmentation Network
von: Tang, Fenghe, et al.
Veröffentlicht: (2022)
von: Tang, Fenghe, et al.
Veröffentlicht: (2022)
Privacy-Preserving Chest X-ray Report Generation via Multimodal Federated Learning with ViT and GPT-2
von: Hossain, Md. Zahid, et al.
Veröffentlicht: (2025)
von: Hossain, Md. Zahid, et al.
Veröffentlicht: (2025)
Volumetric medical image segmentation through dual self-distillation in U-shaped networks
von: Banerjee, Soumyanil, et al.
Veröffentlicht: (2023)
von: Banerjee, Soumyanil, et al.
Veröffentlicht: (2023)
AResNet-ViT: A Hybrid CNN-Transformer Network for Benign and Malignant Breast Nodule Classification in Ultrasound Images
von: Zhao, Xin, et al.
Veröffentlicht: (2024)
von: Zhao, Xin, et al.
Veröffentlicht: (2024)
Pneumonia Image Classification Based on Lightweight Mobile ViT Networks
von: Zhiqiang Zheng, et al.
Veröffentlicht: (2025)
von: Zhiqiang Zheng, et al.
Veröffentlicht: (2025)
ZACH-ViT: Regime-Dependent Inductive Bias in Compact Vision Transformers for Medical Imaging
von: Angelakis, Athanasios
Veröffentlicht: (2026)
von: Angelakis, Athanasios
Veröffentlicht: (2026)
Which images to label for few-shot medical landmark detection?
von: Quan, Quan, et al.
Veröffentlicht: (2021)
von: Quan, Quan, et al.
Veröffentlicht: (2021)
RepViT: Revisiting Mobile CNN From ViT Perspective
von: Wang, Ao, et al.
Veröffentlicht: (2023)
von: Wang, Ao, et al.
Veröffentlicht: (2023)
Are foundation models efficient for medical image segmentation?
von: Ferreira, Danielle, et al.
Veröffentlicht: (2023)
von: Ferreira, Danielle, et al.
Veröffentlicht: (2023)
Rich-U-Net: A medical image segmentation model for fusing spatial depth features and capturing minute structural details
von: Fang, Zhuoyi, et al.
Veröffentlicht: (2026)
von: Fang, Zhuoyi, et al.
Veröffentlicht: (2026)
DS-ViT: Dual-Stream Vision Transformer for Cross-Task Distillation in Alzheimer's Early Diagnosis
von: Chen, Ke, et al.
Veröffentlicht: (2024)
von: Chen, Ke, et al.
Veröffentlicht: (2024)
A hybrid approach for improving U-Net variants in medical image segmentation
von: Gupta, Aitik, et al.
Veröffentlicht: (2023)
von: Gupta, Aitik, et al.
Veröffentlicht: (2023)
ViT Registers and Fractal ViT
von: Chou, Jason Chuan-Chih, et al.
Veröffentlicht: (2026)
von: Chou, Jason Chuan-Chih, et al.
Veröffentlicht: (2026)
ViT-NeBLa: A Hybrid Vision Transformer and Neural Beer-Lambert Framework for Single-View 3D Reconstruction of Oral Anatomy from Panoramic Radiographs
von: Parida, Bikram Keshari, et al.
Veröffentlicht: (2025)
von: Parida, Bikram Keshari, et al.
Veröffentlicht: (2025)
Rethinking domain generalization in medical image segmentation: One image as one domain
von: Hong, Jin, et al.
Veröffentlicht: (2025)
von: Hong, Jin, et al.
Veröffentlicht: (2025)
TransUNext: towards a more advanced U-shaped framework for automatic vessel segmentation in the fundus image
von: Li, Xiang, et al.
Veröffentlicht: (2024)
von: Li, Xiang, et al.
Veröffentlicht: (2024)
MediViSTA: Medical Video Segmentation via Temporal Fusion SAM Adaptation for Echocardiography
von: Kim, Sekeun, et al.
Veröffentlicht: (2023)
von: Kim, Sekeun, et al.
Veröffentlicht: (2023)
High-resolution myelin-water fraction and quantitative relaxation mapping using 3D ViSTa-MR fingerprinting
von: Liao, Congyu, et al.
Veröffentlicht: (2023)
von: Liao, Congyu, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
CLAMP-ViT: Contrastive Data-Free Learning for Adaptive Post-Training Quantization of ViTs
von: Ramachandran, Akshat, et al.
Veröffentlicht: (2024) -
U-RWKV: Lightweight medical image segmentation with direction-adaptive RWKV
von: Ye, Hongbo, et al.
Veröffentlicht: (2025) -
ME-ViT: A Single-Load Memory-Efficient FPGA Accelerator for Vision Transformers
von: Marino, Kyle, et al.
Veröffentlicht: (2024) -
NT-ViT: Neural Transcoding Vision Transformers for EEG-to-fMRI Synthesis
von: Lanzino, Romeo, et al.
Veröffentlicht: (2024) -
PEANO-ViT: Power-Efficient Approximations of Non-Linearities in Vision Transformers
von: Sadeghi, Mohammad Erfan, et al.
Veröffentlicht: (2024)