AdaViT: Adaptive Vision Transformer for Flexible Pretrain and Finetune with Variable 3D Medical Image Modalities
Fuente:
arXiv
Saved in:
| Main Authors: | Das, Badhan Kumar, Zhao, Gengyan, Liu, Han, Re, Thomas J., Comaniciu, Dorin, Gibson, Eli, Maier, Andreas |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Self Pre-training with Adaptive Mask Autoencoders for Variable-Contrast 3D Medical Imaging
by: Das, Badhan Kumar, et al.
Published: (2025)
by: Das, Badhan Kumar, et al.
Published: (2025)
VIViT: Variable-Input Vision Transformer Framework for 3D MR Image Segmentation
by: Das, Badhan Kumar, et al.
Published: (2025)
by: Das, Badhan Kumar, et al.
Published: (2025)
Multi-Plane Vision Transformer for Hemorrhage Classification Using Axial and Sagittal MRI Data
by: Das, Badhan Kumar, et al.
Published: (2025)
by: Das, Badhan Kumar, et al.
Published: (2025)
Invited Paper: BitMedViT: Ternary-Quantized Vision Transformer for Medical AI Assistants on the Edge
by: Walczak, Mikolaj, et al.
Published: (2025)
by: Walczak, Mikolaj, et al.
Published: (2025)
General-Purpose vs. Domain-Adapted Large Language Models for Extraction of Structured Data from Chest Radiology Reports
by: Dhanaliwala, Ali H., et al.
Published: (2023)
by: Dhanaliwala, Ali H., et al.
Published: (2023)
Adaptive-avg-pooling based Attention Vision Transformer for Face Anti-spoofing
by: Yang, Jichen, et al.
Published: (2024)
by: Yang, Jichen, et al.
Published: (2024)
FEFormer: Frequency-enhanced Vision Transformer for Generic Knowledge Extraction and Adaptive Feature Fusion in Volumetric Medical Image Segmentation
by: Yang, Jin, et al.
Published: (2026)
by: Yang, Jin, et al.
Published: (2026)
Towards a vision foundation model for comprehensive assessment of Cardiac MRI
by: Jacob, Athira J, et al.
Published: (2024)
by: Jacob, Athira J, et al.
Published: (2024)
3D-EffiViTCaps: 3D Efficient Vision Transformer with Capsule for Medical Image Segmentation
by: Gan, Dongwei, et al.
Published: (2024)
by: Gan, Dongwei, et al.
Published: (2024)
ME-ViT: A Single-Load Memory-Efficient FPGA Accelerator for Vision Transformers
by: Marino, Kyle, et al.
Published: (2024)
by: Marino, Kyle, et al.
Published: (2024)
ModAn-MulSupCon: Modality-and Anatomy-Aware Multi-Label Supervised Contrastive Pretraining for Medical Imaging
by: Takaya, Eichi, et al.
Published: (2025)
by: Takaya, Eichi, et al.
Published: (2025)
Learning to Optimize Radiotherapy Plans via Fluence Maps Diffusion Model Generation and LSTM-based Optimization
by: Poles, Isabella, et al.
Published: (2026)
by: Poles, Isabella, et al.
Published: (2026)
MediViSTA: Medical Video Segmentation via Temporal Fusion SAM Adaptation for Echocardiography
by: Kim, Sekeun, et al.
Published: (2023)
by: Kim, Sekeun, et al.
Published: (2023)
Variable Resolution Pixel Quantization for Low Power Machine Vision Application on Edge
by: Deb, Senorita, et al.
Published: (2024)
by: Deb, Senorita, et al.
Published: (2024)
ViTGAN: Training GANs with Vision Transformers
by: Lee, Kwonjoon, et al.
Published: (2021)
by: Lee, Kwonjoon, et al.
Published: (2021)
CIMGEN: Controlled Image Manipulation by Finetuning Pretrained Generative Models on Limited Data
by: Gudavalli, Chandrakanth, et al.
Published: (2024)
by: Gudavalli, Chandrakanth, et al.
Published: (2024)
Flexible Variable-Rate Image Feature Compression for Edge-Cloud Systems
by: Hossain, Md Adnan Faisal, et al.
Published: (2024)
by: Hossain, Md Adnan Faisal, et al.
Published: (2024)
NT-ViT: Neural Transcoding Vision Transformers for EEG-to-fMRI Synthesis
by: Lanzino, Romeo, et al.
Published: (2024)
by: Lanzino, Romeo, et al.
Published: (2024)
Enhancing Cross-Modality Synthesis: Subvolume Merging for MRI-to-CT Conversion
by: Fan, Fuxin, et al.
Published: (2024)
by: Fan, Fuxin, et al.
Published: (2024)
SaViD: Spectravista Aesthetic Vision Integration for Robust and Discerning 3D Object Detection in Challenging Environments
by: Dam, Tanmoy, et al.
Published: (2025)
by: Dam, Tanmoy, et al.
Published: (2025)
Motion-Adaptive Inference for Flexible Learned B-Frame Compression
by: Yilmaz, M. Akin, et al.
Published: (2024)
by: Yilmaz, M. Akin, et al.
Published: (2024)
Introduction of Medical Imaging Modalities
by: Islam, S. K. M Shadekul, et al.
Published: (2023)
by: Islam, S. K. M Shadekul, et al.
Published: (2023)
DIOR-ViT: Differential Ordinal Learning Vision Transformer for Cancer Classification in Pathology Images
by: Lee, Ju Cheon, et al.
Published: (2024)
by: Lee, Ju Cheon, et al.
Published: (2024)
SW-ViT: A Spatio-Temporal Vision Transformer Network with Post Denoiser for Sequential Multi-Push Ultrasound Shear Wave Elastography
by: Akash, Ahsan Habib, et al.
Published: (2025)
by: Akash, Ahsan Habib, et al.
Published: (2025)
Versatile Volumetric Medical Image Coding for Human-Machine Vision
by: Chen, Jietao, et al.
Published: (2024)
by: Chen, Jietao, et al.
Published: (2024)
CNN-ViT Hybrid for Pneumonia Detection: Theory and Empiric on Limited Data without Pretraining
by: Basnet, Prashant Singh, et al.
Published: (2025)
by: Basnet, Prashant Singh, et al.
Published: (2025)
Optimizing Prompt Strategies for SAM: Advancing lesion Segmentation Across Diverse Medical Imaging Modalities
by: Wang, Yuli, et al.
Published: (2024)
by: Wang, Yuli, et al.
Published: (2024)
A Simple and Robust Framework for Cross-Modality Medical Image Segmentation applied to Vision Transformers
by: Bastico, Matteo, et al.
Published: (2023)
by: Bastico, Matteo, et al.
Published: (2023)
PEANO-ViT: Power-Efficient Approximations of Non-Linearities in Vision Transformers
by: Sadeghi, Mohammad Erfan, et al.
Published: (2024)
by: Sadeghi, Mohammad Erfan, et al.
Published: (2024)
HistoViT: Vision Transformer for Accurate and Scalable Histopathological Cancer Diagnosis
by: Ahmed, Faisal
Published: (2025)
by: Ahmed, Faisal
Published: (2025)
Equi-ViT: Rotational Equivariant Vision Transformer for Robust Histopathology Analysis
by: Chen, Fuyao, et al.
Published: (2026)
by: Chen, Fuyao, et al.
Published: (2026)
Diffusion Models are Robust Pretrainers
by: Yagoda, Mika, et al.
Published: (2025)
by: Yagoda, Mika, et al.
Published: (2025)
ConvShareViT: Enhancing Vision Transformers with Convolutional Attention Mechanisms for Free-Space Optical Accelerators
by: Ibadulla, Riad, et al.
Published: (2025)
by: Ibadulla, Riad, et al.
Published: (2025)
Variable Rate Learned Wavelet Video Coding using Temporal Layer Adaptivity
by: Meyer, Anna, et al.
Published: (2024)
by: Meyer, Anna, et al.
Published: (2024)
AdaBM: On-the-Fly Adaptive Bit Mapping for Image Super-Resolution
by: Hong, Cheeun, et al.
Published: (2024)
by: Hong, Cheeun, et al.
Published: (2024)
ELFATT: Efficient Linear Fast Attention for Vision Transformers
by: Wu, Chong, et al.
Published: (2025)
by: Wu, Chong, et al.
Published: (2025)
AdaMSS: Adaptive Multi-Modality Segmentation-to-Survival Learning for Survival Outcome Prediction from PET/CT Images
by: Meng, Mingyuan, et al.
Published: (2023)
by: Meng, Mingyuan, et al.
Published: (2023)
Image Provenance Analysis via Graph Encoding with Vision Transformer
by: Zhang, Keyang, et al.
Published: (2024)
by: Zhang, Keyang, et al.
Published: (2024)
Entropy Coding for Non-Rectangular Transform Blocks using Partitioned DCT Dictionaries for AV1
by: Das, Priyanka, et al.
Published: (2025)
by: Das, Priyanka, et al.
Published: (2025)
UNet-3D with Adaptive TverskyCE Loss for Pancreas Medical Image Segmentation
by: Zhang, Xubei, et al.
Published: (2025)
by: Zhang, Xubei, et al.
Published: (2025)
Similar Items
-
Self Pre-training with Adaptive Mask Autoencoders for Variable-Contrast 3D Medical Imaging
by: Das, Badhan Kumar, et al.
Published: (2025) -
VIViT: Variable-Input Vision Transformer Framework for 3D MR Image Segmentation
by: Das, Badhan Kumar, et al.
Published: (2025) -
Multi-Plane Vision Transformer for Hemorrhage Classification Using Axial and Sagittal MRI Data
by: Das, Badhan Kumar, et al.
Published: (2025) -
Invited Paper: BitMedViT: Ternary-Quantized Vision Transformer for Medical AI Assistants on the Edge
by: Walczak, Mikolaj, et al.
Published: (2025) -
General-Purpose vs. Domain-Adapted Large Language Models for Extraction of Structured Data from Chest Radiology Reports
by: Dhanaliwala, Ali H., et al.
Published: (2023)