Masks and Manuscripts: Advancing Medical Pre-training with End-to-End Masking and Narrative Structuring
Fuente:
arXiv
Salvato in:
| Autori principali: | Gowda, Shreyank N, Clifton, David A. |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Distribution-Based Masked Medical Vision-Language Model Using Structured Reports
di: Gowda, Shreyank N, et al.
Pubblicazione: (2025)
di: Gowda, Shreyank N, et al.
Pubblicazione: (2025)
CC-SAM: SAM with Cross-feature Attention and Context for Ultrasound Image Segmentation
di: Gowda, Shreyank N, et al.
Pubblicazione: (2024)
di: Gowda, Shreyank N, et al.
Pubblicazione: (2024)
FE-Adapter: Adapting Image-based Emotion Classifiers to Videos
di: Gowda, Shreyank N, et al.
Pubblicazione: (2024)
di: Gowda, Shreyank N, et al.
Pubblicazione: (2024)
Prototype-Enhanced Confidence Modeling for Cross-Modal Medical Image-Report Retrieval
di: Gowda, Shreyank N, et al.
Pubblicazione: (2025)
di: Gowda, Shreyank N, et al.
Pubblicazione: (2025)
MaskFuser: Masked Fusion of Joint Multi-Modal Tokenization for End-to-End Autonomous Driving
di: Duan, Yiqun, et al.
Pubblicazione: (2024)
di: Duan, Yiqun, et al.
Pubblicazione: (2024)
MaskedCLIP: Bridging the Masked and CLIP Space for Semi-Supervised Medical Vision-Language Pre-training
di: Zhu, Lei, et al.
Pubblicazione: (2025)
di: Zhu, Lei, et al.
Pubblicazione: (2025)
Telling Stories for Common Sense Zero-Shot Action Recognition
di: Gowda, Shreyank N, et al.
Pubblicazione: (2023)
di: Gowda, Shreyank N, et al.
Pubblicazione: (2023)
MLIP: Medical Language-Image Pre-training with Masked Local Representation Learning
di: Liu, Jiarun, et al.
Pubblicazione: (2024)
di: Liu, Jiarun, et al.
Pubblicazione: (2024)
Efficient and Explainable End-to-End Autonomous Driving via Masked Vision-Language-Action Diffusion
di: Zhang, Jiaru, et al.
Pubblicazione: (2026)
di: Zhang, Jiaru, et al.
Pubblicazione: (2026)
Reimagining Reality: A Comprehensive Survey of Video Inpainting Techniques
di: Gowda, Shreyank N, et al.
Pubblicazione: (2024)
di: Gowda, Shreyank N, et al.
Pubblicazione: (2024)
Emerging Property of Masked Token for Effective Pre-training
di: Choi, Hyesong, et al.
Pubblicazione: (2024)
di: Choi, Hyesong, et al.
Pubblicazione: (2024)
Efficient Vision-Language Pre-training by Cluster Masking
di: Wei, Zihao, et al.
Pubblicazione: (2024)
di: Wei, Zihao, et al.
Pubblicazione: (2024)
Dataset Ownership Verification for Pre-trained Masked Models
di: Xie, Yuechen, et al.
Pubblicazione: (2025)
di: Xie, Yuechen, et al.
Pubblicazione: (2025)
Single Point, Full Mask: Velocity-Guided Level Set Evolution for End-to-End Amodal Segmentation
di: Li, Zhixuan, et al.
Pubblicazione: (2025)
di: Li, Zhixuan, et al.
Pubblicazione: (2025)
CLEAR: Context-Aware Learning with End-to-End Mask-Free Inference for Adaptive Video Subtitle Removal
di: He, Qingdong, et al.
Pubblicazione: (2026)
di: He, Qingdong, et al.
Pubblicazione: (2026)
Continual Learning Improves Zero-Shot Action Recognition
di: Gowda, Shreyank N, et al.
Pubblicazione: (2024)
di: Gowda, Shreyank N, et al.
Pubblicazione: (2024)
MaskHOI: Robust 3D Hand-Object Interaction Estimation via Masked Pre-training
di: Xie, Yuechen, et al.
Pubblicazione: (2025)
di: Xie, Yuechen, et al.
Pubblicazione: (2025)
BEEP3D: Box-Supervised End-to-End Pseudo-Mask Generation for 3D Instance Segmentation
di: Yoo, Youngju, et al.
Pubblicazione: (2025)
di: Yoo, Youngju, et al.
Pubblicazione: (2025)
Masked Pre-training Enables Universal Zero-shot Denoiser
di: Ma, Xiaoxiao, et al.
Pubblicazione: (2024)
di: Ma, Xiaoxiao, et al.
Pubblicazione: (2024)
Masked Clustering Prediction for Unsupervised Point Cloud Pre-training
di: Ren, Bin, et al.
Pubblicazione: (2025)
di: Ren, Bin, et al.
Pubblicazione: (2025)
MiM: Mask in Mask Self-Supervised Pre-Training for 3D Medical Image Analysis
di: Zhuang, Jiaxin, et al.
Pubblicazione: (2024)
di: Zhuang, Jiaxin, et al.
Pubblicazione: (2024)
There is No VAE: End-to-End Pixel-Space Generative Modeling via Self-Supervised Pre-training
di: Lei, Jiachen, et al.
Pubblicazione: (2025)
di: Lei, Jiachen, et al.
Pubblicazione: (2025)
Generative Planning with 3D-vision Language Pre-training for End-to-End Autonomous Driving
di: Li, Tengpeng, et al.
Pubblicazione: (2025)
di: Li, Tengpeng, et al.
Pubblicazione: (2025)
Adaptive Data Dropout: Towards Self-Regulated Learning in Deep Neural Networks
di: Gahir, Amar, et al.
Pubblicazione: (2026)
di: Gahir, Amar, et al.
Pubblicazione: (2026)
Focus on Texture: Rethinking Pre-training in Masked Autoencoders for Medical Image Classification
di: Madan, Chetan, et al.
Pubblicazione: (2025)
di: Madan, Chetan, et al.
Pubblicazione: (2025)
SelfMedHPM: Self Pre-training With Hard Patches Mining Masked Autoencoders For Medical Image Segmentation
di: Lv, Yunhao, et al.
Pubblicazione: (2025)
di: Lv, Yunhao, et al.
Pubblicazione: (2025)
Is Temporal Prompting All We Need For Limited Labeled Action Recognition?
di: Gowda, Shreyank N, et al.
Pubblicazione: (2025)
di: Gowda, Shreyank N, et al.
Pubblicazione: (2025)
MaskDiffusion: Exploiting Pre-trained Diffusion Models for Semantic Segmentation
di: Kawano, Yasufumi, et al.
Pubblicazione: (2024)
di: Kawano, Yasufumi, et al.
Pubblicazione: (2024)
Universal Image Restoration Pre-training via Masked Degradation Classification
di: Hu, JiaKui, et al.
Pubblicazione: (2025)
di: Hu, JiaKui, et al.
Pubblicazione: (2025)
$\mathsf{CSMAE~}$:~Cataract Surgical Masked Autoencoder (MAE) based Pre-training
di: Shah, Nisarg A., et al.
Pubblicazione: (2025)
di: Shah, Nisarg A., et al.
Pubblicazione: (2025)
PaCo-FR: Patch-Pixel Aligned End-to-End Codebook Learning for Facial Representation Pre-training
di: Xie, Yin, et al.
Pubblicazione: (2025)
di: Xie, Yin, et al.
Pubblicazione: (2025)
Salience-Based Adaptive Masking: Revisiting Token Dynamics for Enhanced Pre-training
di: Choi, Hyesong, et al.
Pubblicazione: (2024)
di: Choi, Hyesong, et al.
Pubblicazione: (2024)
Data-efficient Event Camera Pre-training via Disentangled Masked Modeling
di: Huang, Zhenpeng, et al.
Pubblicazione: (2024)
di: Huang, Zhenpeng, et al.
Pubblicazione: (2024)
Pan-cancer Histopathology WSI Pre-training with Position-aware Masked Autoencoder
di: Wu, Kun, et al.
Pubblicazione: (2024)
di: Wu, Kun, et al.
Pubblicazione: (2024)
Self Pre-training with Topology- and Spatiality-aware Masked Autoencoders for 3D Medical Image Segmentation
di: Gu, Pengfei, et al.
Pubblicazione: (2024)
di: Gu, Pengfei, et al.
Pubblicazione: (2024)
MaskMed: Decoupled Mask and Class Prediction for Medical Image Segmentation
di: Xie, Bin, et al.
Pubblicazione: (2025)
di: Xie, Bin, et al.
Pubblicazione: (2025)
Self Pre-training with Adaptive Mask Autoencoders for Variable-Contrast 3D Medical Imaging
di: Das, Badhan Kumar, et al.
Pubblicazione: (2025)
di: Das, Badhan Kumar, et al.
Pubblicazione: (2025)
Semantics-enhanced Cross-modal Masked Image Modeling for Vision-Language Pre-training
di: Liu, Haowei, et al.
Pubblicazione: (2024)
di: Liu, Haowei, et al.
Pubblicazione: (2024)
How Effective is Pre-training of Large Masked Autoencoders for Downstream Earth Observation Tasks?
di: Sosa, Jose, et al.
Pubblicazione: (2024)
di: Sosa, Jose, et al.
Pubblicazione: (2024)
Rethinking UMM Visual Generation: Masked Modeling for Efficient Image-Only Pre-training
di: Sun, Peng, et al.
Pubblicazione: (2026)
di: Sun, Peng, et al.
Pubblicazione: (2026)
Documenti analoghi
-
Distribution-Based Masked Medical Vision-Language Model Using Structured Reports
di: Gowda, Shreyank N, et al.
Pubblicazione: (2025) -
CC-SAM: SAM with Cross-feature Attention and Context for Ultrasound Image Segmentation
di: Gowda, Shreyank N, et al.
Pubblicazione: (2024) -
FE-Adapter: Adapting Image-based Emotion Classifiers to Videos
di: Gowda, Shreyank N, et al.
Pubblicazione: (2024) -
Prototype-Enhanced Confidence Modeling for Cross-Modal Medical Image-Report Retrieval
di: Gowda, Shreyank N, et al.
Pubblicazione: (2025) -
MaskFuser: Masked Fusion of Joint Multi-Modal Tokenization for End-to-End Autonomous Driving
di: Duan, Yiqun, et al.
Pubblicazione: (2024)