Boosting Medical Vision-Language Pretraining via Momentum Self-Distillation under Limited Computing Resources
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Pham, Phuc, Pham, Nhu, Ly, Ngoc Quoc |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Recent Advances in Medical Image Classification
von: Dao, Loan, et al.
Veröffentlicht: (2025)
von: Dao, Loan, et al.
Veröffentlicht: (2025)
A Comprehensive Study on Medical Image Segmentation using Deep Neural Networks
von: Dao, Loan, et al.
Veröffentlicht: (2025)
von: Dao, Loan, et al.
Veröffentlicht: (2025)
Enhancing Feature Diversity Boosts Channel-Adaptive Vision Transformers
von: Pham, Chau, et al.
Veröffentlicht: (2024)
von: Pham, Chau, et al.
Veröffentlicht: (2024)
ViCLIP-OT: The First Foundation Vision-Language Model for Vietnamese Image-Text Retrieval with Optimal Transport
von: Tran, Quoc-Khang, et al.
Veröffentlicht: (2026)
von: Tran, Quoc-Khang, et al.
Veröffentlicht: (2026)
Ontology-based knowledge representation for bone disease diagnosis: a foundation for safe and sustainable medical artificial intelligence systems
von: Dao, Loan, et al.
Veröffentlicht: (2025)
von: Dao, Loan, et al.
Veröffentlicht: (2025)
Enhancing Medical Large Vision-Language Models via Alignment Distillation
von: Chang, Aofei, et al.
Veröffentlicht: (2025)
von: Chang, Aofei, et al.
Veröffentlicht: (2025)
STER-VLM: Spatio-Temporal With Enhanced Reference Vision-Language Models
von: Nguyen-Nhu, Tinh-Anh, et al.
Veröffentlicht: (2025)
von: Nguyen-Nhu, Tinh-Anh, et al.
Veröffentlicht: (2025)
CheXmix: Unified Generative Pretraining for Vision Language Models in Medical Imaging
von: Kumar, Ashwin, et al.
Veröffentlicht: (2026)
von: Kumar, Ashwin, et al.
Veröffentlicht: (2026)
Multi-Aspect Knowledge-Enhanced Medical Vision-Language Pretraining with Multi-Agent Data Generation
von: Li, Xieji, et al.
Veröffentlicht: (2025)
von: Li, Xieji, et al.
Veröffentlicht: (2025)
Boost Self-Supervised Dataset Distillation via Parameterization, Predefined Augmentation, and Approximation
von: Yu, Sheng-Feng, et al.
Veröffentlicht: (2025)
von: Yu, Sheng-Feng, et al.
Veröffentlicht: (2025)
Aleatoric Uncertainty Medical Image Segmentation Estimation via Flow Matching
von: Van Nguyen, Phi, et al.
Veröffentlicht: (2025)
von: Van Nguyen, Phi, et al.
Veröffentlicht: (2025)
ALPI: Auto-Labeller with Proxy Injection for 3D Object Detection using 2D Labels Only
von: Lahlali, Saad, et al.
Veröffentlicht: (2024)
von: Lahlali, Saad, et al.
Veröffentlicht: (2024)
Diffusion Model in Latent Space for Medical Image Segmentation Task
von: Ngoc, Huynh Trinh, et al.
Veröffentlicht: (2025)
von: Ngoc, Huynh Trinh, et al.
Veröffentlicht: (2025)
OE3DIS: Open-Ended 3D Point Cloud Instance Segmentation
von: Nguyen, Phuc D. A., et al.
Veröffentlicht: (2024)
von: Nguyen, Phuc D. A., et al.
Veröffentlicht: (2024)
TinySSL: Distilled Self-Supervised Pretraining for Sub-Megabyte MCU Models
von: Wilson, Bibin
Veröffentlicht: (2026)
von: Wilson, Bibin
Veröffentlicht: (2026)
QTSeg: A Query Token-Based Dual-Mix Attention Framework with Multi-Level Feature Distribution for Medical Image Segmentation
von: Tran, Phuong-Nam, et al.
Veröffentlicht: (2024)
von: Tran, Phuong-Nam, et al.
Veröffentlicht: (2024)
Exploring the Application of Visual Question Answering (VQA) for Classroom Activity Monitoring
von: Vu, Sinh Trong, et al.
Veröffentlicht: (2025)
von: Vu, Sinh Trong, et al.
Veröffentlicht: (2025)
Confounder-Aware Medical Data Selection for Fine-Tuning Pretrained Vision Models
von: Ji, Anyang, et al.
Veröffentlicht: (2025)
von: Ji, Anyang, et al.
Veröffentlicht: (2025)
Comparative Study of UNet-based Architectures for Liver Tumor Segmentation in Multi-Phase Contrast-Enhanced Computed Tomography
von: Ly, Doan-Van-Anh, et al.
Veröffentlicht: (2025)
von: Ly, Doan-Van-Anh, et al.
Veröffentlicht: (2025)
GMAT: Grounded Multi-Agent Clinical Description Generation for Text Encoder in Vision-Language MIL for Whole Slide Image Classification
von: Quang, Ngoc Bui Lam, et al.
Veröffentlicht: (2025)
von: Quang, Ngoc Bui Lam, et al.
Veröffentlicht: (2025)
SLIP: Structural-aware Language-Image Pretraining for Vision-Language Alignment
von: Lu, Wenbo
Veröffentlicht: (2025)
von: Lu, Wenbo
Veröffentlicht: (2025)
DINORANKCLIP: DINOv3 Distillation and Injection for Vision-Language Pretraining with High-Order Ranking Consistency
von: Jiang, Shuyang, et al.
Veröffentlicht: (2026)
von: Jiang, Shuyang, et al.
Veröffentlicht: (2026)
LiteNeXt: A Novel Lightweight ConvMixer-based Model with Self-embedding Representation Parallel for Medical Image Segmentation
von: Tran, Ngoc-Du, et al.
Veröffentlicht: (2024)
von: Tran, Ngoc-Du, et al.
Veröffentlicht: (2024)
UMSPU: Universal Multi-Size Phase Unwrapping via Mutual Self-Distillation and Adaptive Boosting Ensemble Segmenters
von: Du, Lintong, et al.
Veröffentlicht: (2024)
von: Du, Lintong, et al.
Veröffentlicht: (2024)
HDC: Hierarchical Distillation for Multi-level Noisy Consistency in Semi-Supervised Fetal Ultrasound Segmentation
von: Le, Tran Quoc Khanh, et al.
Veröffentlicht: (2025)
von: Le, Tran Quoc Khanh, et al.
Veröffentlicht: (2025)
ThyroidEffi 1.0: A Cost-Effective System for High-Performance Multi-Class Thyroid Carcinoma Classification
von: Pham-Ngoc, Hai, et al.
Veröffentlicht: (2025)
von: Pham-Ngoc, Hai, et al.
Veröffentlicht: (2025)
Person Re-Identification System at Semantic Level based on Pedestrian Attributes Ontology
von: Ly, Ngoc Q., et al.
Veröffentlicht: (2025)
von: Ly, Ngoc Q., et al.
Veröffentlicht: (2025)
Leveraging Model Soups to Classify Intangible Cultural Heritage Images from the Mekong Delta
von: Tran, Quoc-Khang, et al.
Veröffentlicht: (2026)
von: Tran, Quoc-Khang, et al.
Veröffentlicht: (2026)
Adversarial Prompt Distillation for Vision-Language Models
von: Luo, Lin, et al.
Veröffentlicht: (2024)
von: Luo, Lin, et al.
Veröffentlicht: (2024)
Text-to-CT Generation via 3D Latent Diffusion Model with Contrastive Vision-Language Pretraining
von: Molino, Daniele, et al.
Veröffentlicht: (2025)
von: Molino, Daniele, et al.
Veröffentlicht: (2025)
Investigating and Mitigating Object Hallucinations in Pretrained Vision-Language (CLIP) Models
von: Liu, Yufang, et al.
Veröffentlicht: (2024)
von: Liu, Yufang, et al.
Veröffentlicht: (2024)
Sim4Seg: Boosting Multimodal Multi-disease Medical Diagnosis Segmentation with Region-Aware Vision-Language Similarity Masks
von: Song, Lingran, et al.
Veröffentlicht: (2025)
von: Song, Lingran, et al.
Veröffentlicht: (2025)
Multimodal Distribution Matching for Vision-Language Dataset Distillation
von: Jeong, Jongoh, et al.
Veröffentlicht: (2026)
von: Jeong, Jongoh, et al.
Veröffentlicht: (2026)
SyncMask: Synchronized Attentional Masking for Fashion-centric Vision-Language Pretraining
von: Song, Chull Hwan, et al.
Veröffentlicht: (2024)
von: Song, Chull Hwan, et al.
Veröffentlicht: (2024)
MAE-Based Self-Supervised Pretraining for Data-Efficient Medical Image Segmentation Using nnFormer
von: Sureddi, R. M. Krishna, et al.
Veröffentlicht: (2026)
von: Sureddi, R. M. Krishna, et al.
Veröffentlicht: (2026)
Accuracy-Robustness Trade Off via Spiking Neural Network Gradient Sparsity Trail
von: Nhan, Luu Trong, et al.
Veröffentlicht: (2025)
von: Nhan, Luu Trong, et al.
Veröffentlicht: (2025)
Unifying Global and Local Scene Entities Modelling for Precise Action Spotting
von: Tran, Kim Hoang, et al.
Veröffentlicht: (2024)
von: Tran, Kim Hoang, et al.
Veröffentlicht: (2024)
Any-to-Any Learning in Computational Pathology via Triplet Multimodal Pretraining
von: Sun, Qichen, et al.
Veröffentlicht: (2025)
von: Sun, Qichen, et al.
Veröffentlicht: (2025)
Single-Teacher View Augmentation: Boosting Knowledge Distillation via Angular Diversity
von: Yu, Seonghoon, et al.
Veröffentlicht: (2025)
von: Yu, Seonghoon, et al.
Veröffentlicht: (2025)
ConPro: Learning Severity Representation for Medical Images using Contrastive Learning and Preference Optimization
von: Nguyen, Hong, et al.
Veröffentlicht: (2024)
von: Nguyen, Hong, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Recent Advances in Medical Image Classification
von: Dao, Loan, et al.
Veröffentlicht: (2025) -
A Comprehensive Study on Medical Image Segmentation using Deep Neural Networks
von: Dao, Loan, et al.
Veröffentlicht: (2025) -
Enhancing Feature Diversity Boosts Channel-Adaptive Vision Transformers
von: Pham, Chau, et al.
Veröffentlicht: (2024) -
ViCLIP-OT: The First Foundation Vision-Language Model for Vietnamese Image-Text Retrieval with Optimal Transport
von: Tran, Quoc-Khang, et al.
Veröffentlicht: (2026) -
Ontology-based knowledge representation for bone disease diagnosis: a foundation for safe and sustainable medical artificial intelligence systems
von: Dao, Loan, et al.
Veröffentlicht: (2025)