ASAP: Advancing Medical Volumetric Representation Learning with Anatomy-aware Semantically-adaptive Pre-training
Fuente:
arXiv
Saved in:
| Main Authors: | Wang, Rongsheng, Tang, Fenghe, Jiang, Zihang, Li, Yingtai, Zhang, Xu, Lai, Haoran, Ma, Wenxin, Wei, Wei, He, Zhiyang, Tao, Xiaodong, Yan, Rui, Yao, Qingsong, Zhou, Shaohua Kevin |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
ECAMP: Entity-centered Context-aware Medical Vision Language Pre-training
by: Wang, Rongsheng, et al.
Published: (2023)
by: Wang, Rongsheng, et al.
Published: (2023)
SimCroP: Radiograph Representation Learning with Similarity-driven Cross-granularity Pre-training
by: Wang, Rongsheng, et al.
Published: (2025)
by: Wang, Rongsheng, et al.
Published: (2025)
Bridged Semantic Alignment for Zero-shot 3D Medical Image Diagnosis
by: Lai, Haoran, et al.
Published: (2025)
by: Lai, Haoran, et al.
Published: (2025)
Pre-Trained LLM is a Semantic-Aware and Generalizable Segmentation Booster
by: Tang, Fenghe, et al.
Published: (2025)
by: Tang, Fenghe, et al.
Published: (2025)
Med3D-R1: Incentivizing Clinical Reasoning in 3D Medical Vision-Language Models for Abnormality Diagnosis
by: Lai, Haoran, et al.
Published: (2026)
by: Lai, Haoran, et al.
Published: (2026)
MambaMIM: Pre-training Mamba with State Space Token Interpolation and its Application to Medical Image Segmentation
by: Tang, Fenghe, et al.
Published: (2024)
by: Tang, Fenghe, et al.
Published: (2024)
E3D-GPT: Enhanced 3D Visual Foundation for Medical Vision-Language Model
by: Lai, Haoran, et al.
Published: (2024)
by: Lai, Haoran, et al.
Published: (2024)
CARZero: Cross-Attention Alignment for Radiology Zero-Shot Classification
by: Lai, Haoran, et al.
Published: (2024)
by: Lai, Haoran, et al.
Published: (2024)
More performant and scalable: Rethinking contrastive vision-language pre-training of radiology in the LLM era
by: Li, Yingtai, et al.
Published: (2025)
by: Li, Yingtai, et al.
Published: (2025)
AA-CLIP: Enhancing Zero-shot Anomaly Detection via Anomaly-Aware CLIP
by: Ma, Wenxin, et al.
Published: (2025)
by: Ma, Wenxin, et al.
Published: (2025)
DiffVP: Differential Visual Semantic Prompting for LLM-Based CT Report Generation
by: Tian, Yuhe, et al.
Published: (2026)
by: Tian, Yuhe, et al.
Published: (2026)
MedReason-R1: Learning to Reason for CT Diagnosis with Reinforcement Learning and Local Zoom
by: Li, Yifan, et al.
Published: (2025)
by: Li, Yifan, et al.
Published: (2025)
Concept-to-Pixel: Prompt-Free Universal Medical Image Segmentation
by: Chen, Haoyun, et al.
Published: (2026)
by: Chen, Haoyun, et al.
Published: (2026)
Hi-End-MAE: Hierarchical encoder-driven masked autoencoders are stronger vision learners for medical image segmentation
by: Tang, Fenghe, et al.
Published: (2025)
by: Tang, Fenghe, et al.
Published: (2025)
GreenRFM: Toward a resource-efficient radiology foundation model
by: Li, Yingtai, et al.
Published: (2026)
by: Li, Yingtai, et al.
Published: (2026)
UCAD: Uncertainty-guided Contour-aware Displacement for semi-supervised medical image segmentation
by: Ding, Chengbo, et al.
Published: (2026)
by: Ding, Chengbo, et al.
Published: (2026)
U-Bench: A Comprehensive Understanding of U-Net through 100-Variant Benchmarking
by: Tang, Fenghe, et al.
Published: (2025)
by: Tang, Fenghe, et al.
Published: (2025)
APPLE: Adversarial Privacy-aware Perturbations on Latent Embedding for Unfairness Mitigation
by: Xu, Zikang, et al.
Published: (2024)
by: Xu, Zikang, et al.
Published: (2024)
A General Knowledge Injection Framework for ICD Coding
by: Zhang, Xu, et al.
Published: (2025)
by: Zhang, Xu, et al.
Published: (2025)
HySparK: Hybrid Sparse Masking for Large Scale Medical Image Pre-Training
by: Tang, Fenghe, et al.
Published: (2024)
by: Tang, Fenghe, et al.
Published: (2024)
Boosting Vision Semantic Density with Anatomy Normality Modeling for Medical Vision-language Pre-training
by: Cao, Weiwei, et al.
Published: (2025)
by: Cao, Weiwei, et al.
Published: (2025)
From Documents to Spans: Scalable Supervision for Evidence-Based ICD Coding with LLMs
by: Zhang, Xu, et al.
Published: (2026)
by: Zhang, Xu, et al.
Published: (2026)
MeDSLIP: Medical Dual-Stream Language-Image Pre-training with Pathology-Anatomy Semantic Alignment
by: Fan, Wenrui, et al.
Published: (2024)
by: Fan, Wenrui, et al.
Published: (2024)
Exploring Simple Open-Vocabulary Semantic Segmentation
by: Lai, Zihang
Published: (2024)
by: Lai, Zihang
Published: (2024)
Advancing Medical Radiograph Representation Learning: A Hybrid Pre-training Paradigm with Multilevel Semantic Granularity
by: Jiang, Hanqi, et al.
Published: (2024)
by: Jiang, Hanqi, et al.
Published: (2024)
Improving Anomalous Sound Detection with Attribute-aware Representation from Domain-adaptive Pre-training
by: Fang, Xin, et al.
Published: (2025)
by: Fang, Xin, et al.
Published: (2025)
Equivariant Sampling for Improving Diffusion Model-based Image Restoration
by: Wu, Chenxu, et al.
Published: (2025)
by: Wu, Chenxu, et al.
Published: (2025)
Towards Accurate Unified Anomaly Segmentation
by: Ma, Wenxin, et al.
Published: (2025)
by: Ma, Wenxin, et al.
Published: (2025)
ASAP: Advancing Semantic Alignment Promotes Multi-Modal Manipulation Detecting and Grounding
by: Zhang, Zhenxing, et al.
Published: (2024)
by: Zhang, Zhenxing, et al.
Published: (2024)
Multi-level Asymmetric Contrastive Learning for Volumetric Medical Image Segmentation Pre-training
by: Zeng, Shuang, et al.
Published: (2023)
by: Zeng, Shuang, et al.
Published: (2023)
3DGR-CAR: Coronary artery reconstruction from ultra-sparse 2D X-ray views with a 3D Gaussians representation
by: Fu, Xueming, et al.
Published: (2024)
by: Fu, Xueming, et al.
Published: (2024)
SRSNetwork: Siamese Reconstruction-Segmentation Networks based on Dynamic-Parameter Convolution
by: Nian, Bingkun, et al.
Published: (2023)
by: Nian, Bingkun, et al.
Published: (2023)
Semi-supervised Medical Image Segmentation via Geometry-aware Consistency Training
by: Liu, Zihang, et al.
Published: (2022)
by: Liu, Zihang, et al.
Published: (2022)
SCALE-VLP: Soft-Weighted Contrastive Volumetric Vision-Language Pre-training with Spatial-Knowledge Semantics
by: Mahdizadeh, Ailar, et al.
Published: (2025)
by: Mahdizadeh, Ailar, et al.
Published: (2025)
Slide-SAM: Medical SAM Meets Sliding Window
by: Quan, Quan, et al.
Published: (2023)
by: Quan, Quan, et al.
Published: (2023)
The Hamiltonian properties of rectangular meshes with at most two faulty nodes
by: Xie, Yingtai
Published: (2025)
by: Xie, Yingtai
Published: (2025)
Pre-training data selection for biomedical domain adaptation using journal impact metrics
by: Laï-king, Mathieu, et al.
Published: (2024)
by: Laï-king, Mathieu, et al.
Published: (2024)
Dyna3DGR: 4D Cardiac Motion Tracking with Dynamic 3D Gaussian Representation
by: Fu, Xueming, et al.
Published: (2025)
by: Fu, Xueming, et al.
Published: (2025)
Superpixel Semantics Representation and Pre-training for Vision-Language Task
by: Zhang, Siyu, et al.
Published: (2023)
by: Zhang, Siyu, et al.
Published: (2023)
Dynamic Self-adaptive Multiscale Distillation from Pre-trained Multimodal Large Model for Efficient Cross-modal Representation Learning
by: Liang, Zhengyang, et al.
Published: (2024)
by: Liang, Zhengyang, et al.
Published: (2024)
Similar Items
-
ECAMP: Entity-centered Context-aware Medical Vision Language Pre-training
by: Wang, Rongsheng, et al.
Published: (2023) -
SimCroP: Radiograph Representation Learning with Similarity-driven Cross-granularity Pre-training
by: Wang, Rongsheng, et al.
Published: (2025) -
Bridged Semantic Alignment for Zero-shot 3D Medical Image Diagnosis
by: Lai, Haoran, et al.
Published: (2025) -
Pre-Trained LLM is a Semantic-Aware and Generalizable Segmentation Booster
by: Tang, Fenghe, et al.
Published: (2025) -
Med3D-R1: Incentivizing Clinical Reasoning in 3D Medical Vision-Language Models for Abnormality Diagnosis
by: Lai, Haoran, et al.
Published: (2026)