G2D: From Global to Dense Radiography Representation Learning via Vision-Language Pre-training
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Liu, Che, Ouyang, Cheng, Cheng, Sibo, Shah, Anand, Bai, Wenjia, Arcucci, Rossella |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2023
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
IMITATE: Clinical Prior Guided Hierarchical Vision-Language Pre-training
von: Liu, Che, et al.
Veröffentlicht: (2023)
von: Liu, Che, et al.
Veröffentlicht: (2023)
Utilizing Synthetic Data for Medical Vision-Language Pre-training: Bypassing the Need for Real Images
von: Liu, Che, et al.
Veröffentlicht: (2023)
von: Liu, Che, et al.
Veröffentlicht: (2023)
Freeze the backbones: A Parameter-Efficient Contrastive Approach to Robust Medical Vision-Language Pre-training
von: Qin, Jiuming, et al.
Veröffentlicht: (2024)
von: Qin, Jiuming, et al.
Veröffentlicht: (2024)
T3D: Advancing 3D Medical Vision-Language Pre-training by Learning Multi-View Visual Consistency
von: Liu, Che, et al.
Veröffentlicht: (2023)
von: Liu, Che, et al.
Veröffentlicht: (2023)
How Far Have Medical Vision-Language Models Come? A Comprehensive Benchmarking Study
von: Liu, Che, et al.
Veröffentlicht: (2025)
von: Liu, Che, et al.
Veröffentlicht: (2025)
Enhancing Abnormality Grounding for Vision Language Models with Knowledge Descriptions
von: Li, Jun, et al.
Veröffentlicht: (2025)
von: Li, Jun, et al.
Veröffentlicht: (2025)
Med-UniC: Unifying Cross-Lingual Medical Vision-Language Pre-Training by Diminishing Bias
von: Wan, Zhongwei, et al.
Veröffentlicht: (2023)
von: Wan, Zhongwei, et al.
Veröffentlicht: (2023)
Zero-Shot ECG Classification with Multimodal Learning and Test-time Clinical Knowledge Enhancement
von: Liu, Che, et al.
Veröffentlicht: (2024)
von: Liu, Che, et al.
Veröffentlicht: (2024)
Can Medical Vision-Language Pre-training Succeed with Purely Synthetic Data?
von: Liu, Che, et al.
Veröffentlicht: (2024)
von: Liu, Che, et al.
Veröffentlicht: (2024)
How Does Diverse Interpretability of Textual Prompts Impact Medical Vision-Language Zero-Shot Tasks?
von: Wang, Sicheng, et al.
Veröffentlicht: (2024)
von: Wang, Sicheng, et al.
Veröffentlicht: (2024)
Knowledge to Sight: Reasoning over Visual Attributes via Knowledge Decomposition for Abnormality Grounding
von: Li, Jun, et al.
Veröffentlicht: (2025)
von: Li, Jun, et al.
Veröffentlicht: (2025)
Knowledge-enhanced Multimodal ECG Representation Learning with Arbitrary-Lead Inputs
von: Liu, Che, et al.
Veröffentlicht: (2025)
von: Liu, Che, et al.
Veröffentlicht: (2025)
Argus: Benchmarking and Enhancing Vision-Language Models for 3D Radiology Report Generation
von: Liu, Che, et al.
Veröffentlicht: (2024)
von: Liu, Che, et al.
Veröffentlicht: (2024)
BIMCV-R: A Landmark Dataset for 3D CT Text-Image Retrieval
von: Chen, Yinda, et al.
Veröffentlicht: (2024)
von: Chen, Yinda, et al.
Veröffentlicht: (2024)
Fire-Image-DenseNet (FIDN) for predicting wildfire burnt area using remote sensing data
von: Pang, Bo, et al.
Veröffentlicht: (2024)
von: Pang, Bo, et al.
Veröffentlicht: (2024)
FMBench: Benchmarking Fairness in Multimodal Large Language Models on Medical Tasks
von: Wu, Peiran, et al.
Veröffentlicht: (2024)
von: Wu, Peiran, et al.
Veröffentlicht: (2024)
Superpixel Semantics Representation and Pre-training for Vision-Language Task
von: Zhang, Siyu, et al.
Veröffentlicht: (2023)
von: Zhang, Siyu, et al.
Veröffentlicht: (2023)
BOTM: Echocardiography Segmentation via Bi-directional Optimal Token Matching
von: Liu, Zhihua, et al.
Veröffentlicht: (2025)
von: Liu, Zhihua, et al.
Veröffentlicht: (2025)
Does DINOv3 Set a New Medical Vision Standard? Benchmarking 2D and 3D Classification, Segmentation, and Registration
von: Liu, Che, et al.
Veröffentlicht: (2025)
von: Liu, Che, et al.
Veröffentlicht: (2025)
SuPreME: A Supervised Pre-training Framework for Multimodal ECG Representation Learning
von: Cai, Mingsheng, et al.
Veröffentlicht: (2025)
von: Cai, Mingsheng, et al.
Veröffentlicht: (2025)
TorchDA: A Python package for performing data assimilation with deep learning forward and transformation functions
von: Cheng, Sibo, et al.
Veröffentlicht: (2024)
von: Cheng, Sibo, et al.
Veröffentlicht: (2024)
MLIP: Medical Language-Image Pre-training with Masked Local Representation Learning
von: Liu, Jiarun, et al.
Veröffentlicht: (2024)
von: Liu, Jiarun, et al.
Veröffentlicht: (2024)
Noise2Noise Denoising of CRISM Hyperspectral Data
von: Platt, Robert, et al.
Veröffentlicht: (2024)
von: Platt, Robert, et al.
Veröffentlicht: (2024)
Multi-fidelity physics constrained neural networks for dynamical systems
von: Zhou, Hao, et al.
Veröffentlicht: (2024)
von: Zhou, Hao, et al.
Veröffentlicht: (2024)
Can LLM-Generated Text Empower Surgical Vision-Language Pre-training?
von: Che, Chengan, et al.
Veröffentlicht: (2026)
von: Che, Chengan, et al.
Veröffentlicht: (2026)
MAA: Meticulous Adversarial Attack against Vision-Language Pre-trained Models
von: Zhang, Peng-Fei, et al.
Veröffentlicht: (2025)
von: Zhang, Peng-Fei, et al.
Veröffentlicht: (2025)
Unsupervised Pre-training with Language-Vision Prompts for Low-Data Instance Segmentation
von: Zhang, Dingwen, et al.
Veröffentlicht: (2024)
von: Zhang, Dingwen, et al.
Veröffentlicht: (2024)
PLIP: Language-Image Pre-training for Person Representation Learning
von: Zuo, Jialong, et al.
Veröffentlicht: (2023)
von: Zuo, Jialong, et al.
Veröffentlicht: (2023)
TIP: Tabular-Image Pre-training for Multimodal Classification with Incomplete Data
von: Du, Siyi, et al.
Veröffentlicht: (2024)
von: Du, Siyi, et al.
Veröffentlicht: (2024)
3D Scene Graph Guided Vision-Language Pre-training
von: Liu, Hao, et al.
Veröffentlicht: (2024)
von: Liu, Hao, et al.
Veröffentlicht: (2024)
Enhancing Representation in Radiography-Reports Foundation Model: A Granular Alignment Algorithm Using Masked Contrastive Learning
von: Huang, Weijian, et al.
Veröffentlicht: (2023)
von: Huang, Weijian, et al.
Veröffentlicht: (2023)
Event Camera Data Dense Pre-training
von: Yang, Yan, et al.
Veröffentlicht: (2023)
von: Yang, Yan, et al.
Veröffentlicht: (2023)
Pre-trained Model Guided Fine-Tuning for Zero-Shot Adversarial Robustness
von: Wang, Sibo, et al.
Veröffentlicht: (2024)
von: Wang, Sibo, et al.
Veröffentlicht: (2024)
SPOT: Scalable 3D Pre-training via Occupancy Prediction for Learning Transferable 3D Representations
von: Yan, Xiangchao, et al.
Veröffentlicht: (2023)
von: Yan, Xiangchao, et al.
Veröffentlicht: (2023)
Boosting 3D Neuron Segmentation with 2D Vision Transformer Pre-trained on Natural Images
von: Cheng, Yik San, et al.
Veröffentlicht: (2024)
von: Cheng, Yik San, et al.
Veröffentlicht: (2024)
ConDense: Consistent 2D/3D Pre-training for Dense and Sparse Features from Multi-View Images
von: Zhang, Xiaoshuai, et al.
Veröffentlicht: (2024)
von: Zhang, Xiaoshuai, et al.
Veröffentlicht: (2024)
Sculpting Holistic 3D Representation in Contrastive Language-Image-3D Pre-training
von: Gao, Yipeng, et al.
Veröffentlicht: (2023)
von: Gao, Yipeng, et al.
Veröffentlicht: (2023)
Efficient Vision-Language Pre-training by Cluster Masking
von: Wei, Zihao, et al.
Veröffentlicht: (2024)
von: Wei, Zihao, et al.
Veröffentlicht: (2024)
Enhancing Vision-Language Pre-training with Rich Supervisions
von: Gao, Yuan, et al.
Veröffentlicht: (2024)
von: Gao, Yuan, et al.
Veröffentlicht: (2024)
Universal Adversarial Perturbations for Vision-Language Pre-trained Models
von: Zhang, Peng-Fei, et al.
Veröffentlicht: (2024)
von: Zhang, Peng-Fei, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
IMITATE: Clinical Prior Guided Hierarchical Vision-Language Pre-training
von: Liu, Che, et al.
Veröffentlicht: (2023) -
Utilizing Synthetic Data for Medical Vision-Language Pre-training: Bypassing the Need for Real Images
von: Liu, Che, et al.
Veröffentlicht: (2023) -
Freeze the backbones: A Parameter-Efficient Contrastive Approach to Robust Medical Vision-Language Pre-training
von: Qin, Jiuming, et al.
Veröffentlicht: (2024) -
T3D: Advancing 3D Medical Vision-Language Pre-training by Learning Multi-View Visual Consistency
von: Liu, Che, et al.
Veröffentlicht: (2023) -
How Far Have Medical Vision-Language Models Come? A Comprehensive Benchmarking Study
von: Liu, Che, et al.
Veröffentlicht: (2025)