Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Lian, Chenyu, Zhou, Hong-Yu, Liang, Dongyun, Qin, Jing, Wang, Liansheng |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Evidential Reasoning Advances Interpretable Real-World Disease Screening
von: Lian, Chenyu, et al.
Veröffentlicht: (2026)
von: Lian, Chenyu, et al.
Veröffentlicht: (2026)
Less Could Be Better: Parameter-efficient Fine-tuning Advances Medical Vision Foundation Models
von: Lian, Chenyu, et al.
Veröffentlicht: (2024)
von: Lian, Chenyu, et al.
Veröffentlicht: (2024)
BenchReAD: A systematic benchmark for retinal anomaly detection
von: Lian, Chenyu, et al.
Veröffentlicht: (2025)
von: Lian, Chenyu, et al.
Veröffentlicht: (2025)
Concept-Guided Noisy Negative Suppression for Zero-Shot Classification and Grounding of Chest X-Ray Findings
von: Lian, Chenyu, et al.
Veröffentlicht: (2026)
von: Lian, Chenyu, et al.
Veröffentlicht: (2026)
AdaptVision: Efficient Vision-Language Models via Adaptive Visual Acquisition
von: Lin, Zichuan, et al.
Veröffentlicht: (2025)
von: Lin, Zichuan, et al.
Veröffentlicht: (2025)
Adapting Vision-Language Models Without Labels: A Comprehensive Survey
von: Dong, Hao, et al.
Veröffentlicht: (2025)
von: Dong, Hao, et al.
Veröffentlicht: (2025)
Adapting Vision-Language Models for Evaluating World Models
von: Hendriksen, Mariya, et al.
Veröffentlicht: (2025)
von: Hendriksen, Mariya, et al.
Veröffentlicht: (2025)
Masking Teacher and Reinforcing Student for Distilling Vision-Language Models
von: Lee, Byung-Kwan, et al.
Veröffentlicht: (2025)
von: Lee, Byung-Kwan, et al.
Veröffentlicht: (2025)
FastVLM: Efficient Vision Encoding for Vision Language Models
von: Vasu, Pavan Kumar Anasosalu, et al.
Veröffentlicht: (2024)
von: Vasu, Pavan Kumar Anasosalu, et al.
Veröffentlicht: (2024)
Improvise, Adapt, Overcome -- Telescopic Adapters for Efficient Fine-tuning of Vision Language Models in Medical Imaging
von: Mishra, Ujjwal, et al.
Veröffentlicht: (2025)
von: Mishra, Ujjwal, et al.
Veröffentlicht: (2025)
Look Through Masks: Towards Masked Face Recognition with De-Occlusion Distillation
von: Li, Chenyu, et al.
Veröffentlicht: (2024)
von: Li, Chenyu, et al.
Veröffentlicht: (2024)
Freeze the backbones: A Parameter-Efficient Contrastive Approach to Robust Medical Vision-Language Pre-training
von: Qin, Jiuming, et al.
Veröffentlicht: (2024)
von: Qin, Jiuming, et al.
Veröffentlicht: (2024)
RIV: Recursive Introspection Mask Diffusion Vision Language Model
von: Li, YuQian, et al.
Veröffentlicht: (2025)
von: Li, YuQian, et al.
Veröffentlicht: (2025)
Visual Modality Prompt for Adapting Vision-Language Object Detectors
von: Medeiros, Heitor R., et al.
Veröffentlicht: (2024)
von: Medeiros, Heitor R., et al.
Veröffentlicht: (2024)
Anatomy-VLM: A Fine-grained Vision-Language Model for Medical Interpretation
von: Gu, Difei, et al.
Veröffentlicht: (2025)
von: Gu, Difei, et al.
Veröffentlicht: (2025)
FG-CLIP 2: A Bilingual Fine-grained Vision-Language Alignment Model
von: Xie, Chunyu, et al.
Veröffentlicht: (2025)
von: Xie, Chunyu, et al.
Veröffentlicht: (2025)
Jailbreaking Vision-Language Models Through the Visual Modality
von: Azulay, Aharon, et al.
Veröffentlicht: (2026)
von: Azulay, Aharon, et al.
Veröffentlicht: (2026)
Enhancing Medical Large Vision-Language Models via Alignment Distillation
von: Chang, Aofei, et al.
Veröffentlicht: (2025)
von: Chang, Aofei, et al.
Veröffentlicht: (2025)
RadAlign: Advancing Radiology Report Generation with Vision-Language Concept Alignment
von: Gu, Difei, et al.
Veröffentlicht: (2025)
von: Gu, Difei, et al.
Veröffentlicht: (2025)
Robust Pre-Training of Medical Vision-and-Language Models with Domain-Invariant Multi-Modal Masked Reconstruction
von: Filvantorkaman, Melika, et al.
Veröffentlicht: (2026)
von: Filvantorkaman, Melika, et al.
Veröffentlicht: (2026)
Post-pre-training for Modality Alignment in Vision-Language Foundation Models
von: Yamaguchi, Shin'ya, et al.
Veröffentlicht: (2025)
von: Yamaguchi, Shin'ya, et al.
Veröffentlicht: (2025)
Learning Emergent Modular Representations in Multi-modality Medical Vision Foundation Models
von: He, Yuting, et al.
Veröffentlicht: (2026)
von: He, Yuting, et al.
Veröffentlicht: (2026)
Efficient Few-Shot Learning in Remote Sensing: Fusing Vision and Vision-Language Models
von: Chua, Jia Yun, et al.
Veröffentlicht: (2025)
von: Chua, Jia Yun, et al.
Veröffentlicht: (2025)
Topological Alignment of Shared Vision-Language Embedding Space
von: You, Junwon, et al.
Veröffentlicht: (2025)
von: You, Junwon, et al.
Veröffentlicht: (2025)
ZipVL: Efficient Large Vision-Language Models with Dynamic Token Sparsification
von: He, Yefei, et al.
Veröffentlicht: (2024)
von: He, Yefei, et al.
Veröffentlicht: (2024)
PETAR: Localized Findings Generation with Mask-Aware Vision-Language Modeling for PET Automated Reporting
von: Maqbool, Danyal, et al.
Veröffentlicht: (2025)
von: Maqbool, Danyal, et al.
Veröffentlicht: (2025)
Improving Medical Large Vision-Language Models with Abnormal-Aware Feedback
von: Zhou, Yucheng, et al.
Veröffentlicht: (2025)
von: Zhou, Yucheng, et al.
Veröffentlicht: (2025)
A Survey on Efficient Vision-Language-Action Models
von: Yu, Zhaoshu, et al.
Veröffentlicht: (2025)
von: Yu, Zhaoshu, et al.
Veröffentlicht: (2025)
Safety Alignment for Vision Language Models
von: Liu, Zhendong, et al.
Veröffentlicht: (2024)
von: Liu, Zhendong, et al.
Veröffentlicht: (2024)
VisionThink: Smart and Efficient Vision Language Model via Reinforcement Learning
von: Yang, Senqiao, et al.
Veröffentlicht: (2025)
von: Yang, Senqiao, et al.
Veröffentlicht: (2025)
Enhancing Visual-Language Modality Alignment in Large Vision Language Models via Self-Improvement
von: Wang, Xiyao, et al.
Veröffentlicht: (2024)
von: Wang, Xiyao, et al.
Veröffentlicht: (2024)
TinyAlign: Boosting Lightweight Vision-Language Models by Mitigating Modal Alignment Bottlenecks
von: Hu, Yuanze, et al.
Veröffentlicht: (2025)
von: Hu, Yuanze, et al.
Veröffentlicht: (2025)
MirrorCheck: Efficient Adversarial Defense for Vision-Language Models
von: Fares, Samar, et al.
Veröffentlicht: (2024)
von: Fares, Samar, et al.
Veröffentlicht: (2024)
Refer to Any Segmentation Mask Group With Vision-Language Prompts
von: Cao, Shengcao, et al.
Veröffentlicht: (2025)
von: Cao, Shengcao, et al.
Veröffentlicht: (2025)
Adapting Vision-Language Models for E-commerce Understanding at Scale
von: Nulli, Matteo, et al.
Veröffentlicht: (2026)
von: Nulli, Matteo, et al.
Veröffentlicht: (2026)
Distilling Out-of-Distribution Robustness from Vision-Language Foundation Models
von: Zhou, Andy, et al.
Veröffentlicht: (2023)
von: Zhou, Andy, et al.
Veröffentlicht: (2023)
Skill-Conditioned Visual Geolocation for Vision-Language Models
von: Yang, Chenjie, et al.
Veröffentlicht: (2026)
von: Yang, Chenjie, et al.
Veröffentlicht: (2026)
Inverse-LLaVA: Eliminating Alignment Pre-training Through Text-to-Vision Mapping
von: Zhan, Xuhui, et al.
Veröffentlicht: (2025)
von: Zhan, Xuhui, et al.
Veröffentlicht: (2025)
Analyzing Fine-Grained Alignment and Enhancing Vision Understanding in Multimodal Language Models
von: Jiang, Jiachen, et al.
Veröffentlicht: (2025)
von: Jiang, Jiachen, et al.
Veröffentlicht: (2025)
To Trust Or Not To Trust Your Vision-Language Model's Prediction
von: Dong, Hao, et al.
Veröffentlicht: (2025)
von: Dong, Hao, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Evidential Reasoning Advances Interpretable Real-World Disease Screening
von: Lian, Chenyu, et al.
Veröffentlicht: (2026) -
Less Could Be Better: Parameter-efficient Fine-tuning Advances Medical Vision Foundation Models
von: Lian, Chenyu, et al.
Veröffentlicht: (2024) -
BenchReAD: A systematic benchmark for retinal anomaly detection
von: Lian, Chenyu, et al.
Veröffentlicht: (2025) -
Concept-Guided Noisy Negative Suppression for Zero-Shot Classification and Grounding of Chest X-Ray Findings
von: Lian, Chenyu, et al.
Veröffentlicht: (2026) -
AdaptVision: Efficient Vision-Language Models via Adaptive Visual Acquisition
von: Lin, Zichuan, et al.
Veröffentlicht: (2025)