Bridging Vision and Language: Optimal Transport-Driven Radiology Report Generation via LLMs
Fuente:
arXiv
Saved in:
| Main Authors: | Zhao, Haifeng, Zhang, Yufei, Ma, Leilei, Xu, Shuo, Sun, Dengdi |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Domain Adaptive Lung Nodule Detection in X-ray Image
by: Zhao, Haifeng, et al.
Published: (2024)
by: Zhao, Haifeng, et al.
Published: (2024)
Fully Automated SAM for Single-source Domain Generalization in Medical Image Segmentation
by: Zhuo, Huanli, et al.
Published: (2025)
by: Zhuo, Huanli, et al.
Published: (2025)
Text-Region Matching for Multi-Label Image Recognition with Missing Labels
by: Ma, Leilei, et al.
Published: (2024)
by: Ma, Leilei, et al.
Published: (2024)
Correlative and Discriminative Label Grouping for Multi-Label Visual Prompt Tuning
by: Ma, LeiLei, et al.
Published: (2025)
by: Ma, LeiLei, et al.
Published: (2025)
Bidirectional Uncertainty-Aware Region Learning for Semi-Supervised Medical Image Segmentation
by: Zhou, Shiwei, et al.
Published: (2025)
by: Zhou, Shiwei, et al.
Published: (2025)
R2GenKG: Hierarchical Multi-modal Knowledge Graph for LLM-based Radiology Report Generation
by: Wang, Futian, et al.
Published: (2025)
by: Wang, Futian, et al.
Published: (2025)
Griffon-G: Bridging Vision-Language and Vision-Centric Tasks via Large Multimodal Models
by: Zhan, Yufei, et al.
Published: (2024)
by: Zhan, Yufei, et al.
Published: (2024)
Visual Alignment of Medical Vision-Language Models for Grounded Radiology Report Generation
by: Bose, Sarosij, et al.
Published: (2025)
by: Bose, Sarosij, et al.
Published: (2025)
Intensive Vision-guided Network for Radiology Report Generation
by: Zheng, Fudan, et al.
Published: (2024)
by: Zheng, Fudan, et al.
Published: (2024)
Dynamic Prompt Adjustment for Multi-Label Class-Incremental Learning
by: Zhao, Haifeng, et al.
Published: (2024)
by: Zhao, Haifeng, et al.
Published: (2024)
Evaluating Vision Language Model Adaptations for Radiology Report Generation in Low-Resource Languages
by: Salmè, Marco, et al.
Published: (2025)
by: Salmè, Marco, et al.
Published: (2025)
RIHA: Report-Image Hierarchical Alignment for Radiology Report Generation
by: Chen, Yucheng, et al.
Published: (2026)
by: Chen, Yucheng, et al.
Published: (2026)
ORID: Organ-Regional Information Driven Framework for Radiology Report Generation
by: Gu, Tiancheng, et al.
Published: (2024)
by: Gu, Tiancheng, et al.
Published: (2024)
Argus: Benchmarking and Enhancing Vision-Language Models for 3D Radiology Report Generation
by: Liu, Che, et al.
Published: (2024)
by: Liu, Che, et al.
Published: (2024)
RaDialog: A Large Vision-Language Model for Radiology Report Generation and Conversational Assistance
by: Pellegrini, Chantal, et al.
Published: (2023)
by: Pellegrini, Chantal, et al.
Published: (2023)
Bridging Different Language Models and Generative Vision Models for Text-to-Image Generation
by: Zhao, Shihao, et al.
Published: (2024)
by: Zhao, Shihao, et al.
Published: (2024)
Improving Medical Visual Representations via Radiology Report Generation
by: Quigley, Keegan, et al.
Published: (2023)
by: Quigley, Keegan, et al.
Published: (2023)
MCA-RG: Enhancing LLMs with Medical Concept Alignment for Radiology Report Generation
by: Xing, Qilong, et al.
Published: (2025)
by: Xing, Qilong, et al.
Published: (2025)
3D-CT-GPT: Generating 3D Radiology Reports through Integration of Large Vision-Language Models
by: Chen, Hao, et al.
Published: (2024)
by: Chen, Hao, et al.
Published: (2024)
Any-to-Any Vision-Language Model for Multimodal X-ray Imaging and Radiological Report Generation
by: Molino, Daniele, et al.
Published: (2025)
by: Molino, Daniele, et al.
Published: (2025)
Visual Prompt Engineering for Vision Language Models in Radiology
by: Denner, Stefan, et al.
Published: (2024)
by: Denner, Stefan, et al.
Published: (2024)
Hierarchical Semantic-Augmented Navigation: Optimal Transport and Graph-Driven Reasoning for Vision-Language Navigation
by: Fang, Xiang, et al.
Published: (2026)
by: Fang, Xiang, et al.
Published: (2026)
HC-LLM: Historical-Constrained Large Language Models for Radiology Report Generation
by: Liu, Tengfei, et al.
Published: (2024)
by: Liu, Tengfei, et al.
Published: (2024)
Scene Graph Aided Radiology Report Generation
by: Wang, Jun, et al.
Published: (2024)
by: Wang, Jun, et al.
Published: (2024)
Rethinking the Efficiency and Effectiveness of Reinforcement Learning for Radiology Report Generation
by: Lu, Zilin, et al.
Published: (2026)
by: Lu, Zilin, et al.
Published: (2026)
ICON: Improving Inter-Report Consistency in Radiology Report Generation via Lesion-aware Mixup Augmentation
by: Hou, Wenjun, et al.
Published: (2024)
by: Hou, Wenjun, et al.
Published: (2024)
RadAlign: Advancing Radiology Report Generation with Vision-Language Concept Alignment
by: Gu, Difei, et al.
Published: (2025)
by: Gu, Difei, et al.
Published: (2025)
Adapting Foundation Vision-Language Models to Medical Diagnosis via Query-Driven Expert Bridging
by: Li, Yitong, et al.
Published: (2025)
by: Li, Yitong, et al.
Published: (2025)
Textual Inversion and Self-supervised Refinement for Radiology Report Generation
by: Luo, Yuanjiang, et al.
Published: (2024)
by: Luo, Yuanjiang, et al.
Published: (2024)
Radiology Report Generation with Layer-Wise Anatomical Attention
by: Muñiz-De-León, Emmanuel D., et al.
Published: (2025)
by: Muñiz-De-León, Emmanuel D., et al.
Published: (2025)
Anatomical Attention Alignment representation for Radiology Report Generation
by: Nguyen, Quang Vinh, et al.
Published: (2025)
by: Nguyen, Quang Vinh, et al.
Published: (2025)
Contrastive Learning with Counterfactual Explanations for Radiology Report Generation
by: Li, Mingjie, et al.
Published: (2024)
by: Li, Mingjie, et al.
Published: (2024)
HERGen: Elevating Radiology Report Generation with Longitudinal Data
by: Wang, Fuying, et al.
Published: (2024)
by: Wang, Fuying, et al.
Published: (2024)
AWT: Transferring Vision-Language Models via Augmentation, Weighting, and Transportation
by: Zhu, Yuhan, et al.
Published: (2024)
by: Zhu, Yuhan, et al.
Published: (2024)
FOCUS: Unified Vision-Language Modeling for Interactive Editing Driven by Referential Segmentation
by: Yang, Fan, et al.
Published: (2025)
by: Yang, Fan, et al.
Published: (2025)
Adapting Lightweight Vision Language Models for Radiological Visual Question Answering
by: Shourya, Aditya, et al.
Published: (2025)
by: Shourya, Aditya, et al.
Published: (2025)
RADAR: Enhancing Radiology Report Generation with Supplementary Knowledge Injection
by: Hou, Wenjun, et al.
Published: (2025)
by: Hou, Wenjun, et al.
Published: (2025)
MAIRA-2: Grounded Radiology Report Generation
by: Bannur, Shruthi, et al.
Published: (2024)
by: Bannur, Shruthi, et al.
Published: (2024)
Memory-based Cross-modal Semantic Alignment Network for Radiology Report Generation
by: Tao, Yitian, et al.
Published: (2024)
by: Tao, Yitian, et al.
Published: (2024)
MIND-Edit: MLLM Insight-Driven Editing via Language-Vision Projection
by: Wang, Shuyu, et al.
Published: (2025)
by: Wang, Shuyu, et al.
Published: (2025)
Similar Items
-
Domain Adaptive Lung Nodule Detection in X-ray Image
by: Zhao, Haifeng, et al.
Published: (2024) -
Fully Automated SAM for Single-source Domain Generalization in Medical Image Segmentation
by: Zhuo, Huanli, et al.
Published: (2025) -
Text-Region Matching for Multi-Label Image Recognition with Missing Labels
by: Ma, Leilei, et al.
Published: (2024) -
Correlative and Discriminative Label Grouping for Multi-Label Visual Prompt Tuning
by: Ma, LeiLei, et al.
Published: (2025) -
Bidirectional Uncertainty-Aware Region Learning for Semi-Supervised Medical Image Segmentation
by: Zhou, Shiwei, et al.
Published: (2025)