Any-to-Any Vision-Language Model for Multimodal X-ray Imaging and Radiological Report Generation
Fuente:
arXiv
Saved in:
| Main Authors: | Molino, Daniele, di Feola, Francesco, Shen, Linlin, Soda, Paolo, Guarrasi, Valerio |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Evaluating Vision Language Model Adaptations for Radiology Report Generation in Low-Resource Languages
by: Salmè, Marco, et al.
Published: (2025)
by: Salmè, Marco, et al.
Published: (2025)
Sample-Aware Test-Time Adaptation for Medical Image-to-Image Translation
by: Iele, Irene, et al.
Published: (2025)
by: Iele, Irene, et al.
Published: (2025)
Text-to-CT Generation via 3D Latent Diffusion Model with Contrastive Vision-Language Pretraining
by: Molino, Daniele, et al.
Published: (2025)
by: Molino, Daniele, et al.
Published: (2025)
Retrieval-Augmented Anatomical Guidance for Text-to-CT Generation
by: Molino, Daniele, et al.
Published: (2026)
by: Molino, Daniele, et al.
Published: (2026)
Concept-Enhanced Multimodal RAG: Towards Interpretable and Accurate Radiology Report Generation
by: Salmè, Marco, et al.
Published: (2026)
by: Salmè, Marco, et al.
Published: (2026)
Multi-Scale Texture Loss for CT denoising with GANs
by: Di Feola, Francesco, et al.
Published: (2024)
by: Di Feola, Francesco, et al.
Published: (2024)
Resilient Vision-Tabular Multimodal Learning under Modality Missingness
by: Caruso, Camillo Maria, et al.
Published: (2026)
by: Caruso, Camillo Maria, et al.
Published: (2026)
Benchmarking Foundation Models and Parameter-Efficient Fine-Tuning for Prognosis Prediction in Medical Imaging
by: Ruffini, Filippo, et al.
Published: (2025)
by: Ruffini, Filippo, et al.
Published: (2025)
Texture-Aware StarGAN for CT data harmonisation
by: Di Feola, Francesco, et al.
Published: (2025)
by: Di Feola, Francesco, et al.
Published: (2025)
Multi-objective optimization determines when, which and how to fuse deep networks: an application to predict COVID-19 outcomes
by: Guarrasi, Valerio, et al.
Published: (2022)
by: Guarrasi, Valerio, et al.
Published: (2022)
Can Foundation Models Really Segment Tumors? A Benchmarking Odyssey in Lung CT Imaging
by: Ayllón, Elena Mulero, et al.
Published: (2025)
by: Ayllón, Elena Mulero, et al.
Published: (2025)
Timing Is Everything: Finding the Optimal Fusion Points in Multimodal Medical Imaging
by: Guarrasi, Valerio, et al.
Published: (2025)
by: Guarrasi, Valerio, et al.
Published: (2025)
Beyond a Single Mode: GAN Ensembles for Diverse Medical Data Generation
by: Tronchin, Lorenzo, et al.
Published: (2025)
by: Tronchin, Lorenzo, et al.
Published: (2025)
Class Balancing Diversity Multimodal Ensemble for Alzheimer's Disease Diagnosis and Early Detection
by: Francesconi, Arianna, et al.
Published: (2024)
by: Francesconi, Arianna, et al.
Published: (2024)
Multimodal Stepwise Clinically-Guided Attention Learning for Pathological Complete Response Prediction in Breast Cancer
by: Caragliano, Alice Natalina, et al.
Published: (2026)
by: Caragliano, Alice Natalina, et al.
Published: (2026)
Beyond the Generative Learning Trilemma: Generative Model Assessment in Data Scarcity Domains
by: Salmè, Marco, et al.
Published: (2025)
by: Salmè, Marco, et al.
Published: (2025)
Longitudinal NSCLC Treatment Progression via Multimodal Generative Models
by: Mantegna, Massimiliano, et al.
Published: (2026)
by: Mantegna, Massimiliano, et al.
Published: (2026)
XGeM: A Multi-Prompt Foundation Model for Multimodal Medical Data Generation
by: Molino, Daniele, et al.
Published: (2025)
by: Molino, Daniele, et al.
Published: (2025)
A Systematic Benchmark of GAN Architectures for MRI-to-CT Synthesis
by: Pesci, Alessandro, et al.
Published: (2026)
by: Pesci, Alessandro, et al.
Published: (2026)
Augmented Intelligence for Multimodal Virtual Biopsy in Breast Cancer Using Generative Artificial Intelligence
by: Rofena, Aurora, et al.
Published: (2025)
by: Rofena, Aurora, et al.
Published: (2025)
Virtual Scanning for NSCLC Histology: Investigating the Discriminatory Power of Synthetic PET
by: Aksu, Fatih, et al.
Published: (2026)
by: Aksu, Fatih, et al.
Published: (2026)
Multi-Dataset Multi-Task Learning for COVID-19 Prognosis
by: Ruffini, Filippo, et al.
Published: (2024)
by: Ruffini, Filippo, et al.
Published: (2024)
Depth Any Canopy: Leveraging Depth Foundation Models for Canopy Height Estimation
by: Cambrin, Daniele Rege, et al.
Published: (2024)
by: Cambrin, Daniele Rege, et al.
Published: (2024)
AnySynth: Harnessing the Power of Image Synthetic Data Generation for Generalized Vision-Language Tasks
by: Li, You, et al.
Published: (2024)
by: Li, You, et al.
Published: (2024)
Lesion-Aware Generative Artificial Intelligence for Virtual Contrast-Enhanced Mammography in Breast Cancer
by: Rofena, Aurora, et al.
Published: (2025)
by: Rofena, Aurora, et al.
Published: (2025)
UniM: A Unified Any-to-Any Interleaved Multimodal Benchmark
by: Li, Yanlin, et al.
Published: (2026)
by: Li, Yanlin, et al.
Published: (2026)
Unaligning Everything: Or Aligning Any Text to Any Image in Multimodal Models
by: Salman, Shaeke, et al.
Published: (2024)
by: Salman, Shaeke, et al.
Published: (2024)
AnySR: Realizing Image Super-Resolution as Any-Scale, Any-Resource
by: Zhan, Wengyi, et al.
Published: (2024)
by: Zhan, Wengyi, et al.
Published: (2024)
HorizonForge: Driving Scene Editing with Any Trajectories and Any Vehicles
by: Wang, Yifan, et al.
Published: (2026)
by: Wang, Yifan, et al.
Published: (2026)
AnyTrans: Translate AnyText in the Image with Large Scale Models
by: Qian, Zhipeng, et al.
Published: (2024)
by: Qian, Zhipeng, et al.
Published: (2024)
DiffusionVL: Translating Any Autoregressive Models into Diffusion Vision Language Models
by: Zeng, Lunbin, et al.
Published: (2025)
by: Zeng, Lunbin, et al.
Published: (2025)
Spider: Any-to-Many Multimodal LLM
by: Lai, Jinxiang, et al.
Published: (2024)
by: Lai, Jinxiang, et al.
Published: (2024)
AnyI2V: Animating Any Conditional Image with Motion Control
by: Li, Ziye, et al.
Published: (2025)
by: Li, Ziye, et al.
Published: (2025)
AnyEdit: Mastering Unified High-Quality Image Editing for Any Idea
by: Yu, Qifan, et al.
Published: (2024)
by: Yu, Qifan, et al.
Published: (2024)
Any-to-Any Learning in Computational Pathology via Triplet Multimodal Pretraining
by: Sun, Qichen, et al.
Published: (2025)
by: Sun, Qichen, et al.
Published: (2025)
Any2Any: Incomplete Multimodal Retrieval with Conformal Prediction
by: Li, Po-han, et al.
Published: (2024)
by: Li, Po-han, et al.
Published: (2024)
X-Pose: Detecting Any Keypoints
by: Yang, Jie, et al.
Published: (2023)
by: Yang, Jie, et al.
Published: (2023)
Any-Shift Prompting for Generalization over Distributions
by: Xiao, Zehao, et al.
Published: (2024)
by: Xiao, Zehao, et al.
Published: (2024)
Segment Any Events with Language
by: Lee, Seungjun, et al.
Published: (2026)
by: Lee, Seungjun, et al.
Published: (2026)
AnyMo: Scaling Any-Modality Conditional Motion Generation with Masked Modeling
by: Li, Yiheng, et al.
Published: (2026)
by: Li, Yiheng, et al.
Published: (2026)
Similar Items
-
Evaluating Vision Language Model Adaptations for Radiology Report Generation in Low-Resource Languages
by: Salmè, Marco, et al.
Published: (2025) -
Sample-Aware Test-Time Adaptation for Medical Image-to-Image Translation
by: Iele, Irene, et al.
Published: (2025) -
Text-to-CT Generation via 3D Latent Diffusion Model with Contrastive Vision-Language Pretraining
by: Molino, Daniele, et al.
Published: (2025) -
Retrieval-Augmented Anatomical Guidance for Text-to-CT Generation
by: Molino, Daniele, et al.
Published: (2026) -
Concept-Enhanced Multimodal RAG: Towards Interpretable and Accurate Radiology Report Generation
by: Salmè, Marco, et al.
Published: (2026)