Text-to-CT Generation via 3D Latent Diffusion Model with Contrastive Vision-Language Pretraining
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Molino, Daniele, Caruso, Camillo Maria, Ruffini, Filippo, Soda, Paolo, Guarrasi, Valerio |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Retrieval-Augmented Anatomical Guidance for Text-to-CT Generation
von: Molino, Daniele, et al.
Veröffentlicht: (2026)
von: Molino, Daniele, et al.
Veröffentlicht: (2026)
Resilient Vision-Tabular Multimodal Learning under Modality Missingness
von: Caruso, Camillo Maria, et al.
Veröffentlicht: (2026)
von: Caruso, Camillo Maria, et al.
Veröffentlicht: (2026)
Any-to-Any Vision-Language Model for Multimodal X-ray Imaging and Radiological Report Generation
von: Molino, Daniele, et al.
Veröffentlicht: (2025)
von: Molino, Daniele, et al.
Veröffentlicht: (2025)
Evaluating Vision Language Model Adaptations for Radiology Report Generation in Low-Resource Languages
von: Salmè, Marco, et al.
Veröffentlicht: (2025)
von: Salmè, Marco, et al.
Veröffentlicht: (2025)
Benchmarking Foundation Models and Parameter-Efficient Fine-Tuning for Prognosis Prediction in Medical Imaging
von: Ruffini, Filippo, et al.
Veröffentlicht: (2025)
von: Ruffini, Filippo, et al.
Veröffentlicht: (2025)
Multi-objective optimization determines when, which and how to fuse deep networks: an application to predict COVID-19 outcomes
von: Guarrasi, Valerio, et al.
Veröffentlicht: (2022)
von: Guarrasi, Valerio, et al.
Veröffentlicht: (2022)
MARIA: a Multimodal Transformer Model for Incomplete Healthcare Data
von: Caruso, Camillo Maria, et al.
Veröffentlicht: (2024)
von: Caruso, Camillo Maria, et al.
Veröffentlicht: (2024)
Not Another Imputation Method: A Transformer-based Model for Missing Values in Tabular Datasets
von: Caruso, Camillo Maria, et al.
Veröffentlicht: (2024)
von: Caruso, Camillo Maria, et al.
Veröffentlicht: (2024)
A Systematic Benchmark of GAN Architectures for MRI-to-CT Synthesis
von: Pesci, Alessandro, et al.
Veröffentlicht: (2026)
von: Pesci, Alessandro, et al.
Veröffentlicht: (2026)
Multi-Scale Texture Loss for CT denoising with GANs
von: Di Feola, Francesco, et al.
Veröffentlicht: (2024)
von: Di Feola, Francesco, et al.
Veröffentlicht: (2024)
Beyond a Single Mode: GAN Ensembles for Diverse Medical Data Generation
von: Tronchin, Lorenzo, et al.
Veröffentlicht: (2025)
von: Tronchin, Lorenzo, et al.
Veröffentlicht: (2025)
Multi-Dataset Multi-Task Learning for COVID-19 Prognosis
von: Ruffini, Filippo, et al.
Veröffentlicht: (2024)
von: Ruffini, Filippo, et al.
Veröffentlicht: (2024)
Beyond the Generative Learning Trilemma: Generative Model Assessment in Data Scarcity Domains
von: Salmè, Marco, et al.
Veröffentlicht: (2025)
von: Salmè, Marco, et al.
Veröffentlicht: (2025)
Texture-Aware StarGAN for CT data harmonisation
von: Di Feola, Francesco, et al.
Veröffentlicht: (2025)
von: Di Feola, Francesco, et al.
Veröffentlicht: (2025)
Probabilistic NDVI Forecasting from Sparse Satellite Time Series and Weather Covariates
von: Iele, Irene, et al.
Veröffentlicht: (2026)
von: Iele, Irene, et al.
Veröffentlicht: (2026)
Can Foundation Models Really Segment Tumors? A Benchmarking Odyssey in Lung CT Imaging
von: Ayllón, Elena Mulero, et al.
Veröffentlicht: (2025)
von: Ayllón, Elena Mulero, et al.
Veröffentlicht: (2025)
Lesion-Aware Generative Artificial Intelligence for Virtual Contrast-Enhanced Mammography in Breast Cancer
von: Rofena, Aurora, et al.
Veröffentlicht: (2025)
von: Rofena, Aurora, et al.
Veröffentlicht: (2025)
Sample-Aware Test-Time Adaptation for Medical Image-to-Image Translation
von: Iele, Irene, et al.
Veröffentlicht: (2025)
von: Iele, Irene, et al.
Veröffentlicht: (2025)
Concept-Enhanced Multimodal RAG: Towards Interpretable and Accurate Radiology Report Generation
von: Salmè, Marco, et al.
Veröffentlicht: (2026)
von: Salmè, Marco, et al.
Veröffentlicht: (2026)
A Deep Learning Approach for Overall Survival Prediction in Lung Cancer with Missing Values
von: Caruso, Camillo Maria, et al.
Veröffentlicht: (2023)
von: Caruso, Camillo Maria, et al.
Veröffentlicht: (2023)
Virtual Scanning for NSCLC Histology: Investigating the Discriminatory Power of Synthetic PET
von: Aksu, Fatih, et al.
Veröffentlicht: (2026)
von: Aksu, Fatih, et al.
Veröffentlicht: (2026)
Timing Is Everything: Finding the Optimal Fusion Points in Multimodal Medical Imaging
von: Guarrasi, Valerio, et al.
Veröffentlicht: (2025)
von: Guarrasi, Valerio, et al.
Veröffentlicht: (2025)
Multimodal Stepwise Clinically-Guided Attention Learning for Pathological Complete Response Prediction in Breast Cancer
von: Caragliano, Alice Natalina, et al.
Veröffentlicht: (2026)
von: Caragliano, Alice Natalina, et al.
Veröffentlicht: (2026)
A Deep Learning Approach for Virtual Contrast Enhancement in Contrast Enhanced Spectral Mammography
von: Rofena, Aurora, et al.
Veröffentlicht: (2023)
von: Rofena, Aurora, et al.
Veröffentlicht: (2023)
A Systematic Review of Intermediate Fusion in Multimodal Deep Learning for Biomedical Applications
von: Guarrasi, Valerio, et al.
Veröffentlicht: (2024)
von: Guarrasi, Valerio, et al.
Veröffentlicht: (2024)
Augmented Intelligence for Multimodal Virtual Biopsy in Breast Cancer Using Generative Artificial Intelligence
von: Rofena, Aurora, et al.
Veröffentlicht: (2025)
von: Rofena, Aurora, et al.
Veröffentlicht: (2025)
Leveraging MIMIC Datasets for Better Digital Health: A Review on Open Problems, Progress Highlights, and Future Promises
von: Khaled, Afifa, et al.
Veröffentlicht: (2025)
von: Khaled, Afifa, et al.
Veröffentlicht: (2025)
Learning from Limited and Incomplete Data: A Multimodal Framework for Predicting Pathological Response in NSCLC
von: Caragliano, Alice Natalina, et al.
Veröffentlicht: (2026)
von: Caragliano, Alice Natalina, et al.
Veröffentlicht: (2026)
Class Balancing Diversity Multimodal Ensemble for Alzheimer's Disease Diagnosis and Early Detection
von: Francesconi, Arianna, et al.
Veröffentlicht: (2024)
von: Francesconi, Arianna, et al.
Veröffentlicht: (2024)
Doctor-in-the-Loop: An Explainable, Multi-View Deep Learning Framework for Predicting Pathological Response in Non-Small Cell Lung Cancer
von: Caragliano, Alice Natalina, et al.
Veröffentlicht: (2025)
von: Caragliano, Alice Natalina, et al.
Veröffentlicht: (2025)
Handling Missing Modalities in Multimodal Survival Prediction for Non-Small Cell Lung Cancer
von: Ruffini, Filippo, et al.
Veröffentlicht: (2026)
von: Ruffini, Filippo, et al.
Veröffentlicht: (2026)
Longitudinal NSCLC Treatment Progression via Multimodal Generative Models
von: Mantegna, Massimiliano, et al.
Veröffentlicht: (2026)
von: Mantegna, Massimiliano, et al.
Veröffentlicht: (2026)
CLAP: Contrastive Latent Action Pretraining for Learning Vision-Language-Action Models from Human Videos
von: Zhang, Chubin, et al.
Veröffentlicht: (2026)
von: Zhang, Chubin, et al.
Veröffentlicht: (2026)
Grounded Compositional and Diverse Text-to-3D with Pretrained Multi-View Diffusion Model
von: Li, Xiaolong, et al.
Veröffentlicht: (2024)
von: Li, Xiaolong, et al.
Veröffentlicht: (2024)
Vision-Language Models for Automated 3D PET/CT Report Generation
von: Jiao, Wenpei, et al.
Veröffentlicht: (2025)
von: Jiao, Wenpei, et al.
Veröffentlicht: (2025)
Anatomy-Aware Low-Dose CT Denoising via Pretrained Vision Models and Semantic-Guided Contrastive Learning
von: Wang, Runze, et al.
Veröffentlicht: (2025)
von: Wang, Runze, et al.
Veröffentlicht: (2025)
Hierarchical Vision-Language Alignment for Text-to-Image Generation via Diffusion Models
von: Johnson, Emily, et al.
Veröffentlicht: (2025)
von: Johnson, Emily, et al.
Veröffentlicht: (2025)
Prometheus: 3D-Aware Latent Diffusion Models for Feed-Forward Text-to-3D Scene Generation
von: Yang, Yuanbo, et al.
Veröffentlicht: (2024)
von: Yang, Yuanbo, et al.
Veröffentlicht: (2024)
GenerateCT: Text-Conditional Generation of 3D Chest CT Volumes
von: Hamamci, Ibrahim Ethem, et al.
Veröffentlicht: (2023)
von: Hamamci, Ibrahim Ethem, et al.
Veröffentlicht: (2023)
Learning to Read Where to Look: Disease-Aware Vision-Language Pretraining for 3D CT
von: Ging, Simon, et al.
Veröffentlicht: (2026)
von: Ging, Simon, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Retrieval-Augmented Anatomical Guidance for Text-to-CT Generation
von: Molino, Daniele, et al.
Veröffentlicht: (2026) -
Resilient Vision-Tabular Multimodal Learning under Modality Missingness
von: Caruso, Camillo Maria, et al.
Veröffentlicht: (2026) -
Any-to-Any Vision-Language Model for Multimodal X-ray Imaging and Radiological Report Generation
von: Molino, Daniele, et al.
Veröffentlicht: (2025) -
Evaluating Vision Language Model Adaptations for Radiology Report Generation in Low-Resource Languages
von: Salmè, Marco, et al.
Veröffentlicht: (2025) -
Benchmarking Foundation Models and Parameter-Efficient Fine-Tuning for Prognosis Prediction in Medical Imaging
von: Ruffini, Filippo, et al.
Veröffentlicht: (2025)