Causal Disentanglement for Robust Long-tail Medical Image Generation

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Nie, Weizhi, Zhang, Zichun, Wang, Weijie, Lepri, Bruno, Liu, Anan, Sebe, Nicu
Format: Preprint
Published: 2025
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!
_version_ 1866916704945176576
author Nie, Weizhi
Zhang, Zichun
Wang, Weijie
Lepri, Bruno
Liu, Anan
Sebe, Nicu
author_facet Nie, Weizhi
Zhang, Zichun
Wang, Weijie
Lepri, Bruno
Liu, Anan
Sebe, Nicu
contents Counterfactual medical image generation effectively addresses data scarcity and enhances the interpretability of medical images. However, due to the complex and diverse pathological features of medical images and the imbalanced class distribution in medical data, generating high-quality and diverse medical images from limited data is significantly challenging. Additionally, to fully leverage the information in limited data, such as anatomical structure information and generate more structurally stable medical images while avoiding distortion or inconsistency. In this paper, in order to enhance the clinical relevance of generated data and improve the interpretability of the model, we propose a novel medical image generation framework, which generates independent pathological and structural features based on causal disentanglement and utilizes text-guided modeling of pathological features to regulate the generation of counterfactual images. First, we achieve feature separation through causal disentanglement and analyze the interactions between features. Here, we introduce group supervision to ensure the independence of pathological and identity features. Second, we leverage a diffusion model guided by pathological findings to model pathological features, enabling the generation of diverse counterfactual images. Meanwhile, we enhance accuracy by leveraging a large language model to extract lesion severity and location from medical reports. Additionally, we improve the performance of the latent diffusion model on long-tailed categories through initial noise optimization.
format Preprint
id arxiv_https___arxiv_org_abs_2504_14450
institution arXiv
publishDate 2025
record_format arxiv
spellingShingle Causal Disentanglement for Robust Long-tail Medical Image Generation
Nie, Weizhi
Zhang, Zichun
Wang, Weijie
Lepri, Bruno
Liu, Anan
Sebe, Nicu
Computer Vision and Pattern Recognition
Counterfactual medical image generation effectively addresses data scarcity and enhances the interpretability of medical images. However, due to the complex and diverse pathological features of medical images and the imbalanced class distribution in medical data, generating high-quality and diverse medical images from limited data is significantly challenging. Additionally, to fully leverage the information in limited data, such as anatomical structure information and generate more structurally stable medical images while avoiding distortion or inconsistency. In this paper, in order to enhance the clinical relevance of generated data and improve the interpretability of the model, we propose a novel medical image generation framework, which generates independent pathological and structural features based on causal disentanglement and utilizes text-guided modeling of pathological features to regulate the generation of counterfactual images. First, we achieve feature separation through causal disentanglement and analyze the interactions between features. Here, we introduce group supervision to ensure the independence of pathological and identity features. Second, we leverage a diffusion model guided by pathological findings to model pathological features, enabling the generation of diverse counterfactual images. Meanwhile, we enhance accuracy by leveraging a large language model to extract lesion severity and location from medical reports. Additionally, we improve the performance of the latent diffusion model on long-tailed categories through initial noise optimization.
title Causal Disentanglement for Robust Long-tail Medical Image Generation
topic Computer Vision and Pattern Recognition
url https://arxiv.org/abs/2504.14450