Leveraging Pre-trained Models for FF-to-FFPE Histopathological Image Translation

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Zhang, Qilai, Li, Jiawen, Liao, Peiran, Hu, Jiali, Guan, Tian, Han, Anjia, He, Yonghong
Format: Preprint
Published: 2024
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!
_version_ 1866915016976891904
author Zhang, Qilai
Li, Jiawen
Liao, Peiran
Hu, Jiali
Guan, Tian
Han, Anjia
He, Yonghong
author_facet Zhang, Qilai
Li, Jiawen
Liao, Peiran
Hu, Jiali
Guan, Tian
Han, Anjia
He, Yonghong
contents The two primary types of Hematoxylin and Eosin (H&E) slides in histopathology are Formalin-Fixed Paraffin-Embedded (FFPE) and Fresh Frozen (FF). FFPE slides offer high quality histopathological images but require a labor-intensive acquisition process. In contrast, FF slides can be prepared quickly, but the image quality is relatively poor. Our task is to translate FF images into FFPE style, thereby improving the image quality for diagnostic purposes. In this paper, we propose Diffusion-FFPE, a method for FF-to-FFPE histopathological image translation using a pre-trained diffusion model. Specifically, we utilize a one-step diffusion model as the generator, which we fine-tune using LoRA adapters within an adversarial learning framework. To enable the model to effectively capture both global structural patterns and local details, we introduce a multi-scale feature fusion module that leverages two VAE encoders to extract features at different image resolutions, performing feature fusion before inputting them into the UNet. Additionally, a pre-trained vision-language model for histopathology serves as the backbone for the discriminator, enhancing model performance. Our FF-to-FFPE translation experiments on the TCGA-NSCLC dataset demonstrate that the proposed approach outperforms existing methods. The code and models are released at https://github.com/QilaiZhang/Diffusion-FFPE.
format Preprint
id arxiv_https___arxiv_org_abs_2406_18054
institution arXiv
publishDate 2024
record_format arxiv
spellingShingle Leveraging Pre-trained Models for FF-to-FFPE Histopathological Image Translation
Zhang, Qilai
Li, Jiawen
Liao, Peiran
Hu, Jiali
Guan, Tian
Han, Anjia
He, Yonghong
Image and Video Processing
Computer Vision and Pattern Recognition
The two primary types of Hematoxylin and Eosin (H&E) slides in histopathology are Formalin-Fixed Paraffin-Embedded (FFPE) and Fresh Frozen (FF). FFPE slides offer high quality histopathological images but require a labor-intensive acquisition process. In contrast, FF slides can be prepared quickly, but the image quality is relatively poor. Our task is to translate FF images into FFPE style, thereby improving the image quality for diagnostic purposes. In this paper, we propose Diffusion-FFPE, a method for FF-to-FFPE histopathological image translation using a pre-trained diffusion model. Specifically, we utilize a one-step diffusion model as the generator, which we fine-tune using LoRA adapters within an adversarial learning framework. To enable the model to effectively capture both global structural patterns and local details, we introduce a multi-scale feature fusion module that leverages two VAE encoders to extract features at different image resolutions, performing feature fusion before inputting them into the UNet. Additionally, a pre-trained vision-language model for histopathology serves as the backbone for the discriminator, enhancing model performance. Our FF-to-FFPE translation experiments on the TCGA-NSCLC dataset demonstrate that the proposed approach outperforms existing methods. The code and models are released at https://github.com/QilaiZhang/Diffusion-FFPE.
title Leveraging Pre-trained Models for FF-to-FFPE Histopathological Image Translation
topic Image and Video Processing
Computer Vision and Pattern Recognition
url https://arxiv.org/abs/2406.18054