To protect patient privacy, the ability of diffusion models to generate medical images from noise has become a key focus for enriching datasets. However, due to the high precision required for medical image anatomical structures, generative models designed for natural scenes fail to meet the stringent standards for medical logic. We propose MedMaskDiff, a Mamba-based semantic image synthesis model that generates medical images from masks, which takes full advantage of Mamba’s capability to capture long-range medical features. Additionally, it utilizes an evolutionary condition-guide method to enhance the quality and medical logic of the generated target regions. MedMaskDiff outperforms other advanced methods in synthesizing liver CT, thyroid nodule ultrasound, and low-grade intraepithelial neoplasia microscopy images. By utilizing masks from other liver CT datasets for semantic synthesis and data augmentation, comparative experiments demonstrate that MedMaskDiff effectively safeguards patient privacy while enhancing downstream medical image segmentation tasks, significantly improving the performance of segmentation models. Our code is available on https://github.com/Jiacheng-Han/MedMaskDiff .

错误:搜索内容不能为空,请输入英文关键词
错误:关键词超出字数限制,请精简
高级检索

MedMaskDiff: Mamba-Based Medical Semantic Image Synthesis for Segmentation

  • Jiacheng Han,
  • Ke Niu,
  • Jiuyun Cai

摘要

To protect patient privacy, the ability of diffusion models to generate medical images from noise has become a key focus for enriching datasets. However, due to the high precision required for medical image anatomical structures, generative models designed for natural scenes fail to meet the stringent standards for medical logic. We propose MedMaskDiff, a Mamba-based semantic image synthesis model that generates medical images from masks, which takes full advantage of Mamba’s capability to capture long-range medical features. Additionally, it utilizes an evolutionary condition-guide method to enhance the quality and medical logic of the generated target regions. MedMaskDiff outperforms other advanced methods in synthesizing liver CT, thyroid nodule ultrasound, and low-grade intraepithelial neoplasia microscopy images. By utilizing masks from other liver CT datasets for semantic synthesis and data augmentation, comparative experiments demonstrate that MedMaskDiff effectively safeguards patient privacy while enhancing downstream medical image segmentation tasks, significantly improving the performance of segmentation models. Our code is available on https://github.com/Jiacheng-Han/MedMaskDiff .