Reply to comment on “Feasibility of improving vocal fold pathology image classification with synthetic images generated by DDPM-based GenAI: a pilot study”, by Daungsupawong and Wiwanitkit
摘要
In response to the recent commentary on our pilot study using denoising diffusion probabilistic models (DDPMs) for vocal fold structural pathology image augmentation, we clarify our study’s rationale and methods. The original work evaluated the feasibility of synthetic image generation to address data scarcity and imbalance. We affirm that several concerns raised were already addressed and note that statistical testing was not conducted due to the study’s exploratory nature. Additional clarifications are provided in this letter to address points raised in the commentary and to outline directions for future research.