<p>Balancing data sharing and patient privacy is essential in medical imaging. Face redaction tools anonymize head CTs by removing identifiable features, but their impact on deep learning (DL) model performance remains a concern. We present an open-source face redaction tool designed to enhance data-sharing security while preserving DL performance, validated through a Kaggle competition on age prediction from brain CTs. This study aims to evaluate whether models trained on redacted images perform comparably on both redacted and non-redacted test sets, and how they compare to similar models trained on non-redacted images reported in the literature. A Kaggle challenge was conducted between March 2 and April 30, 2024, to crowdsource age prediction models. The dataset comprised 2377 redacted head CT studies for training and 148 for testing, sourced from multiple institutions. In a post-hoc analysis, the top-performing models were evaluated on both redacted and non-redacted formats of the test set, with performance measured using mean absolute error (MAE). The two best models achieved MAEs of 2.8 and 3.4 years on redacted test data. On the non-redacted format, MAEs increased to 3.2 and 3.8, respectively. Paired <i>t</i>-tests showed a significant performance drop for one model (<i>p</i> = 0.038) but not the other (<i>p</i> = 0.051). There was no significant difference between the models on the redacted test set (<i>p</i> = 0.610). Models trained on redacted data may show minimal performance decline when applied to non-redacted images, yet still outperform existing benchmarks. Our tool enables secure data sharing with a limited impact on DL accuracy.</p>

错误:搜索内容不能为空,请输入英文关键词
错误:关键词超出字数限制,请精简
高级检索

Pixel Tampering: Does Face Redaction Harm Medical AI Performance?

  • Eduardo M. J. M. Farina,
  • Felipe A. Matsuoka,
  • Gustavo Corradi,
  • Yosuke Yamagishi,
  • Masatoshi Abe,
  • Maximilian Pfeiffer,
  • Andrea S. Souza,
  • Raquel Moreno,
  • Ivanei Bramati,
  • Fernanda Moll,
  • Almir Bitencourt,
  • Carlos Sacomani,
  • Soraia Quaranta Damião,
  • Rubens Chojniak,
  • Nitamar Abdala,
  • Rodrigo Ragazzini,
  • Henrique Carrete Jr.,
  • Paulo E. A. Kuriki,
  • Marcelo Straus Takahashi,
  • Nelson Caserta,
  • Cesar H. Nomura,
  • Felipe C. Kitamura

摘要

Balancing data sharing and patient privacy is essential in medical imaging. Face redaction tools anonymize head CTs by removing identifiable features, but their impact on deep learning (DL) model performance remains a concern. We present an open-source face redaction tool designed to enhance data-sharing security while preserving DL performance, validated through a Kaggle competition on age prediction from brain CTs. This study aims to evaluate whether models trained on redacted images perform comparably on both redacted and non-redacted test sets, and how they compare to similar models trained on non-redacted images reported in the literature. A Kaggle challenge was conducted between March 2 and April 30, 2024, to crowdsource age prediction models. The dataset comprised 2377 redacted head CT studies for training and 148 for testing, sourced from multiple institutions. In a post-hoc analysis, the top-performing models were evaluated on both redacted and non-redacted formats of the test set, with performance measured using mean absolute error (MAE). The two best models achieved MAEs of 2.8 and 3.4 years on redacted test data. On the non-redacted format, MAEs increased to 3.2 and 3.8, respectively. Paired t-tests showed a significant performance drop for one model (p = 0.038) but not the other (p = 0.051). There was no significant difference between the models on the redacted test set (p = 0.610). Models trained on redacted data may show minimal performance decline when applied to non-redacted images, yet still outperform existing benchmarks. Our tool enables secure data sharing with a limited impact on DL accuracy.