<p>Speech data inherently contains personally identifiable information. Anonymization strategies to obscure this while preserving essential characteristics all represent a tradeoff between privacy and utility. We examine this balancing act of modifying voice characteristics, masking identity, and eliminating identifiable content by showcasing challenges with the common techniques—generalization, suppression, anatomization, permutation, and perturbation—in the context of preserving utility for individual level speech data analyses in clinical research.</p>

错误:搜索内容不能为空,请输入英文关键词
错误:关键词超出字数限制,请精简
高级检索

Navigating the tradeoff between personal privacy and data utility in speech anonymization for clinical research

  • Catherine Diaz-Asper,
  • Lars Ailo Bongo,
  • Brita Elvevåg

摘要

Speech data inherently contains personally identifiable information. Anonymization strategies to obscure this while preserving essential characteristics all represent a tradeoff between privacy and utility. We examine this balancing act of modifying voice characteristics, masking identity, and eliminating identifiable content by showcasing challenges with the common techniques—generalization, suppression, anatomization, permutation, and perturbation—in the context of preserving utility for individual level speech data analyses in clinical research.