Background <p>Accurate landmark localization is important for three-dimensional (3D) cephalometric analysis. Although deep learning has shown promising performance for 3D landmark localization, the high computational burden of processing volumetric data remains a challenge. The 2.5D networks have emerged to provide the good performance while mitigating computational and memory requirements in the medical domain. Therefore, we compared the performance of 2D, 2.5D and 3D network-based landmark localization.</p> Methods <p>We collected landmark datasets from the volumetric computed tomography (CT) scans of 40 patients. We implemented the 2D, 2.5D and 3D networks for 3D landmark localization. Additionally, we designed a global-to-local loss to mitigate foreground-background imbalance, and employed both soft and hard voting in a network ensemble to improve the robustness. We evaluated each network’s performance in terms of accuracy and computational load.</p> Results <p>The 2.5D network-based landmark localization achieved a mean radial error (MRE) of 1.19<InlineEquation ID="IEq1"> <EquationSource Format="TEX">\(\:\pm\:\)</EquationSource> </InlineEquation>0.65 <i>mm</i> and a successful detection rate (SDR) of 86.46% at 2<i>mm</i>, with a favorable computational load. These results outperformed those of the 2D and 3D networks. Furthermore, using the global-to-local loss led to higher performance compared to using the global loss alone. Soft voting proved the most robust performance among voting methods for landmark localization.</p> Conclusions <p>Comprehensive experiments demonstrate that the 2.5D network offers an optimal trade-off between computational load and accuracy. These findings highlight the potential for more efficient and reliable 3D cephalometry under limited computational resources.</p>

错误:搜索内容不能为空,请输入英文关键词
错误:关键词超出字数限制,请精简
高级检索

Comparison of 2D, 2.5D, and 3D landmark localization networks for 3D cephalometry in CT images

  • Soo-Min Kim,
  • Min-Hyuk Choi,
  • Seung-Hee Han,
  • Min-Jin Kim,
  • Jo-Eun Kim,
  • Kyung-Hoe Huh,
  • Sam-Sun Lee,
  • Min-Suk Heo,
  • Won-Jin Yi

摘要

Background

Accurate landmark localization is important for three-dimensional (3D) cephalometric analysis. Although deep learning has shown promising performance for 3D landmark localization, the high computational burden of processing volumetric data remains a challenge. The 2.5D networks have emerged to provide the good performance while mitigating computational and memory requirements in the medical domain. Therefore, we compared the performance of 2D, 2.5D and 3D network-based landmark localization.

Methods

We collected landmark datasets from the volumetric computed tomography (CT) scans of 40 patients. We implemented the 2D, 2.5D and 3D networks for 3D landmark localization. Additionally, we designed a global-to-local loss to mitigate foreground-background imbalance, and employed both soft and hard voting in a network ensemble to improve the robustness. We evaluated each network’s performance in terms of accuracy and computational load.

Results

The 2.5D network-based landmark localization achieved a mean radial error (MRE) of 1.19 \(\:\pm\:\) 0.65 mm and a successful detection rate (SDR) of 86.46% at 2mm, with a favorable computational load. These results outperformed those of the 2D and 3D networks. Furthermore, using the global-to-local loss led to higher performance compared to using the global loss alone. Soft voting proved the most robust performance among voting methods for landmark localization.

Conclusions

Comprehensive experiments demonstrate that the 2.5D network offers an optimal trade-off between computational load and accuracy. These findings highlight the potential for more efficient and reliable 3D cephalometry under limited computational resources.