Enhancing Video Colorization with Deep Learning: A Comprehensive Analysis of Training Loss Functions
摘要
Traditional colorization approaches rely on the expertise of artists or researchers that meticulously paint or digitally add colors to an image (or video frames), which is often a time-consuming, laborious, and error-prone task. Automatic methods, based on deep learning techniques, have replaced such approaches to colorization. Despite the advances toward improving their accuracy, there is no consensus regarding the best training procedures for existing artificial neural network techniques proposed for video colorization. In this paper, in order to fill this gap, we focus on the impact of selecting an appropriate loss function. We investigate seven loss functions to find the combination that gives the best results with Deep Learning Video Colorization (DLVC) using a U-Net topology and an attention mechanism trained on the DAVIS dataset. An investigation of the current validation metrics for colorization results was also conducted to analyze their ability to accurately judge colors between frames.