Automatic Speech Recognition of Finnish-Swedish Dialects: A Comparison of Three Cutting-Edge Technologies
摘要
This paper explores the performance of two different automatic speech recognition models for the Finnish-Swedish language. The first model, Whisper V1 released by OpenAI and the second, the KBLab model trained using a large dataset by the National Library of Sweden. These models were trained initially using data from the Swedish language from Sweden, and the results were compared with previous work trained using a dataset of Finnish-Swedish audio. Our results indicate that general models perform at the same level, opening up the possibility of using these in Finland for the inclusion of the Finnish-Swedish minority.