We propose a multi-task learning approach that integrates Automatic Speech Recognition (ASR) and Accent Recognition (AR) tasks. By utilizing a shared ASR encoder alongside a Transformer-based AR decoder, our model enhances the extraction of relevant accent features, overcoming challenges posed by non-joint frameworks. With contrastive learning, geographical information is effectively leveraged. Experimental results validate the effectiveness of our proposed method, demonstrating improved accent classification performance and indicate the features important for AR task.

错误:搜索内容不能为空,请输入英文关键词
错误:关键词超出字数限制,请精简
高级检索

Accent Recognition with Auxiliary Task and Contrastive Learning

  • Shutao Liu,
  • Yanping Li,
  • Xi Shao

摘要

We propose a multi-task learning approach that integrates Automatic Speech Recognition (ASR) and Accent Recognition (AR) tasks. By utilizing a shared ASR encoder alongside a Transformer-based AR decoder, our model enhances the extraction of relevant accent features, overcoming challenges posed by non-joint frameworks. With contrastive learning, geographical information is effectively leveraged. Experimental results validate the effectiveness of our proposed method, demonstrating improved accent classification performance and indicate the features important for AR task.