Your browser doesn't support javascript.
loading
Adaptive selection of local and non-local attention mechanisms for speech enhancement.
Xu, Xinmeng; Tu, Weiping; Yang, Yuhong.
Afiliação
  • Xu X; National Engineering Research Center for Multimedia Software, School of Computer Science, Wuhan University, China.
  • Tu W; National Engineering Research Center for Multimedia Software, School of Computer Science, Wuhan University, China; Hubei Luojia Laboratory, China; Hubei Key Laboratory of Multimedia and Network Communication Engineering, Wuhan University, China. Electronic address: tuweiping@whu.edu.cn.
  • Yang Y; National Engineering Research Center for Multimedia Software, School of Computer Science, Wuhan University, China; Hubei Key Laboratory of Multimedia and Network Communication Engineering, Wuhan University, China. Electronic address: yangyuhong@whu.edu.cn.
Neural Netw ; 174: 106236, 2024 Jun.
Article em En | MEDLINE | ID: mdl-38518710
ABSTRACT
In speech enhancement tasks, local and non-local attention mechanisms have been significantly improved and well studied. However, a natural speech signal contains many dynamic and fast-changing acoustic features, and focusing on one type of attention mechanism (local or non-local) cannot precisely capture the most discriminative information for estimating target speech from background interference. To address this issue, we introduce an adaptive selection network to dynamically select an appropriate route that determines whether to use the attention mechanisms and which to use for the task. We train the adaptive selection network using reinforcement learning with a developed difficulty-adjusted reward that is related to the performance, complexity, and difficulty of target speech estimation from the noisy mixtures. Consequently, we propose an Attention Selection Speech Enhancement Network (ASSENet) with the innovative dynamic block that consists of an adaptive selection network and a local and non-local attention based speech enhancement network. In particular, the ASSENet incorporates both local and non-local attention and develops the attention mechanism selection technique to explore the appropriate route of local and non-local attention mechanisms for speech enhancement tasks. The results show that our method achieves comparable and superior performance to existing approaches with attractive computational costs.
Assuntos
Palavras-chave

Texto completo: 1 Coleções: 01-internacional Base de dados: MEDLINE Assunto principal: Fala / Aprendizagem Idioma: En Revista: Neural Netw / Neural netw / Neural networks Assunto da revista: NEUROLOGIA Ano de publicação: 2024 Tipo de documento: Article País de afiliação: China País de publicação: Estados Unidos

Texto completo: 1 Coleções: 01-internacional Base de dados: MEDLINE Assunto principal: Fala / Aprendizagem Idioma: En Revista: Neural Netw / Neural netw / Neural networks Assunto da revista: NEUROLOGIA Ano de publicação: 2024 Tipo de documento: Article País de afiliação: China País de publicação: Estados Unidos