Speech Enhancement using Multi-Microphone Array based Source Separation and Deep Learning

2022 International Conference on Smart Generation Computing, Communication and Networking (SMART GENCON)(2022)

引用 0|浏览0
暂无评分
摘要
Speech Enhancement using tensor decomposition-based source separation and convolutional, bidirectional recurrent neural network (CNN-biRNN) architecture is investigated in this paper. An acoustic receiver comprising uniform linear array (ULA) of microphone sensors is considered, where the ULA performs CANDECOMP/PARAFAC (CP) tensor decomposition to separate the individual speech source signals from the received mixture of multi-channel signals, followed by single channel de-reverberation by a variant of the CNN-biRNN referred to as DenseNet-biLSTM to enhance the target speech signal-of-interest (SOI). While the source separation module based on CP-tensor decomposition is responsible for extracting the target SOI, the subsequent deep learning framework based on DenseNet-biLSTM enhances the extracted SOI by performing de-noising and de-reverberation. It is demonstrated by computer simulations that the proposed approach leads to good performance under multiple interfering speakers and reverberation.
更多
查看译文
关键词
Speech Enhancement,Reverberation,PARAFAC,Microphone Array,DenseNet-LSTM
AI 理解论文
溯源树
样例
生成溯源树,研究论文发展脉络
Chat Paper
正在生成论文摘要