Projection Back onto Filtered Observations for Speech Separation with Distributed Microphone Array

In the context of assisted human, identifying and enhancing non-stationary speech targets speech in various noise environments, such as a cocktail party, is an important issue for real-time speech separation. Previous studies mostly used microphone signal processing to perform target speech separation and analysis, such as feature recognition through a large amount of training data and supervised machine learning. The method was suitable for stationary noise suppression, but relatively limited for non-stationary noise and difficult to meet the real-time processing requirement. In this study, we propose a real-time speech separation method based on an approach that combines an optical camera and a microphone array. The method was divided into two stages. Stage 1 used computer vision technology with the camera to detect and identify interest targets and evaluate source angles and distance. Stage 2 used beamforming technology with microphone array to enhance and separate the target speech sound. The asynchronous update function was utilized to integrate the beamforming control and speech processing to reduce the effect of the processing delay. The experimental results show that the noise reduction in various stationary and non-stationary noise environments were 6.1 dB and 5.2 dB respectively. The response time of speech processing was less than 10ms, which meets the requirements of a real-time system. The proposed method has high potential to be applied in auxiliary listening systems or machine language processing like intelligent personal assistant.

Download Full-text

1A1-W08 Investigation on Cooperation of Microphone Array with Depth Image Sensor : For the Speech Separation of the Multiple Humans that Positions are Unsettled

The Proceedings of JSME annual Conference on Robotics and Mechatronics (Robomec) ◽

10.1299/jsmermd.2015._1a1-w08_1 ◽

2015 ◽

Vol 2015 (0) ◽

pp. _1A1-W08_1-_1A1-W08_2

Author(s):

Takahiro KIGAWA ◽

Takeki OGITSU ◽

Hiroshi TAKEMURA ◽

Hiroshi MIZOGUCHI

Keyword(s):

Microphone Array ◽

Image Sensor ◽

Depth Image ◽

Speech Separation

Download Full-text

Blind speech separation by integrating three pairs of Phase Differences of equilateral triangular microphone array

2010 International Symposium on Intelligent Signal Processing and Communication Systems ◽

10.1109/ispacs.2010.5704649 ◽

2010 ◽

Author(s):

Masashi Yoshida ◽

Ding Ning ◽

Nozomu Hamada

Keyword(s):

Microphone Array ◽

Speech Separation ◽

Phase Differences

Download Full-text

A microphone array beamforming-based system for multi-talker speech separation

International Journal of Signal and Imaging Systems Engineering ◽

10.1504/ijsise.2016.078257 ◽

2016 ◽

Vol 9 (4/5) ◽

pp. 209

Author(s):

Adel Hidri ◽

Hamid Amiri

Keyword(s):

Microphone Array ◽

Speech Separation ◽

Array Beamforming

Download Full-text

Microphone Array Beamforming Approach to Blind Speech Separation

Machine Learning for Multimodal Interaction - Lecture Notes in Computer Science ◽

10.1007/978-3-540-78155-4_26 ◽

2008 ◽

pp. 295-305 ◽

Cited By ~ 6

Author(s):

Ivan Himawan ◽

Iain McCowan ◽

Mike Lincoln

Keyword(s):

Microphone Array ◽

Speech Separation ◽

Array Beamforming

Download Full-text

Microphone Array Speech Separation Algorithm Based on TC-ResNet

Computers Materials & Continua ◽

10.32604/cmc.2021.017080 ◽

2021 ◽

Vol 69 (2) ◽

pp. 2705-2716

Author(s):

Lin Zhou ◽

Yue Xu ◽

Tianyi Wang ◽

Kun Feng ◽

Jingang Shi

Keyword(s):

Microphone Array ◽

Speech Separation ◽

Separation Algorithm

Download Full-text

Speech separation microphone array based on law of causality and frequency domain processing

2009 9th International Symposium on Communications and Information Technology ◽

10.1109/iscit.2009.5341153 ◽

2009 ◽

Author(s):

Takuto Yoshioka ◽

Kensaku Hujii ◽

Mitsuji Muneyasu

Keyword(s):

Frequency Domain ◽

Microphone Array ◽

Speech Separation ◽

Frequency Domain Processing

Download Full-text

Adaptive Speech Separation Based on Beamforming and Frequency Domain-Independent Component Analysis

Applied Sciences ◽

10.3390/app10072593 ◽

2020 ◽

Vol 10 (7) ◽

pp. 2593

Author(s):

Ke Zhang ◽

Yangjie Wei ◽

Dan Wu ◽

Yi Wang

Keyword(s):

Independent Component Analysis ◽

Frequency Domain ◽

Source Localization ◽

Sound Source ◽

Microphone Array ◽

Component Analysis ◽

Independent Component ◽

Sound Source Localization ◽

Speech Separation ◽

Sound Sources

Voice signals acquired by a microphone array often include considerable noise and mutual interference, seriously degrading the accuracy and speed of speech separation. Traditional beamforming is simple to implement, but its source interference suppression is not adequate. In contrast, independent component analysis (ICA) can improve separation, but imposes an iterative and time-consuming process to calculate the separation matrix. As a supporting method, principle component analysis (PCA) contributes to reduce the dimension, retrieve fast results, and disregard false sound sources. Considering the sparsity of frequency components in a mixed signal, we propose an adaptive fast speech separation algorithm based on multiple sound source localization as preprocessing to select between beamforming and frequency domain ICA according to different mixing conditions per frequency bin. First, a fast positioning algorithm allows calculating the maximum number of components per frequency bin of a mixed speech signal to prevent the occurrence of false sound sources. Then, PCA reduces the dimension to adaptively adjust the weight of beamforming and ICA for speech separation. Subsequently, the ICA separation matrix is initialized based on the sound source localization to notably reduce the iteration time and mitigate permutation ambiguity. Simulation and experimental results verify the effectiveness and speedup of the proposed algorithm.

Download Full-text