AES E-Library

Distant Speech Beamforming Improving Multi-User ASR

In this paper an algorithm that improves side speaker attenuation for super directive beamformer like MVDR (Minimum Variance Distortionless Response) is presented. This technique can be utilized in a scenario where there are multiple people in a room intending to interact with an ASR- (automatic speech recognition) enabled device, e.g., smart speaker. The experiments show that the proposed solution gives a reduction of WER (word error rate) up to 23.93% calculated for command uttered by one user when a second user was treated as the side speaker.

 

Author (s):
Affiliation: (See document for exact affiliation information.)
AES Convention: Paper Number:
Publication Date:
Session subject:
Permalink: https://aes2.org/publications/elibrary-page/?id=19542


(1502KB)


Click to purchase paper as a non-member or login as an AES member. If your company or school subscribes to the E-Library then switch to the institutional version. If you are not an AES member Join the AES. If you need to check your member status, login to the Member Portal.

Type:
E-Libary location:
16938
Choose your country of residence from this list:










Skip to content