Direction-of-arrival (DOA) estimation based on microphone arrays has been a hot research topic in recent years. Transfer function (TF) based DOA method performs well because it considers both time difference and intensity difference. However, obtaining transfer function is a difficult task and transfer function based method is susceptible to noise. In this paper, an autoencoder network structure is proposed for DOA estimation task. The network is used to learn the characteristics of the transfer function, which considers both time difference information and intensity difference information for DOA estimation. The proposed unsupervised training method helps minimize the burden for labeling training data. The evaluation experiments show that our method performs better than TF-based method in the noisy environment.
Authors:
Wang, Yiwen; Wu, Xihong; Qu, Tianshu
Affiliation:
Peking University
AES Convention:
148 (May 2020)
Paper Number:
10370
Publication Date:
May 28, 2020
Subject:
Spatial Audio
Click to purchase paper as a non-member or you can login as an AES member to see more options.
No AES members have commented on this paper yet.
To be notified of new comments on this paper you can
subscribe to this RSS feed.
Forum users should login to see additional options.
If you are not yet an AES member and have something important to say about this paper then we urge you to join the AES today and make your voice heard. You can join online today by clicking here.