跳到主要导航 跳到搜索 跳到主要内容

Auditory filterbanks benefit universal sound source separation

  • Northwestern Polytechnical University Xian
  • Technical University of Munich

科研成果: 书/报告/会议事项章节会议稿件同行评审

7 引用 (Scopus)

摘要

For separating two arbitrary sources from monaural recordings, the encoder-separator-decoder framework is popular in recent years. We investigated three kinds of filterbanks in the encoder: free, parameterized, and fixed. We proposed parameterized Gammatone and Gammachirp filterbanks, which improved performance with fewer parameters and better interpretability. Next, the properties of different filterbanks were investigated. Through training the network, an entirely freely learned filterbank emerges with properties similar to a series of bandpass filters spaced on a nonlinear scale - similar to the auditory system. We also explored the underlying separation mechanisms learned by the network through a classic auditory segregation experiment, revealing that the model separates mixtures based on the general principle (proximity of frequency and time). In summary, results demonstrate that the separation network automatically picks up the filterbank properties and separation mechanisms that are similar to those which have developed over millions of years in humans.

源语言英语
主期刊名2021 IEEE International Conference on Acoustics, Speech, and Signal Processing, ICASSP 2021 - Proceedings
出版商Institute of Electrical and Electronics Engineers Inc.
181-185
页数5
ISBN(电子版)9781728176055
DOI
出版状态已出版 - 2021
活动2021 IEEE International Conference on Acoustics, Speech, and Signal Processing, ICASSP 2021 - Virtual, Toronto, 加拿大
期限: 6 6月 202111 6月 2021

出版系列

姓名ICASSP, IEEE International Conference on Acoustics, Speech and Signal Processing - Proceedings
2021-June
ISSN(印刷版)1520-6149

会议

会议2021 IEEE International Conference on Acoustics, Speech, and Signal Processing, ICASSP 2021
国家/地区加拿大
Virtual, Toronto
时期6/06/2111/06/21

学术指纹

探究 'Auditory filterbanks benefit universal sound source separation' 的科研主题。它们共同构成独一无二的学术指纹。

引用此