Robust Source Counting and DOA Estimation Using Spatial Pseudo-Spectrum and Convolutional Neural Network
Thi Ngoc Tho Nguyen, Woon‐Seng Gan, Rishabh Ranjan, Douglas L. Jones
Nanyang Technological University University of Illinois Urbana-Champaign
阅读操作
确认中在文库中上传 PDF 后可生成中文音频讲解。
摘要与影响
Many signal processing-based methods for sound source direction-of-arrival estimation produce a spatial pseudo-spectrum of which the local maxima strongly indicate the source directions. Due to different levels of noise, reverberation and different number of overlapping sources, the spatial pseudo-spectra are noisy even after smoothing. In addition, the number of sources is often unknown. As a result, selecting the peaks from these spectra is susceptible to error. Convolutional neural network has been successfully applied to many image processing problems in general and direction-of-arrival estimation in particular. In addition, deep learning-based methods for direction-of-arrival estimation show good generalization to different environments. We propose to use a 2D convolutional neural network with multi-task learning to robustly estimate the number of sources and the directions-of-arrival from short-time spatial pseudo-spectra, which have useful directional information from audio input signals. This approach reduces the tendency of the neural network to learn unwanted association between sound classes and directional information, and helps the network generalize to unseen sound classes. The simulation and experimental results show that the proposed methods outperform other directional-of-arrival estimation methods in different levels of noise and reverberation, and different number of sources.
逐年被引趋势
关键指标
同类平均 = 1
同领域 · 同年份 · 同类型
Google Scholar 与 OpenAlex 的被引统计范围不同,数值存在差异属正常。
AI 辅助阅读
依据:摘要
可就本文提问;依据不足时会说明。
学术脉络
学科主题
计算机 / AISpeech and Audio Processing
Underwater Acoustics Research · Speech Recognition and Synthesis
参考文献 33
此处列出前 3 条
引用本文 92
按被引量排序,此处列出前 3 条