Multi-Source DOA Estimation Through Pattern Recognition of the Modal Coherence of a Reverberant Soundfield

Abdullah Fahim; Prasanga N. Samarasinghe; Thushara D. Abhayapala

首页> 外文期刊>Audio, Speech, and Language Processing, IEEE/ACM Transactions on >Multi-Source DOA Estimation Through Pattern Recognition of the Modal Coherence of a Reverberant Soundfield

【24h】

Multi-Source DOA Estimation Through Pattern Recognition of the Modal Coherence of a Reverberant Soundfield

机译：通过模式识别识别混响声场的模态一致性的多源DOA估计

获取原文

获取原文并翻译 | 示例

开具论文收录证明 >>

页面导航

摘要
著录项
引文网络
相似文献
相关主题

摘要

We propose a novel multi-source direction of arrival (DOA) estimation technique using a convolutional neural network algorithm which learns the modal coherence patterns of an incident soundfield through measured spherical harmonic coefficients. We train our model for individual time-frequency bins in the short-time Fourier transform spectrum by analyzing the unique snapshot of modal coherence for each desired direction. The proposed method is capable of estimating simultaneously active multiple sound sources on a 3D space using a single-source training scheme. This single-source training scheme reduces the training time and resource requirements as well as allows the reuse of the same trained model for different multi-source combinations. The method is evaluated against various simulated and practical noisy and reverberant environments with varying acoustic criteria and found to outperform the baseline methods in terms of DOA estimation accuracy. Furthermore, the proposed algorithm allows independent training of azimuth and elevation during a full DOA estimation over 3D space which significantly improves its training efficiency without affecting the overall estimation accuracy.

机译：我们提出了一种使用卷积神经网络算法提出了一种新的多源到达（DOA）估计技术，该算法通过测量的球形谐波系数来学习入射声结构场的模态相干模式。通过分析每个所需方向的模态相干的唯一快照，我们在短时间傅立叶变换频谱中培训我们的模型。所提出的方法能够使用单源训练方案在3D空间上同时估计在3D空间上的同时活动多个声源。这种单源训练方案减少了培训时间和资源要求，并允许重用相同的多源组合的训练模型。该方法针对具有不同声学标准的各种模拟和实用噪声和混响环境，并发现在DOA估计准确度方面优于基线方法。此外，所提出的算法允许在3D空间的完整DOA估计期间独立训练方位角和高程，这显着提高了其训练效率而不影响整体估计精度。

著录项

来源
《Audio, Speech, and Language Processing, IEEE/ACM Transactions on》 |2020年第2020期|605-618|共14页
作者
Abdullah Fahim; Prasanga N. Samarasinghe; Thushara D. Abhayapala;
展开▼
作者单位

Audio and Acoustic Signal Processing Group The Australian National University Canberra ACT Australia;

Audio and Acoustic Signal Processing Group The Australian National University Canberra ACT Australia;

Audio and Acoustic Signal Processing Group The Australian National University Canberra ACT Australia;

展开▼
收录信息
原文格式 PDF
正文语种 eng
中图分类
关键词
Direction-of-arrival estimation; Estimation; Harmonic analysis; Training; Signal processing algorithms; Coherence; Acoustics;

机译：到达方向估计;估计;谐波分析;训练;信号处理算法;一致性;声学;

相似文献

外文文献
中文文献
专利

1. Multi-source DOA estimation in reverberant environments using potential single-source points enhancement [J] . Jia Maoshen, Jia Yitian, Gao Shang, Applied Acoustics . 2021,第Mara期

机译：使用潜在单源点增强的混响环境中的多源DOA估计
2. Multi-source TDOA estimation in reverberant audio using angular spectra and clustering [J] . Charles Blandin, Alexey Ozerov, Emmanuel Vincent Signal processing . 2012,第8期

机译：使用角谱和聚类的混响音频中的多源TDOA估计
3. Square Root-Based Multi-Source Early PSD Estimation and Recursive RETF Update in Reverberant Environments by Means of the Orthogonal Procrustes Problem [J] . Thomas Dietzen, Simon Doclo, Marc Moonen, Audio, Speech, and Language Processing, IEEE/ACM Transactions on . 2020,第期

机译：通过正交促进问题解决方形基于根的多源早期PSD估计和递归RETF更新
4. Multi-source DOA estimation in a reverberant room [C] . Ali, A.M., Hudson, Personal, Indoor and Mobile Radio Communications,2005 IEEE 16th International Symposium on . 2008

机译：混响室内的多源DOA估计
5. Abstraction and Recognition of Dynamic Modulation Patterns within and between Sensory Modalities [D] . Rector, Jeffrey David 2018

机译：感觉模态内部和之间的动态调制模式的抽象和识别
6. Human Body Mixed Motion Pattern Recognition Method Based on Multi-Source Feature Parameter Fusion [O] . Jiyuan Song, Aibin Zhu, Yao Tu, 2020

机译：基于多源特征参数融合的人体混合运动模式识别方法
7. Multi-source TDOA estimation in reverberant audio using angular spectra and clustering 1 [O] . Charles Bl, Alexey Ozerov, Emmanuel Vincent 2012

机译：使用角谱和聚类的混响音频中的多源TDOA估计1

Multi-Source DOA Estimation Through Pattern Recognition of the Modal Coherence of a Reverberant Soundfield

摘要

著录项

引文网络

相似文献

相关主题

期刊订阅