Multi-source doa estimation through pattern recognition of the modal coherence of a reverberant soundfield
| dc.contributor.author | Fahim, Abdullah | |
| dc.contributor.author | Samarasinghe, Prasanga | |
| dc.contributor.author | Abhayapala, Thushara | |
| dc.date.accessioned | 2024-02-20T02:34:53Z | |
| dc.date.issued | 2020 | |
| dc.date.updated | 2022-10-02T07:20:22Z | |
| dc.description.abstract | We propose a novel multi-source direction of arrival (DOA) estimation technique using a convolutional neural network algorithm which learns the modal coherence patterns of an incident soundfield through measured spherical harmonic coefficients. We train our model for individual time-frequency bins in the short-time Fourier transform spectrum by analyzing the unique snapshot of modal coherence for each desired direction. The proposed method is capable of estimating simultaneously active multiple sound sources on a 3D space using a single-source training scheme. This single-source training scheme reduces the training time and resource requirements as well as allows the reuse of the same trained model for different multi-source combinations. The method is evaluated against various simulated and practical noisy and reverberant environments with varying acoustic criteria and found to outperform the baseline methods in terms of DOA estimation accuracy. Furthermore, the proposed algorithm allows independent training of azimuth and elevation during a full DOA estimation over 3D space which significantly improves its training efficiency without affecting the overall estimation accuracy. | en_AU |
| dc.format.mimetype | application/pdf | en_AU |
| dc.identifier.issn | 2329-9290 | en_AU |
| dc.identifier.uri | http://hdl.handle.net/1885/313766 | |
| dc.language.iso | en_AU | en_AU |
| dc.publisher | IEEE Signal Processing Society | en_AU |
| dc.rights | © 2019 IEEE | en_AU |
| dc.source | IEEE/ACM Transactions on Audio, Speech, and Language Processing | en_AU |
| dc.subject | Convolutional neural network | en_AU |
| dc.subject | DOA estimation | en_AU |
| dc.subject | spatial audio processing | en_AU |
| dc.subject | spherical harmonics | en_AU |
| dc.title | Multi-source doa estimation through pattern recognition of the modal coherence of a reverberant soundfield | en_AU |
| dc.type | Journal article | en_AU |
| local.bibliographicCitation.lastpage | 14 | en_AU |
| local.bibliographicCitation.startpage | 1 | en_AU |
| local.contributor.affiliation | Fahim, Abdullah, College of Engineering and Computer Science, ANU | en_AU |
| local.contributor.affiliation | Samarasinghe, Prasanga, College of Engineering and Computer Science, ANU | en_AU |
| local.contributor.affiliation | Abhayapala, Thushara, College of Engineering and Computer Science, ANU | en_AU |
| local.contributor.authoruid | Fahim, Abdullah, u5898301 | en_AU |
| local.contributor.authoruid | Samarasinghe, Prasanga, u4801876 | en_AU |
| local.contributor.authoruid | Abhayapala, Thushara, u9701943 | en_AU |
| local.description.embargo | 2099-12-31 | |
| local.description.notes | Imported from ARIES | en_AU |
| local.identifier.absfor | 400607 - Signal processing | en_AU |
| local.identifier.absfor | 460302 - Audio processing | en_AU |
| local.identifier.ariespublication | u6269649xPUB711 | en_AU |
| local.identifier.citationvolume | 28 | en_AU |
| local.identifier.doi | 10.1109/TASLP.2019.2960734 | en_AU |
| local.identifier.scopusID | 2-s2.0-85078725718 | |
| local.identifier.thomsonID | WOS:000526685200004 | |
| local.publisher.url | https://ieeexplore.ieee.org/ | en_AU |
| local.type.status | Published Version | en_AU |
Downloads
Original bundle
1 - 1 of 1
Loading...
- Name:
- Multi-Source_DOA_Estimation_Through_Pattern_Recognition_of_the_Modal_Coherence_of_a_Reverberant_Soundfield.pdf
- Size:
- 2 MB
- Format:
- Adobe Portable Document Format
- Description: