Active Binaural Auditory Perceptual System for a Socially Interactive Humanoid Robot

Sohaib Siddique Butt; Mahnoor Fatima; Ali Asghar; Wasif Muhammad

doi:10.3390/engproc2021012083

Active Binaural Auditory Perceptual System for a Socially Interactive Humanoid Robot

Engineering Proceedings ◽

10.3390/engproc2021012083 ◽

2022 ◽

Vol 12 (1) ◽

pp. 83

Author(s):

Sohaib Siddique Butt ◽

Mahnoor Fatima ◽

Ali Asghar ◽

Wasif Muhammad

Keyword(s):

Neural Network ◽

Source Localization ◽

Sound Source ◽

Humanoid Robot ◽

Cross Correlation ◽

Perceptual System ◽

Sound Source Localization ◽

Signal Processing Technique ◽

Gaze Shift ◽

Perception System

Sound Source Localization (SSL) and gaze shift to the sound source behavior is an integral part of a socially interactive humanoid robot perception system. In noisy and reverberant environments, it is non-trivial to estimate the location of a sound source and accurately shift gaze in its direction. Previous SSL algorithms are deficient in the optimum approximation of distance to audio sources and to accurately detect, interpret, and differentiate the actual sound from comparable sound sources due to challenging acoustic environments. In this article, a learning-based model is presented to achieve noiseless and reverberation-resistant sound source localization in the real-world scenarios. The proposed system utilizes a multi-layered Gaussian Cross-Correlation with Phase Transform (GCC-PHAT) signal processing technique as a baseline for a Generalized Cross Correlation Convolution Neural Network (GCC-CNN) model. The proposed model is integrated with an efficient rotation algorithm to predict and orient toward the sound source. The performance of the proposed method is compared with the state-of-art deep network-based sound source localization methods. The findings of the proposed method outperform the existing neural network-based approaches by achieving the highest accuracy of 96.21% for an active binaural auditory perceptual system.

Download Full-text

An Approach for Sound Source Localization by Complex-Valued Neural Network

IEICE Transactions on Information and Systems ◽

10.1587/transinf.e96.d.2257 ◽

2013 ◽

Vol E96.D (10) ◽

pp. 2257-2265 ◽

Cited By ~ 14

Author(s):

Hirofumi TSUZUKI ◽

Mauricio KUGLER ◽

Susumu KUROYANAGI ◽

Akira IWATA

Keyword(s):

Neural Network ◽

Source Localization ◽

Sound Source ◽

Sound Source Localization ◽

Complex Valued

Download Full-text

Enhancing direct‐path relative transfer function using deep neural network for robust sound source localization

CAAI Transactions on Intelligence Technology ◽

10.1049/cit2.12024 ◽

2021 ◽

Author(s):

Bing Yang ◽

Runwei Ding ◽

Yutong Ban ◽

Xiaofei Li ◽

Hong Liu

Keyword(s):

Neural Network ◽

Transfer Function ◽

Source Localization ◽

Sound Source ◽

Deep Neural Network ◽

Sound Source Localization ◽

Direct Path ◽

Relative Transfer ◽

Relative Transfer Function

Download Full-text

Sound-Source Localization System for Robotics and Industrial Automatic Control Systems Based on Neural Network

2008 International Conference on Smart Manufacturing Application ◽

10.1109/icsma.2008.4505664 ◽

2008 ◽

Cited By ~ 1

Author(s):

Yang Geng ◽

Jongdae Jung

Keyword(s):

Neural Network ◽

Automatic Control ◽

Control Systems ◽

Source Localization ◽

Sound Source ◽

Sound Source Localization ◽

Localization System ◽

Automatic Control Systems

Download Full-text

3 dimension sound source localization with cross-correlation and CORDIC algorithm on FPGA

2016 International Symposium on Electronics and Smart Devices (ISESD) ◽

10.1109/isesd.2016.7886749 ◽

2016 ◽

Author(s):

Agung Nuza Dwiputra ◽

Riko Hasiando Goknipasu Nainggolan ◽

Muhammad Arief Ma'ruf Nasution

Keyword(s):

Source Localization ◽

Sound Source ◽

Cross Correlation ◽

Sound Source Localization ◽

Cordic Algorithm

Download Full-text

Indoor Sound Source Localization Algorithm Based on BP Neural Network

10.1109/icct52962.2021.9658082 ◽

2021 ◽

Author(s):

Lan Wang ◽

Kun Zhang ◽

Chong Shen ◽

Chai Wang ◽

Xixi Fu

Keyword(s):

Neural Network ◽

Source Localization ◽

Sound Source ◽

Bp Neural Network ◽

Sound Source Localization ◽

Localization Algorithm

Download Full-text

Sound source localization by microphone array on a mobile robot using eigen-structure based generalized cross correlation

2008 IEEE Workshop on Advanced robotics and Its Social Impacts ◽

10.1109/arso.2008.4653625 ◽

2008 ◽

Cited By ~ 3

Author(s):

Jwu-Sheng Hu ◽

Chia-Hsing Yang ◽

Cheng-Kang Wang

Keyword(s):

Mobile Robot ◽

Source Localization ◽

Sound Source ◽

Cross Correlation ◽

Microphone Array ◽

Sound Source Localization

Download Full-text

Time-domain generalized cross correlation phase transform sound source localization for small microphone arrays

2012 5th European DSP Education and Research Conference (EDERC) ◽

10.1109/ederc.2012.6532229 ◽

2012 ◽

Cited By ~ 11

Author(s):

B. Van Den Broeck ◽

A. Bertrand ◽

P. Karsmakers ◽

B. Vanrumste ◽

H. Van hamme ◽

...

Keyword(s):

Source Localization ◽

Sound Source ◽

Time Domain ◽

Cross Correlation ◽

Microphone Arrays ◽

Sound Source Localization ◽

Phase Transform

Download Full-text

Indoor Sound Source Localization With Probabilistic Neural Network

IEEE Transactions on Industrial Electronics ◽

10.1109/tie.2017.2786219 ◽

2018 ◽

Vol 65 (8) ◽

pp. 6403-6413 ◽

Cited By ~ 34

Author(s):

Yingxiang Sun ◽

Jiajia Chen ◽

Chau Yuen ◽

Susanto Rahardja

Keyword(s):

Neural Network ◽

Source Localization ◽

Sound Source ◽

Probabilistic Neural Network ◽

Sound Source Localization

Download Full-text

DOANet: a deep dilated convolutional neural network approach for search and rescue with drone-embedded sound source localization

EURASIP Journal on Audio Speech and Music Processing ◽

10.1186/s13636-020-00184-2 ◽

2020 ◽

Vol 2020 (1) ◽

Author(s):

Alif Bin Abdul Qayyum ◽

K. M. Naimul Hassan ◽

Adrita Anika ◽

Md. Farhan Shadiq ◽

Md Mushfiqur Rahman ◽

...

Keyword(s):

Neural Network ◽

Deep Learning ◽

Convolutional Neural Network ◽

Source Localization ◽

Sound Source ◽

Performance Indicator ◽

Angular Spectrum ◽

Search And Rescue ◽

Sound Source Localization ◽

A Performance

Abstract Drone-embedded sound source localization (SSL) has interesting application perspective in challenging search and rescue scenarios due to bad lighting conditions or occlusions. However, the problem gets complicated by severe drone ego-noise that may result in negative signal-to-noise ratios in the recorded microphone signals. In this paper, we present our work on drone-embedded SSL using recordings from an 8-channel cube-shaped microphone array embedded in an unmanned aerial vehicle (UAV). We use angular spectrum-based TDOA (time difference of arrival) estimation methods such as generalized cross-correlation phase-transform (GCC-PHAT), minimum-variance-distortion-less-response (MVDR) as baseline, which are state-of-the-art techniques for SSL. Though we improve the baseline method by reducing ego-noise using speed correlated harmonics cancellation (SCHC) technique, our main focus is to utilize deep learning techniques to solve this challenging problem. Here, we propose an end-to-end deep learning model, called DOANet, for SSL. DOANet is based on a one-dimensional dilated convolutional neural network that computes the azimuth and elevation angles of the target sound source from the raw audio signal. The advantage of using DOANet is that it does not require any hand-crafted audio features or ego-noise reduction for DOA estimation. We then evaluate the SSL performance using the proposed and baseline methods and find that the DOANet shows promising results compared to both the angular spectrum methods with and without SCHC. To evaluate the different methods, we also introduce a well-known parameter—area under the curve (AUC) of cumulative histogram plots of angular deviations—as a performance indicator which, to our knowledge, has not been used as a performance indicator for this sort of problem before.

Download Full-text

Robust Sound Source Localization Using Convolutional Neural Network Based on Microphone Array

Intelligent Automation & Soft Computing ◽

10.32604/iasc.2021.018823 ◽

2021 ◽

Vol 29 (3) ◽

pp. 361-371

Author(s):

Xiaoyan Zhao ◽

Lin Zhou ◽

Ying Tong ◽

Yuxiao Qi ◽

Jingang Shi

Keyword(s):

Neural Network ◽

Convolutional Neural Network ◽

Source Localization ◽

Sound Source ◽

Microphone Array ◽

Sound Source Localization

Download Full-text