A Study on Text-Independent Speaker Recognition Systems in Emotional Conditions Using Different Pattern Recognition Models

In the modern world, human recognition systems play an important role to improve security by reducing chances of evasion. Human ear is used for person identification .In the Empirical study on research on human ear, 10000 images are taken to find the uniqueness of the ear. Ear based system is one of the few biometric systems which can provides stable characteristics over the age. In this paper, ear images are taken from mathematical analysis of images (AMI) ear data base and the analysis is done on ear pattern recognition based on the Expectation maximization algorithm and k means algorithm. Pattern of ears affected with different types of noises are recognized based on Principle component analysis (PCA) algorithm.

Download Full-text

AFRL/HECP Speaker Recognition Systems for the 2004 NIST Speaker Recognition Evaluation

10.21236/ada430750 ◽

2004 ◽

Author(s):

Raymond E. Slyh ◽

Eric G. Hansen ◽

Timothy R. Anderson

Keyword(s):

Speaker Recognition ◽

Recognition Systems

Download Full-text

Wavelet-based denoising for EEG-based pattern recognition systems

2020 IEEE Symposium Series on Computational Intelligence (SSCI) ◽

10.1109/ssci47803.2020.9308421 ◽

2020 ◽

Author(s):

Binh Nguyen ◽

Wanli Ma ◽

Dat Tran ◽

Younjin Chung

Keyword(s):

Pattern Recognition ◽

Pattern Recognition Systems ◽

Recognition Systems

Download Full-text

Impact of Variability Issues of Resistive Memory Synapses on Pattern Recognition Systems

2020 International SoC Design Conference (ISOCC) ◽

10.1109/isocc50952.2020.9333029 ◽

2020 ◽

Author(s):

Jiyong Woo ◽

Miyoung Lee ◽

Jeong Hun Kim ◽

Jong-Pil Im ◽

Solyee Im ◽

...

Keyword(s):

Pattern Recognition ◽

Resistive Memory ◽

Pattern Recognition Systems ◽

Recognition Systems

Download Full-text

Experimental estimation of the recognition reliability in the optical pattern recognition systems

10.1117/12.720177 ◽

2007 ◽

Author(s):

Veacheslav L. Perju ◽

David P. Casasent ◽

Igor A. Mardare ◽

Oleg V. Chirca

Keyword(s):

Pattern Recognition ◽

Optical Pattern Recognition ◽

Experimental Estimation ◽

Optical Pattern ◽

Pattern Recognition Systems ◽

Recognition Systems

Download Full-text

Enhancing intellectual power of recognition systems based on new pattern recognition theory

Proceedings 2002 IEEE International Conference on Artificial Intelligence Systems (ICAIS 2002) ◽

10.1109/icais.2002.1048086 ◽

2003 ◽

Cited By ~ 1

Author(s):

N.G. Fedotov ◽

L.A. Shulga

Keyword(s):

Pattern Recognition ◽

Pattern Recognition Theory ◽

Recognition Theory ◽

Recognition Systems ◽

Intellectual Power

Download Full-text

A robust data simulation technique to improve early detection performance of a classifier in control chart pattern recognition systems

Information Sciences ◽

10.1016/j.ins.2020.09.059 ◽

2021 ◽

Vol 548 ◽

pp. 18-36

Author(s):

Ramazan Ünlü

Keyword(s):

Pattern Recognition ◽

Early Detection ◽

Control Chart ◽

Detection Performance ◽

Simulation Technique ◽

Data Simulation ◽

Pattern Recognition Systems ◽

Recognition Systems ◽

Control Chart Pattern

Download Full-text

Maximum Likelihood and Maximum a Posteriori Adaptation for Distributed Speaker Recognition Systems

Biometric Authentication - Lecture Notes in Computer Science ◽

10.1007/978-3-540-25948-0_87 ◽

2004 ◽

pp. 640-647 ◽

Cited By ~ 1

Author(s):

Chin-Hung Sit ◽

Man-Wai Mak ◽

Sun-Yuan Kung

Keyword(s):

Maximum Likelihood ◽

Speaker Recognition ◽

Maximum A Posteriori ◽

A Posteriori ◽

Recognition Systems

Download Full-text

Performance Evaluation of Mel and Bark Scale based Features for Text-Independent Speaker Identification

International Journal of Innovative Technology and Exploring Engineering - Special Issue ◽

10.35940/ijitee.k1999.0981119 ◽

2019 ◽

Vol 8 (11) ◽

pp. 3734-3738

Keyword(s):

Speaker Recognition ◽

Filter Bank ◽

Work Performance ◽

Speaker Identification ◽

Recognition Rate ◽

Identification System ◽

Bank Structure ◽

Speech Features ◽

Recognition Systems ◽

Human Ear

The performance of Mel scale and Bark scale is evaluated for text-independent speaker identification system. Mel scale and Bark scale are designed according to human auditory system. The filter bank structure is defined using Mel and Bark scales for speech and speaker recognition systems to extract speaker specific speech features. In this work, performance of Mel scale and Bark scale is evaluated for text-independent speaker identification system. It is found that Bark scale centre frequencies are more effective than Mel scale centre frequencies in case of Indian dialect speaker databases. Mel scale is defined as per interpretation of pitch by human ear and Bark scale is based on critical band selectivity at which loudness becomes significantly different. The recognition rate achieved using Bark scale filter bank is 96% for AISSMSIOIT database and 95% for Marathi database.

Download Full-text

U-Vectors: Generating Clusterable Speaker Embedding from Unlabeled Data

Applied Sciences ◽

10.3390/app112110079 ◽

2021 ◽

Vol 11 (21) ◽

pp. 10079

Author(s):

Muhammad Firoz Mridha ◽

Abu Quwsar Ohi ◽

Muhammad Mostafa Monowar ◽

Md. Abdul Hamid ◽

Md. Rashedul Islam ◽

...

Keyword(s):

Speaker Recognition ◽

Large Scale ◽

English Language ◽

Domain Adaptation ◽

Recognition System ◽

Extraction Process ◽

Unlabeled Data ◽

Training Strategy ◽

Speech Segment ◽

Recognition Systems

Speaker recognition deals with recognizing speakers by their speech. Most speaker recognition systems are built upon two stages, the first stage extracts low dimensional correlation embeddings from speech, and the second performs the classification task. The robustness of a speaker recognition system mainly depends on the extraction process of speech embeddings, which are primarily pre-trained on a large-scale dataset. As the embedding systems are pre-trained, the performance of speaker recognition models greatly depends on domain adaptation policy, which may reduce if trained using inadequate data. This paper introduces a speaker recognition strategy dealing with unlabeled data, which generates clusterable embedding vectors from small fixed-size speech frames. The unsupervised training strategy involves an assumption that a small speech segment should include a single speaker. Depending on such a belief, a pairwise constraint is constructed with noise augmentation policies, used to train AutoEmbedder architecture that generates speaker embeddings. Without relying on domain adaption policy, the process unsupervisely produces clusterable speaker embeddings, termed unsupervised vectors (u-vectors). The evaluation is concluded in two popular speaker recognition datasets for English language, TIMIT, and LibriSpeech. Also, a Bengali dataset is included to illustrate the diversity of the domain shifts for speaker recognition systems. Finally, we conclude that the proposed approach achieves satisfactory performance using pairwise architectures.

Download Full-text