IADIS International Journal on Computer Science and Information Systems

Published by IADIS (International Association for Development of the Information Society) • ISSN (Online): 1646-3692 • ISSN (Print): 1646-3692
100% Open Access
Double-Blind Peer Review
Crossref DOI Persistent IDs
Open Access Peer-Reviewed Original Research

Improved Voice-based Biometrics Using Multi-channel Transfer Learning

Youssouf Ismail Cherifi *
Abdelhakim Dahimene *
2 1Universite de M’Hamed Bougara *
Boumerdes Researcher *
Algeria Researcher *
* 2Signal and System laboratory, Boumerdes, Algeria (Portugal)
* 2Signal and System laboratory, Boumerdes, Algeria (Portugal)
* 2Signal and System laboratory, Boumerdes, Algeria (Portugal)
* 2Signal and System laboratory, Boumerdes, Algeria (Portugal)
* 2Signal and System laboratory, Boumerdes, Algeria (Portugal)

Abstract

Identifying the speaker has become more of an imperative thing to do in the modern age. Especially since most personal and professional appliances rely on voice commands or speech in general terms to operate. These systems need to discern the identity of the speaker rather than just the words that have been said to be both smart and safe. Especially if we consider the numerous advanced methods that have been developed to generate fake speech segments. The objective of this paper is to improve upon the existing voice-based biometrics to keep up with these synthesizers. The proposed method focuses on defining a novel and more speaker adapted features by implying artificial neural networks and transfer learning. The approach uses pre -trained networks to define a mapping from two complementary acoustic features to a speaker adapted phonetic features. The complementary acoustics features are paired to provide both information about how the speech segments are perceived (type 1 feature) and produced (type 2 feature). The approach was evaluated using both a small and large closed-speaker data set. Primary results are encouraging and confirm the usefulness of such an approach to extract speaker adapted features whether for classical machine learning algorithms or advanced neural structures such as LSTM or CNN.

Keywords

Speech Analysis Transfer Learning Pattern Recognition Speaker Recognition Feature Extraction
Full-Text PDF Available

Read Complete Peer-Reviewed Manuscript

Includes full econometric models, data tables, policy recommendations, declarations, and citations.

Declarations & Ethics

Funding: This research received academic dissemination support through ESCAP / JournalsHub publishing programs.
Conflicts of Interest: The authors declare no competing financial or institutional interests.
Peer Review: Double-blind peer reviewed by international subject specialists.
License: Creative Commons Attribution 4.0 International (CC BY 4.0).
How to Cite This Article
APA / MLA / BibTeX
Cherifi, et al. (2020). Improved Voice-based Biometrics Using Multi-channel Transfer Learning. IADIS International Journal on Computer Science and Information Systems, 15(1). https://doi.org/10.33965/ijcsis_2020_v15i1_09
Cherifi, et al. "Improved Voice-based Biometrics Using Multi-channel Transfer Learning." IADIS International Journal on Computer Science and Information Systems, vol. 15, no. 1, 2020. https://doi.org/10.33965/ijcsis_2020_v15i1_09
Cherifi, et al. "Improved Voice-based Biometrics Using Multi-channel Transfer Learning." IADIS International Journal on Computer Science and Information Systems 15, no. 1 (2020). https://doi.org/10.33965/ijcsis_2020_v15i1_09