Deep Learning Methods in Speaker Recognition: A Review

Beke, András; Szaszák, György; Sztahó, Dávid

Deep Learning Methods in Speaker Recognition: A Review

Authors: András Beke
György Szaszák
Dávid Sztahó
Publication date: 29 October 2021
Publisher: Budapest University of Technology and Economics (BME)

Abstract

This paper reviews the applied Deep Learning (DL) practices in the field of Speaker Recognition (SR), both in verification and identification. Speaker Recognition has been a widely used topic of speech technology. Many research works have been carried out and little progress has been achieved in the past 5–6 years. However, as Deep Learning techniques do advance in most machine learning fields, the former state-of-the-art methods are getting replaced by them in Speaker Recognition too. It seems that Deep Learning becomes the now state-of-the-art solution for both Speaker Verification (SV) and identification. The standard x-vectors, additional to i-vectors, are used as baseline in most of the novel works. The increasing amount of gathered data opens up the territory to Deep Learning, where they are the most effective

Similar works

Full text

Open in the Core reader

Download PDF

Available Versions

Periodica Polytechnica (Budapest University of Technology and Economics)

oai:ojs.pkp.sfu.ca:article/170...

Last time updated on 23/11/2023