Result Details

Unscented Transform For Ivector-based Noisy Speaker Recognition

MARTÍNEZ GONZÁLEZ, D.; BURGET, L.; STAFYLAKIS, T.; LEI, Y.; KENNY, P.; LLEIDA, E. Unscented Transform For Ivector-based Noisy Speaker Recognition. In Proceedings of ICASSP 2014. Florencie: IEEE Signal Processing Society, 2014. p. 4070-4074. ISBN: 978-1-4799-2892-7.
Type
conference paper
Language
English
Authors
Martínez González David, FIT (FIT)
Burget Lukáš, doc. Ing., Ph.D., DCGM (FIT)
Stafylakis Themos
Lei Yun
Kenny Patrick
LLeida Eduardo
Abstract

This article is about new version of an unscented transform for Ivector-based noisy speaker recognition.

Keywords

Noise Robust Speaker Recognition, UnscentedTransform, Vector Taylor Series, iVector

URL
Annotation

Recently, a new version of the iVector modelling has been proposed for noise robust speaker recognition, where the nonlinear function that relates clean and noisy cepstral coefficients is approximated by a first order vector Taylor series (VTS). In this paper, it is proposed to substitute the first order VTS by an unscented transform, where unlike VTS, the nonlinear function is not applied over the clean model parameters directly, but over a set of sampled points. The resulting points in the transformed space are then used to calculate the model parameters. For very low signal-to-noise ratio improvements in equal error rate of about 7% for a clean backend and of 14.50% for a multistyle backend are obtained.

Published
2014
Pages
4070–4074
Proceedings
Proceedings of ICASSP 2014
Conference
The 39th International Conference on Acoustics, Speech, and Signal Processing (ICASSP)
ISBN
978-1-4799-2892-7
Publisher
IEEE Signal Processing Society
Place
Florencie
DOI
UT WoS
000343655304013
EID Scopus
BibTeX
@inproceedings{BUT111555,
  author="David {Martínez González} and Lukáš {Burget} and Themos {Stafylakis} and Yun {Lei} and Patrick {Kenny} and Eduardo {LLeida}",
  title="Unscented Transform For Ivector-based Noisy Speaker Recognition",
  booktitle="Proceedings of ICASSP 2014",
  year="2014",
  pages="4070--4074",
  publisher="IEEE Signal Processing Society",
  address="Florencie",
  doi="10.1109/ICASSP.2014.6854361",
  isbn="978-1-4799-2892-7",
  url="https://www.fit.vut.cz/research/publication/10573/"
}
Files
Projects
Centrum excelence IT4Innovations, MŠMT, Operační program Výzkum a vývoj pro inovace, ED1.1.00/02.0070, start: 2011-01-01, end: 2015-12-31, completed
DARPA Robust Automatic Transcription of Speech (RATS) - RATS Patrol I, BBN, start: 2010-09-23, end: 2014-06-30, completed
Technologies of speech processing for efficient human-machine communication, TAČR, Program aplikovaného výzkumu a experimentálního vývoje ALFA, TA01011328, start: 2011-01-01, end: 2014-12-31, completed
Research groups
Departments
Back to top