Článek ve sborníku konference

 
Martínez, G., D., Burget, L., Ferrer, L., Scheffer, N.: Ivector-Based Prosodic System For Language Identification, In: Proc. International Conference on Acoustics, Speec, Kyoto, JP, IEEESP, 2012, s. 4861-4864, ISBN 978-1-4673-0044-5
Jazyk publikace:angličtina
Název publikace:Ivector-Based Prosodic System For Language Identification
Název (cs):I-vektorový prozodický systém pro rozpoznávání jazyka
Strany:4861-4864
Sborník:Proc. International Conference on Acoustics, Speec
Konference:The 37th International Conference on Acoustics, Speech, and Signal Processing
Místo vydání:Kyoto, JP
Rok:2012
ISBN:978-1-4673-0044-5
Vydavatel:IEEE Signal Processing Society
URL:http://www.fit.vutbr.cz/research/groups/speech/publi/2012/martinez_icassp2012_0004861.pdf [PDF]
Klíčová slova
Language Identification, Prosody, iVectors, Joint Factor Analysis
Anotace
Článek pojednává o I-vektorovém prozodickém systému pro rozpoznávání jazyka. Byl zkoumán rytmus, důraz a intonace jazyka.
Abstrakt
Prosody is the part of speech where rhythm, stress, and intonation are reflected. In language identification tasks, these characteristics are assumed to be language dependent, and thus the language can be identified from them. In this paper, an automatic language recognition system that extracts prosody information from speech and makes decisions about the language with a generative classifier based on iVectors is built. The system is tested on the NIST LRE09 dataset. The results are still not comparable to state-of-the-art acoustic and phonotactic systems. However, they are promising and the fusion of the new approach with an iVector-based acoustic system is found to bring further improvements over the latter.
BibTeX:
@INPROCEEDINGS{
   author = {David González Martínez and Lukáš Burget and Luciana Ferrer
	and Nicolas Scheffer},
   title = {Ivector-Based Prosodic System For Language Identification},
   pages = {4861--4864},
   booktitle = {Proc. International Conference on Acoustics, Speec},
   year = {2012},
   location = {Kyoto, JP},
   publisher = {IEEE Signal Processing Society},
   ISBN = {978-1-4673-0044-5},
   language = {english},
   url = {http://www.fit.vutbr.cz/research/view_pub.php?id=9997}
}