IARPA-BAA-11-02 - Babel Program (BABELON Project)

Název v češtině:"Babelon" - IARPA - výzva: BAA-11-02 - Program Babel
Hlavní řešitel:Matějka Pavel
Spoluřešitelé:Burget Lukáš, Černocký Jan, Glembek Ondřej, Grézl František, Hannemann Mirko, Karafiát Martin, Plchot Oldřich, Szőke Igor
Další řešitelé:Andrla Petr, Cipr Tomáš, Kesiraju Santosh (IIIT), Novotný Ondřej, Ondel Lucas, Skála František, Veselý Karel
Agentura:Raytheon BBN Technologies Corp.
Kód:IARPA-BABEL
Zahájení:2012-03-05
Ukončení:2016-11-04
Klíčová slova:speech recognition, speaker recognition, language recognition, LVCSR, feature extraction, acoustic modelling, neural-network
Anotace:
Cílem Babel programu je vyvinout agilní a robustní technologii pro rozpoznávání řeči, která může být rychle aplikována na jakoukoli mluvenou řeč, tak aby poskytla účinnou vyhledávací kapacitu analytikům pro efektivní zpracování záznamů velmi objemných souborů dat spontánní řeči.

Publikace

2016BRUMMER Niko, SWART Albert du Preez, PRIETO Jesús J., GARCIA Perera Leibny Paola, MATĚJKA Pavel, PLCHOT Oldřich, DIEZ Sánchez Mireia, SILNOVA Anna, JIANG Xiaowei, NOVOTNÝ Ondřej, ROHDIN Johan A., GLEMBEK Ondřej, GRÉZL František, BURGET Lukáš, ONDEL Lucas, PEŠÁN Jan, ČERNOCKÝ Jan, KENNY Patrick, ALAM Jahangir, BHATTACHARYA Gautam a ZEINALI Hossein et al. ABC NIST SRE 2016 SYSTEM DESCRIPTION. San Diego: United States Department of Commerce, National Institute of Standards and Technology, 2016.
 GRÉZL František a KARAFIÁT Martin. Boosting Performance on Low-resource Languages by Standard Corpora: AN ANALYSIS. In: Proceeding of SLT 2016. San Diego: IEEE Signal Processing Society, 2016, s. 629-636. ISBN 978-1-5090-4903-5.
 GRÉZL František a KARAFIÁT Martin. Bottle-Neck Feature Extraction Structures for Multilingual Training and Porting. In: Procedia Computer Science. Yogyakarta: Elsevier Science, 2016, s. 144-151. ISSN 1877-0509.
 GRÉZL František, EGOROVA Ekaterina a KARAFIÁT Martin. Study of Large Data Resources for Multilingual Training and System Porting. In: Procedia Computer Science. Yogyakarta: Elsevier Science, 2016, s. 15-22. ISSN 1877-0509.
 KARAFIÁT Martin, BASKAR Murali K., MATĚJKA Pavel, VESELÝ Karel, GRÉZL František a ČERNOCKÝ Jan. Multilingual BLSTM and Speaker-Specific Vector Adaptation in 2016 BUT BABEL SYSTEM. In: Proceedings of SLT 2016. San Diego: IEEE Signal Processing Society, 2016, s. 637-643. ISBN 978-1-5090-4903-5.
 KARAFIÁT Martin, BURGET Lukáš, GRÉZL František, VESELÝ Karel a ČERNOCKÝ Jan. Multilingual Region-Dependent Transforms. In: Proceedings of the 41th IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP 2016), 2016. Shanghai: IEEE Signal Processing Society, 2016, s. 5430-5434. ISBN 978-1-4799-9988-0.
 MATĚJKA Pavel, GLEMBEK Ondřej, NOVOTNÝ Ondřej, PLCHOT Oldřich, GRÉZL František, BURGET Lukáš a ČERNOCKÝ Jan. Analysis Of DNN Approaches To Speaker Identification. In: Proceedings of the 41th IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP 2016), 2016. Shanghai: IEEE Signal Processing Society, 2016, s. 5100-5104. ISBN 978-1-4799-9988-0.
 NOVOTNÝ Ondřej, MATĚJKA Pavel, GLEMBEK Ondřej, PLCHOT Oldřich, GRÉZL František, BURGET Lukáš a ČERNOCKÝ Jan. Analysis of the DNN-Based SRE Systems in Multi-language Conditions. In: Proceedings of SLT 2016. San Diego: IEEE Signal Processing Society, 2016, s. 199-204. ISBN 978-1-5090-4903-5.
 PLCHOT Oldřich, MATĚJKA Pavel, FÉR Radek, GLEMBEK Ondřej, NOVOTNÝ Ondřej, PEŠÁN Jan, VESELÝ Karel, ONDEL Lucas, KARAFIÁT Martin, GRÉZL František, KESIRAJU Santosh, BURGET Lukáš, BRUMMER Niko, SWART Albert du Preez, CUMANI Sandro, MALLIDI Sri Harish a LI Ruizhi. BAT System Description for NIST LRE 2015. In: Proceedings of Odyssey 2016, The Speaker and Language Recognition Workshop. Bilbao: International Speech Communication Association, 2016, s. 166-173. ISSN 2312-2846.
2015FÉR Radek, MATĚJKA Pavel, GRÉZL František, PLCHOT Oldřich a ČERNOCKÝ Jan. Multilingual Bottleneck Features for Language Recognition. In: Proceedings of Interspeech 2015. Dresden: International Speech Communication Association, 2015, s. 389-393. ISBN 978-1-5108-1790-6. ISSN 1990-9772.
 HEŘMANSKÝ Hynek, BURGET Lukáš, COHEN Jordan, DUPOUX Emmanuel, FELDMAN Naomi, GODFREY John, KHUDANPUR Sanjeev, MACIEJEWSKI Matthew, MALLIDI Sri Harish, MENON Anjali, OGAWA Tetsuji, PEDDINTI Vijayaditya, ROSE Richard, STERN Richard, WIESNER Matthew a VESELÝ Karel. Towards Machines That Know When They Do Not Know: Summary of Work Done at 2014 FREDERICK JELINEK MEMORIAL WORKSHOP. In: Proceedings of 2015 IEEE International Conference on Acoustics, Speech and Signal Processing. South Brisbane, Queensland: IEEE Signal Processing Society, 2015, s. 5009-5013. ISBN 978-1-4673-6997-8.
 HSIAO Roger, MA Jeff, HARTMANN William, KARAFIÁT Martin, GRÉZL František, BURGET Lukáš, SZŐKE Igor, ČERNOCKÝ Jan, WATANABE Shinji, CHEN Zhuo, MALLIDI Sri Harish, HEŘMANSKÝ Hynek, TSAKALIDIS Stavros a SCHWARTZ Richard. Robust Speech Recognition in Unknown Reverberant and Noisy Conditions. In: Proceedings of 2015 IEEE Automatic Speech Recognition and Understanding Workshop. Scottsdale, Arizona: IEEE Signal Processing Society, 2015, s. 533-538. ISBN 978-1-4799-7291-3.
 MALLIDI Sri Harish, OGAWA Tetsuji, VESELÝ Karel, NIDADAVOLU Phani S. a HEŘMANSKÝ Hynek. Autoencoder based multi-stream combination for noise robust speech recognition. In: Proceeding of Interspeech 2015. Dresden: International Speech Communication Association, 2015, s. 3551-3555. ISBN 978-1-5108-1790-6. ISSN 1990-9772.
 PEŠÁN Jan, BURGET Lukáš, HEŘMANSKÝ Hynek a VESELÝ Karel. DNN derived filters for processing of modulation spectrum of speech. In: Proceedings of Interspeech 2015. Dresden: International Speech Communication Association, 2015, s. 1908-1911. ISBN 978-1-5108-1790-6. ISSN 1990-9772.
2014GRÉZL František a KARAFIÁT Martin. Adapting Multilingual Neural Network Hierarchy to a New Language. In: Proceedings of the 4th International Workshop on Spoken Language Technologies for Under- resourced Languages SLTU-2014. – St. Petersburg, Russia, 2014. St. Petersburg: International Speech Communication Association, 2014, s. 39-45. ISBN 978-5-8088-0908-6.
 GRÉZL František a KARAFIÁT Martin. Combination of Multilingual and Semi-Supervised Training for Under-Resourced Languages. In: Proceedings of Interspeech 2014. Singapore: International Speech Communication Association, 2014, s. 820-824. ISBN 978-1-63439-435-2.
 GRÉZL František, EGOROVA Ekaterina a KARAFIÁT Martin. Further Investigation into Multilingual Training and Adaptation of Stacked Bottle-neck Neural Network Structure. In: Proceedings of 2014 Spoken Language Technology Workshop. South Lake Tahoe, Nevada: IEEE Signal Processing Society, 2014, s. 48-53. ISBN 978-1-4799-7129-9.
 GRÉZL František, KARAFIÁT Martin a VESELÝ Karel. Adaptation of Multilingual Stacked Bottle-neck Neural Network Structure for New Language. In: Proceedings of ICASSP 2014. Florencie: IEEE Signal Processing Society, 2014, s. 7704-7708. ISBN 978-1-4799-2892-7.
 KARAFIÁT Martin, GRÉZL František, HANNEMANN Mirko a ČERNOCKÝ Jan. BUT Neural Network Features for Spontaneous Vietnamese in BABEL. In: Proceedings of ICASSP 2014. Florencie: IEEE Signal Processing Society, 2014, s. 5659-5663. ISBN 978-1-4799-2892-7.
 KARAFIÁT Martin, GRÉZL František, VESELÝ Karel, HANNEMANN Mirko, SZŐKE Igor a ČERNOCKÝ Jan. BUT 2014 Babel System: Analysis of adaptation in NN based systems. In: Proceedings of Interspeech 2014. Singapore: International Speech Communication Association, 2014, s. 3002-3006. ISBN 978-1-63439-435-2.
 KARAFIÁT Martin, VESELÝ Karel, SZŐKE Igor, BURGET Lukáš, GRÉZL František, HANNEMANN Mirko a ČERNOCKÝ Jan. BUT ASR System for BABEL Surprise Evaluation 2014. In: Proceedings of 2014 Spoken Language Technology Workshop. South Lake Tahoe, Nevada: IEEE Signal Processing Society, 2014, s. 501-506. ISBN 978-1-4799-7129-9.
2013BURGET Lukáš, PLCHOT Oldřich a SZŐKE Igor. 2013 Summary report of project "Processing and analysis of speech, automatic speaker identification". Brno: Raytheon BBN Technologies Corp., 2013.
 GRÉZL František a KARAFIÁT Martin. Semi-Supervised Bootstrapping Approach For Neural Network Feature Extractor Training. In: Proceedings of ASRU 2013. Olomouc: IEEE Signal Processing Society, 2013, s. 470-475. ISBN 978-1-4799-2755-5.
 HANNEMANN Mirko, POVEY Daniel a ZWEIG Geoffrey. Combining Forward and Backward Search in Decoding. In: Proceedings of ICASSP 2013. Vancouver: IEEE Signal Processing Society, 2013, s. 6739-6743. ISBN 978-1-4799-0355-9.
 HSIAO Roger, NG Tim, GRÉZL František, KARAKOS Damianos, TSAKALIDIS Stavros, NGUYEN Long a SCHWARTZ Richard. Discriminative Semi-supervised Training for Keyword Search in Low Resource Languages. In: Proceedings of ASRU 2013. Olomouc: IEEE Signal Processing Society, 2013, s. 440-445. ISBN 978-1-4799-2755-5.
 KARAFIÁT Martin, GRÉZL František, HANNEMANN Mirko, VESELÝ Karel a ČERNOCKÝ Jan. BUT BABEL System for Spontaneous Cantonese. In: Proceedings of Interspeech 2013. Lyon: International Speech Communication Association, 2013, s. 2589-2593. ISBN 978-1-62993-443-3. ISSN 2308-457X.
 KARAKOS Damianos, SCHWARTZ Richard, TSAKALIDIS Stavros, ZHANG Le, RANJAN Shivesh, NG Tim, HSIAO Roger, NGUYEN Long, GRÉZL František, HANNEMANN Mirko, KARAFIÁT Martin, SZŐKE Igor a VESELÝ Karel et al. Score Normalization and System Combination for Improved Keyword Spotting. In: Proceedings of ASRU 2013. Olomouc: IEEE Signal Processing Society, 2013, s. 210-215. ISBN 978-1-4799-2755-5.
 LEI Yun, BURGET Lukáš a SCHEFFER Nicolas. A Noise Robust I-Vector Extractor Using Vector Taylor Series For Speaker Recognition. In: Proceedings of ICASSP 2013. Vancouver: IEEE Signal Processing Society, 2013, s. 6788-6791. ISBN 978-1-4799-0355-9.
 VESELÝ Karel, GHOSHAL Arnab, BURGET Lukáš a POVEY Daniel. Sequence-discriminative Training of Deep Neural Networks. In: Proceedings of Interspeech 2013. Lyon: International Speech Communication Association, 2013, s. 2345-2349. ISBN 978-1-62993-443-3. ISSN 2308-457X.
 VESELÝ Karel, HANNEMANN Mirko a BURGET Lukáš. Semi-supervised Training of Deep Neural Networks. In: Proceedings of ASRU 2013. Olomouc: IEEE Signal Processing Society, 2013, s. 267-272. ISBN 978-1-4799-2755-5.
2012BURGET Lukáš, GLEMBEK Ondřej, MATĚJKA Pavel a PLCHOT Oldřich. 2012 Summary report of project "Processing and analysis of speech, automatic speaker identification". Cambridge: Raytheon BBN Technologies Corp., 2012.
 VESELÝ Karel, KARAFIÁT Martin, GRÉZL František, JANDA Miloš a EGOROVA Ekaterina. The Language-Independent Bottleneck Features. In: Proceedings of IEEE 2012 Workshop on Spoken Language Technology. Miami: IEEE Signal Processing Society, 2012, s. 336-341. ISBN 978-1-4673-5124-9.

Vaše IPv4 adresa: 54.225.47.94
Přepnout na IPv6 spojení

DNSSEC [dnssec]