Szczegóły publikacji

Opis bibliograficzny

Application of morphosyntactic and class-based language models in automatic speech recognition of Polish / Aleksander SMYWIŃSKI-POHL, Bartosz ZIÓŁKO // International Journal on Artificial Intelligence Tools ; ISSN 0218-2130. — 2016 — vol. 25 no. 2 art. no. 1650006, s. 1650006-1–1650006-24. — Bibliogr. s. 1650006-22–1650006-24. — Publikacja dostępna online od: 2016-04-22. — Dod. afiliacja autorów: Jagiellonian University

Autorzy (2)

Słowa kluczowe

word clusteringASRrerankingmorphosyntactic language modelautomatic speech recognitionclass-based language model

Dane bibliometryczne

ID BaDAP98054
Data dodania do BaDAP2016-06-21
DOI10.1142/S0218213016500068
Rok publikacji2016
Typ publikacjiartykuł w czasopiśmie
Otwarty dostęptak
Czasopismo/seriaInternational Journal on Artificial Intelligence Tools

Abstract

In this paper we investigate the usefulness of morphosyntactic information as well as clustering in modeling Polish for automatic speech recognition. Polish is an inflectional language, thus we investigate the usefulness of an N-gram model based on morphosyntactic features. We present how individual types of features influence the model and which types of features are best suited for building a language model for automatic speech recognition. We compared the results of applying them with a class-based model that is automatically derived from the training corpus. We show that our approach towards clustering performs significantly better than frequently used SRI LM clustering method. However, this difference is apparent only for smaller corpora.

Publikacje, które mogą Cię zainteresować

fragment książki
#78957Data dodania: 21.1.2014
A comparison of Polish taggers in the application for automatic speech recognition / Aleksander POHL, Bartosz ZIÓŁKO // W: Human language technologies as a challenge for computer science and linguistics : 6th language & technology conference : December 7–9, 2013, Poznań : proceedings / eds. Zygmunt Vetulani, Hans Uszkoreit. — Poznań : Fundacja Uniwersytetu im. A. Mickiewicza, 2013 + CD. — ISBN: 978-83-932640-3-2; e-ISBN: 978-83-932640-4-9. — S. 294–298. — Bibliogr. s. 298, Abstr.
artykuł
#50169Data dodania: 5.2.2010
Language modeling and large vocabulary continuous speech recognition / Leszek Gajecki, Ryszard TADEUSIEWICZ // Journal of Applied Computer Science ; ISSN 1507-0360. — 2009 — vol. 17 no. 2, s. 57–70. — Bibliogr. s. 69–70, Abstr.