Skrybot – A System for Automatic Speech Recognition of Polish Language

  • Lesław Pawlaczyk
  • Paweł Bosky
Part of the Advances in Intelligent and Soft Computing book series (AINSC, volume 59)


In this article we present a system for clustering and indexing of automatically recognised radio and television news spoken in Polish language. The aim of the system is to quickly navigate and search for information which is not available in standard internet search engines. The system comprises of speech recognition, alignment and indexing module. The recognition part is trained using dozens of hours of transcribed audio and millions of words representing modern Polish language. The training audio and text is then converted into acoustic and language model, where we apply techniques such as Hidden Markov Models and statistical language processing. The audio is decoded and later submitted into indexing engine which extracts summary information about the spoken topic. The system presents a significant potential in many areas such as media monitoring, university lectures indexing, automated telephone centres and security enhancements.


speech recognition speech processing pattern recognition 


Unable to display preview. Download preview PDF.

Unable to display preview. Download preview PDF.


  1. 1.
    Clarkson, P.R., Rosenfeld, R.: Statistical language modeling using the CMU-Cambridge toolkit. In: Proceedings of the European Conference on Speech Communication and Technology (1997)Google Scholar
  2. 2.
    Formey, G.D.: The Viterbi algorithm. Proceedings of the IEEE 61, 268–278 (1973)CrossRefGoogle Scholar
  3. 3.
    Jurafsky, D., Martin, J.H.: Machine translation. In: Ward, N., Jurafsky, D. (eds.) Speech and Language Processing. Prentice-Hall, Englewood Cliffs (2000)Google Scholar
  4. 4.
    Lee, A., Kawahar, T., Shikano, K.: Julius – an open source real-time large vocabulary recognition engine. In: Proceedings of the European Conference on Speech Communication and Technology, pp. 1691–1694 (2001)Google Scholar
  5. 5.
    Young, S., et al.: The HTK book (for HTK version 3.4). Cambridge University Engineering Department (2006)Google Scholar

Copyright information

© Springer-Verlag Berlin Heidelberg 2009

Authors and Affiliations

  • Lesław Pawlaczyk
    • 1
  • Paweł Bosky
  1. 1.Insitute of InformaticsSilesian University of TechnologyGliwicePoland

Personalised recommendations