Incorporating Knowledge Sources into Statistical Speech Recognition
Book information
Description
Incorporating Knowledge Sources into Statistical Speech Recognition offers solutions for enhancing the robustness of a statistical automatic speech recognition (ASR) system by incorporating various additional knowledge sources while keeping the training and recognition effort feasible. The authors provide an efficient general framework for incorporating knowledge sources into state-of-the-art statistical ASR systems. This framework, which is called GFIKS (graphical framework to incorporate additional knowledge sources), was designed by utilizing the concept of the Bayesian network (BN) framework. This framework allows probabilistic relationships among different information sources to be learned, various kinds of knowledge sources to be incorporated, and a probabilistic function of the model to be formulated. Incorporating Knowledge Sources into Statistical Speech Recognition demonstrates how the statistical speech recognition system may incorporate additional information sources by utilizing GFIKS at different levels of ASR. The incorporation of various knowledge sources, including background noises, accent, gender and wide phonetic knowledge information, in modeling is discussed theoretically and analyzed experimentally.
Similar books
Applied Op Amp Circuits: Analysis and Design with NI® Multisim™
2023 · EPUB
Applied Op Amp Circuits: Analysis and Design with NI® Multisim™
2023 · PDF
Handbook Of Digital Face Manipulation And Detection: From DeepFakes To Morphing Attacks
2022 · PDF
Memristor Emulator Circuits
2021 · PDF
Image Processing and Computer Vision in iOS
2020 · PDF
Fashion Recommender Systems
2020 · PDF
Games and Learning Alliance: 9th International Conference, GALA 2020, Laval, France, December 9–10, 2020, Proceedings
2020 · PDF
Intelligent Systems: 9th Brazilian Conference, BRACIS 2020, Rio Grande, Brazil, October 20–23, 2020, Proceedings, Part I
2020 · PDF