Building and Using Comparable Corpora
Book information
Description
The 1990s saw a paradigm change in the use of corpus-driven methods in NLP. In the field of multilingual NLP (such as machine translation and terminology mining) this implied the use of parallel corpora. However, parallel resources are relatively scarce: many more texts are produced daily by native speakers of any given language than translated. This situation resulted in a natural drive towards the use of comparable corpora, i.e. non-parallel texts in the same domain or genre. Nevertheless, this research direction has not produced a single authoritative source suitable for researchers and students coming to the field. The proposed volume provides a reference source, identifying the state of the art in the field as well as future trends. The book is intended for specialists and students in natural language processing, machine translation and computer-assisted translation.
Similar books
Sanskrit Computational Linguistics: Third International Symposium, Hyderabad, India, January 15-17, 2009. Proceedings
2009 · PDF
Linguistic Expressions and Semantic Processing: A Practical Approach
2015 · PDF
Advanced Applications of Natural Language Processing for Performing Information Extraction
2015 · PDF
Computational Linguistics and Intelligent Text Processing: 16th International Conference, CICLing 2015, Cairo, Egypt, April 14-20, 2015, Proceedings, Part II
2015 · PDF
Computational Linguistics and Intelligent Text Processing: 16th International Conference, CICLing 2015, Cairo, Egypt, April 14-20, 2015, Proceedings, Part I
2015 · PDF
Language Processing with Perl and Prolog: Theories, Implementation, and Application
2014 · PDF
The Welsh Language in the Digital Age
2014 · PDF
Natural Language Processing of Semitic Languages
2014 · PDF