ENGLISH

Information Extraction: Algorithms and Prospects in a Retrieval Context

Book information

Publisher
Springer
Year
2006
ISBN
3540740473, 9783540740476
LCC
ML74 .M86 2007
Open Library ID
OL22664951M
Language
english
Format
PDF
Filesize
5 MB (5711500 bytes)
Series
The Information Retrieval Series
Edition
1
Pages
254\254
Time added
2010-02-18 13:16:04

Description

Information extraction regards the processes of structuring and combining content that is explicitly stated or implied in one or multiple unstructured information sources. It involves a semantic classification and linking of certain pieces of information and is considered as a light form of content understanding by the machine. Currently, there is a considerable interest in integrating the results of information extraction in retrieval systems, because of the growing demand for search engines that return precise answers to flexible information queries. Advanced retrieval models satisfy that need and they rely on tools that automatically build a probabilistic model of the content of a (multi-media) document. The book focuses on content recognition in text. It elaborates on the past and current most successful algorithms and their application in a variety of domains (e.g., news filtering, mining of biomedical text, intelligence gathering, competitive intelligence, legal information searching, and processing of informal text). An important part discusses current statistical and machine learning algorithms for information detection and classification and integrates their results in probabilistic retrieval models. The book also reveals a number of ideas towards an advanced understanding and synthesis of textual content. The book is aimed at researchers and software developers interested in information extraction and retrieval, but the many illustrations and real world examples make it also suitable as a handbook for students.

Similar books