2012The MIT Press eBooksRequires access

Mining the Biomedical Literature

Hagit Shatkay, Mark Craven

Open publisher page 22 citations

Abstract

A concise introduction to fundamental methods for finding and extracting relevant information from the ever-increasing amounts of biomedical text available The introduction of high-throughput methods has transformed biology into a data-rich science. Knowledge about biological entities and processes has traditionally been acquired by thousands of scientists through decades of experimentation and analysis. The current abundance of biomedical data is accompanied by the creation and quick dissemination of new information. Much of this information and knowledge, however, is represented only in text form—in the biomedical literature, lab notebooks, Web pages, and other sources. Researchers' need to find relevant information in the vast amounts of text has created a surge of interest in automated text-analysis. In this book, Hagit Shatkay and Mark Craven offer a concise and accessible introduction to key ideas in biomedical text mining. The chapters cover such topics as the relevant sources of biomedical text; text-analysis methods in natural language processing; the tasks of information extraction, information retrieval, and text categorization; and methods for empirically assessing text-mining systems. Finally, the authors describe several applications that recognize entities in text and link them to other entities and data resources, support the curation of structured databases, and make use of text to enable further prediction and discovery.

About this research paper

What this paper is about

A concise introduction to fundamental methods for finding and extracting relevant information from the ever-increasing amounts of biomedical text available The introduction of high-throughput methods has transformed biology into a data-rich science. Knowledge about biological entities and processes has traditionally been acquired by thousands of scientists through decades of experimentation and analysis. The current abundance of biomedical data is accompanied by the creation and quick dissemination of new information. Much of this information and knowledge, however, is represented only in text form—in the biomedical literature, lab notebooks, Web pages, and other sources. Researchers' need to find relevant information in the vast amounts of text has created a surge of interest in automated text-analysis. In this book, Hagit Shatkay and Mark Craven offer a concise and accessible introduction to key ideas in biomedical text mining. The chapters cover such topics as the relevant sources of biomedical text; text-analysis methods in natural language processing; the tasks of information extraction, information retrieval, and text categorization; and methods for empirically assessing text-mining systems. Finally, the authors describe several applications that recognize entities in text and link them to other entities and data resources, support the curation of structured databases, and make use of text to enable further prediction and discovery.

Why it matters

OpenAlex reports 22 citations for this work. Citation counts describe recorded attention and do not establish research quality.

Key contribution

A contribution statement is not available in the OpenAlex record.

Method / approach

Method details are not available in the OpenAlex metadata.

Main findings

Findings are not separately available in the OpenAlex metadata.

Limitations

Limitations are not available in the OpenAlex metadata.

Applications

Application details are not available in the OpenAlex metadata.

Available abstract

A concise introduction to fundamental methods for finding and extracting relevant information from the ever-increasing amounts of biomedical text available The introduction of high-throughput methods has transformed biology into a data-rich science. Knowledge about biological entities and processes has traditionally been acquired by thousands of scientists through decades of experimentation and analysis. The current abundance of biomedical data is accompanied by the creation and quick dissemination of new information. Much of this information and knowledge, however, is represented only in text form—in the biomedical literature, lab notebooks, Web pages, and other sources. Researchers' need to find relevant information in the vast amounts of text has created a surge of interest in automated text-analysis. In this book, Hagit Shatkay and Mark Craven offer a concise and accessible introduction to key ideas in biomedical text mining. The chapters cover such topics as the relevant sources of biomedical text; text-analysis methods in natural language processing; the tasks of information extraction, information retrieval, and text categorization; and methods for empirically assessing text-mining systems. Finally, the authors describe several applications that recognize entities in text and link them to other entities and data resources, support the curation of structured databases, and make use of text to enable further prediction and discovery.

Key concepts: Biomedical text mining, Computer science, Information retrieval, Data science, Text mining, Information extraction, Question answering, Categorization

Related papers

Back to paper searchBrowse research topicsOriginal source
Mining the Biomedical Literature — Research Paper | ScholarLens