2016•Journal of Data Mining in Genomics & ProteomicsOpen access

Comparative genomics: From genome sequences to genome biology

Michael Y. Galperin

Open full text 1 citations

Abstract

Genome biology aims at using the complete genome sequences to reconstruct all metabolic and signaling pathways that could operate in the target organisms and identify the likely regulatory hubs and potential drug targets. Such analysis requires comprehensive functional annotation of all proteins encoded in each sequenced genome. Standard sequence analysis typically fails to provide (confident) functional assignment for at least a third of the genes even in the relatively small prokaryotic genomes. As a result, comparative genomics has to deal with the constantly growing numbers of proteins whose functions remain unknown. This talk will discuss using comparative genomics to improve our understanding of microbial metabolic and signaling pathways, including some recent examples of identification of missing enzymes and prediction of alternative enzyme variants. It will show that the number of truly enigmatic conserved hypothetical proteins is relatively small, particularly in the reduced genomes of pathogenic bacteria, which suggests that most of their cellular functions are already accounted for. In contrast, the number of uncharacterized genes in free-living organisms remains quite large and their functions remain obscure. Our current hypothesis is that many of these genes have “house-cleaning” function, which is almost as important as house-keeping, particularly for aerobic bacteria and for eukaryotic cells. We shall also briefly discuss how comparative genomics could be used for identification of priority targets for future research and the challenges in characterization of their functions

About this research paper

What this paper is about

Genome biology aims at using the complete genome sequences to reconstruct all metabolic and signaling pathways that could operate in the target organisms and identify the likely regulatory hubs and potential drug targets. Such analysis requires comprehensive functional annotation of all proteins encoded in each sequenced genome. Standard sequence analysis typically fails to provide (confident) functional assignment for at least a third of the genes even in the relatively small prokaryotic genomes. As a result, comparative genomics has to deal with the constantly growing numbers of proteins whose functions remain unknown. This talk will discuss using comparative genomics to improve our understanding of microbial metabolic and signaling pathways, including some recent examples of identification of missing enzymes and prediction of alternative enzyme variants. It will show that the number of truly enigmatic conserved hypothetical proteins is relatively small, particularly in the reduced genomes of pathogenic bacteria, which suggests that most of their cellular functions are already accounted for. In contrast, the number of uncharacterized genes in free-living organisms remains quite large and their functions remain obscure. Our current hypothesis is that many of these genes have “house-cleaning” function, which is almost as important as house-keeping, particularly for aerobic bacteria and for eukaryotic cells. We shall also briefly discuss how comparative genomics could be used for identification of priority targets for future research and the challenges in characterization of their functions

Why it matters

OpenAlex reports 1 citations for this work. Citation counts describe recorded attention and do not establish research quality.

Key contribution

A contribution statement is not available in the OpenAlex record.

Method / approach

Method details are not available in the OpenAlex metadata.

Main findings

Findings are not separately available in the OpenAlex metadata.

Limitations

Limitations are not available in the OpenAlex metadata.

Applications

Application details are not available in the OpenAlex metadata.

Available abstract

Genome biology aims at using the complete genome sequences to reconstruct all metabolic and signaling pathways that could operate in the target organisms and identify the likely regulatory hubs and potential drug targets. Such analysis requires comprehensive functional annotation of all proteins encoded in each sequenced genome. Standard sequence analysis typically fails to provide (confident) functional assignment for at least a third of the genes even in the relatively small prokaryotic genomes. As a result, comparative genomics has to deal with the constantly growing numbers of proteins whose functions remain unknown. This talk will discuss using comparative genomics to improve our understanding of microbial metabolic and signaling pathways, including some recent examples of identification of missing enzymes and prediction of alternative enzyme variants. It will show that the number of truly enigmatic conserved hypothetical proteins is relatively small, particularly in the reduced genomes of pathogenic bacteria, which suggests that most of their cellular functions are already accounted for. In contrast, the number of uncharacterized genes in free-living organisms remains quite large and their functions remain obscure. Our current hypothesis is that many of these genes have “house-cleaning” function, which is almost as important as house-keeping, particularly for aerobic bacteria and for eukaryotic cells. We shall also briefly discuss how comparative genomics could be used for identification of priority targets for future research and the challenges in characterization of their functions

Key concepts: Genome, Comparative genomics, Biology, Genomics, Computational biology, Function (biology), Gene, Identification (biology)

Related papers

Back to paper searchBrowse research topicsOriginal source
Comparative genomics: From genome sequences to genome biology — Research Paper | ScholarLens