2001Unpublished venueRequires access

iProClass: An Integrated Protein Classification Database for Proteomics

Yini Huang, Winona C. Barker

Open publisher page 0 citations

Abstract

Advanced databases are essential for gaining insight into protein structure and function from the voluminous, heterogeneous, and distributed molecular data. Protein family classification is now well recognized as an effective means for large-scale functional characterization of genes. The iProClass database [1] is an integrated resource that provides comprehensive family relationships at both global (whole protein) and local (domain and motif/site) levels, as well as structural/functional classifications and features of proteins. The PIR superfamily/family organization allows complete and nonoverlapping clustering of all proteins. The iProClass consists of more than 266,000 nonredundant PIR and Swiss-Prot proteins organized with more than 30,000 superfamilies, 100,000 families, 2600 domains, 1300 motifs, 280 post-translational modification sites, and links to over 40 databases of protein families, structures, functions, genes, genomes, literature, and taxonomy. Protein and superfamily summary reports provide rich annotations, including membership information with length, taxonomy, and keyword statistics, comprehensive enzyme and PDB cross-references, and graphical feature display. The database facilitates classification-driven annotation for protein sequences and complete genomes, and supports proteomics research. The iProClass is implemented in Oracle 8i object-relational system and available for sequence/text search and report

About this research paper

What this paper is about

Advanced databases are essential for gaining insight into protein structure and function from the voluminous, heterogeneous, and distributed molecular data. Protein family classification is now well recognized as an effective means for large-scale functional characterization of genes. The iProClass database [1] is an integrated resource that provides comprehensive family relationships at both global (whole protein) and local (domain and motif/site) levels, as well as structural/functional classifications and features of proteins. The PIR superfamily/family organization allows complete and nonoverlapping clustering of all proteins. The iProClass consists of more than 266,000 nonredundant PIR and Swiss-Prot proteins organized with more than 30,000 superfamilies, 100,000 families, 2600 domains, 1300 motifs, 280 post-translational modification sites, and links to over 40 databases of protein families, structures, functions, genes, genomes, literature, and taxonomy. Protein and superfamily summary reports provide rich annotations, including membership information with length, taxonomy, and keyword statistics, comprehensive enzyme and PDB cross-references, and graphical feature display. The database facilitates classification-driven annotation for protein sequences and complete genomes, and supports proteomics research. The iProClass is implemented in Oracle 8i object-relational system and available for sequence/text search and report

Why it matters

A significance statement is not available in the OpenAlex record.

Key contribution

A contribution statement is not available in the OpenAlex record.

Method / approach

Method details are not available in the OpenAlex metadata.

Main findings

Findings are not separately available in the OpenAlex metadata.

Limitations

Limitations are not available in the OpenAlex metadata.

Applications

Application details are not available in the OpenAlex metadata.

Available abstract

Advanced databases are essential for gaining insight into protein structure and function from the voluminous, heterogeneous, and distributed molecular data. Protein family classification is now well recognized as an effective means for large-scale functional characterization of genes. The iProClass database [1] is an integrated resource that provides comprehensive family relationships at both global (whole protein) and local (domain and motif/site) levels, as well as structural/functional classifications and features of proteins. The PIR superfamily/family organization allows complete and nonoverlapping clustering of all proteins. The iProClass consists of more than 266,000 nonredundant PIR and Swiss-Prot proteins organized with more than 30,000 superfamilies, 100,000 families, 2600 domains, 1300 motifs, 280 post-translational modification sites, and links to over 40 databases of protein families, structures, functions, genes, genomes, literature, and taxonomy. Protein and superfamily summary reports provide rich annotations, including membership information with length, taxonomy, and keyword statistics, comprehensive enzyme and PDB cross-references, and graphical feature display. The database facilitates classification-driven annotation for protein sequences and complete genomes, and supports proteomics research. The iProClass is implemented in Oracle 8i object-relational system and available for sequence/text search and report

Key concepts: Structural Classification of Proteins database, Protein Data Bank (RCSB PDB), Protein structure database, UniProt, RefSeq, Protein family, Relational database, Protein domain

Related papers

Back to paper searchBrowse research topicsOriginal source
iProClass: An Integrated Protein Classification Database for Proteomics — Research Paper | ScholarLens