2021bioRxiv (Cold Spring Harbor Laboratory)Open access

AlphaFold Protein Structure Database for Sequence-Independent Molecular Replacement

Lawrence Chai, Ping Zhu, Jin Chai, Changxu Pang, Babak Andi, Seán McSweeney, John Shanklin, Qun Liu

Open full text 0 citations

Abstract

Abstract Crystallographic phasing recovers the phase information that is lost during a diffraction experiment. Molecular replacement is a dominant phasing method for the crystal structures in the protein data bank. In one form it uses a protein sequence to search a structure database for finding suitable templates for phasing. However, such sequence information is not always available such as when proteins are crystallized with unknown binding partner proteins or when the crystal is that of a contaminant. The recent development of AlphaFold has resulted in the availability of predicted protein structures for all proteins from twenty species. In this work, we tested whether AlphaFold-predicted E. coli protein structures were accurate enough for sequence-independent phasing of diffraction data from two crystallization contaminants for which we had not identified the protein. Using each of more than 4000 predicted structures as a search model, robust molecular replacement solutions were obtained which allowed the identification and structure determination of both structures, YncE and YadF. Our results advocate a general utility of AlphaFold-predicted structure database with respect to crystallographic phasing.

Open-access reader

About this research paper

What this paper is about

Abstract Crystallographic phasing recovers the phase information that is lost during a diffraction experiment. Molecular replacement is a dominant phasing method for the crystal structures in the protein data bank. In one form it uses a protein sequence to search a structure database for finding suitable templates for phasing. However, such sequence information is not always available such as when proteins are crystallized with unknown binding partner proteins or when the crystal is that of a contaminant. The recent development of AlphaFold has resulted in the availability of predicted protein structures for all proteins from twenty species. In this work, we tested whether AlphaFold-predicted E. coli protein structures were accurate enough for sequence-independent phasing of diffraction data from two crystallization contaminants for which we had not identified the protein. Using each of more than 4000 predicted structures as a search model, robust molecular replacement solutions were obtained which allowed the identification and structure determination of both structures, YncE and YadF. Our results advocate a general utility of AlphaFold-predicted structure database with respect to crystallographic phasing.

Why it matters

A significance statement is not available in the OpenAlex record.

Key contribution

A contribution statement is not available in the OpenAlex record.

Method / approach

Method details are not available in the OpenAlex metadata.

Main findings

Findings are not separately available in the OpenAlex metadata.

Limitations

Limitations are not available in the OpenAlex metadata.

Applications

Application details are not available in the OpenAlex metadata.

Available abstract

Abstract Crystallographic phasing recovers the phase information that is lost during a diffraction experiment. Molecular replacement is a dominant phasing method for the crystal structures in the protein data bank. In one form it uses a protein sequence to search a structure database for finding suitable templates for phasing. However, such sequence information is not always available such as when proteins are crystallized with unknown binding partner proteins or when the crystal is that of a contaminant. The recent development of AlphaFold has resulted in the availability of predicted protein structures for all proteins from twenty species. In this work, we tested whether AlphaFold-predicted E. coli protein structures were accurate enough for sequence-independent phasing of diffraction data from two crystallization contaminants for which we had not identified the protein. Using each of more than 4000 predicted structures as a search model, robust molecular replacement solutions were obtained which allowed the identification and structure determination of both structures, YncE and YadF. Our results advocate a general utility of AlphaFold-predicted structure database with respect to crystallographic phasing.

Key concepts: Phaser, Molecular replacement, Protein Data Bank, Sequence (biology), Protein crystallization, Protein structure, Crystal structure, Crystallography

Related papers

Back to paper searchBrowse research topicsOriginal source
AlphaFold Protein Structure Database for Sequence-Independent Molecular Replacement — Research Paper | ScholarLens