Crossref journal-article
Wiley
Protein Science (311)
Abstract

AbstractModeling the inherent flexibility of the protein backbone as part of computational protein design is necessary to capture the behavior of real proteins and is a prerequisite for the accurate exploration of protein sequence space. We present the results of a broad exploration of sequence space, with backbone flexibility, through a novel approach: large‐scale protein design to structural ensembles. A distributed computing architecture has allowed us to generate hundreds of thousands of diverse sequences for a set of 253 naturally occurring proteins, allowing exciting insights into the nature of protein sequence space. Designing to a structural ensemble produces a much greater diversity of sequences than previous studies have reported, and homology searches using profiles derived from the designed sequences against the Protein Data Bank show that the relevance and quality of the sequences is not diminished. The designed sequences have greater overall diversity than corresponding natural sequence alignments, and no direct correlations are seen between the diversity of natural sequence alignments and the diversity of the corresponding designed sequences. For structures in the same fold, the sequence entropies of the designed sequences cluster together tightly. This tight clustering of sequence entropies within a fold and the separation of sequence entropy distributions for different folds suggest that the diversity of designed sequences is primarily determined by a structure's overall fold, and that the designability principle postulated from studies of simple models holds in real proteins. This has important implications for experimental protein design and engineering, as well as providing insight into protein evolution.

Bibliography

Larson, S. M., England, J. L., Desjarlais, J. R., & Pande, V. S. (2002). Thoroughly sampling sequence space: Large‐scale protein design of structural ensembles. Protein Science, 11(12), 2804–2813. Portico.

Dates
Type When
Created 22 years, 9 months ago (Nov. 19, 2002, 7:14 p.m.)
Deposited 1 year, 10 months ago (Oct. 9, 2023, 10:41 a.m.)
Indexed 1 month ago (July 26, 2025, 5:04 a.m.)
Issued 22 years, 8 months ago (Dec. 1, 2002)
Published 22 years, 8 months ago (Dec. 1, 2002)
Published Online 16 years, 4 months ago (April 13, 2009)
Published Print 22 years, 8 months ago (Dec. 1, 2002)
Funders 0

None

@article{Larson_2002, title={Thoroughly sampling sequence space: Large‐scale protein design of structural ensembles}, volume={11}, ISSN={1469-896X}, url={http://dx.doi.org/10.1110/ps.0203902}, DOI={10.1110/ps.0203902}, number={12}, journal={Protein Science}, publisher={Wiley}, author={Larson, Stefan M. and England, Jeremy L. and Desjarlais, John R. and Pande, Vijay S.}, year={2002}, month=dec, pages={2804–2813} }