Structural analysis of a set of proteins resulting from a bacterial genomics project
Citations Over TimeTop 1% of 2005 papers
Abstract
The targets of the Structural GenomiX (SGX) bacterial genomics project were proteins conserved in multiple prokaryotic organisms with no obvious sequence homolog in the Protein Data Bank of known structures. The outcome of this work was 80 structures, covering 60 unique sequences and 49 different genes. Experimental phase determination from proteins incorporating Se-Met was carried out for 45 structures with most of the remainder solved by molecular replacement using members of the experimentally phased set as search models. An automated tool was developed to deposit these structures in the Protein Data Bank, along with the associated X-ray diffraction data (including refined experimental phases) and experimentally confirmed sequences. BLAST comparisons of the SGX structures with structures that had appeared in the Protein Data Bank over the intervening 3.5 years since the SGX target list had been compiled identified homologs for 49 of the 60 unique sequences represented by the SGX structures. This result indicates that, for bacterial structures that are relatively easy to express, purify, and crystallize, the structural coverage of gene space is proceeding rapidly. More distant sequence-structure relationships between the SGX and PDB structures were investigated using PDB-BLAST and Combinatorial Extension (CE). Only one structure, SufD, has a truly unique topology compared to all folds in the PDB.
Related Papers
- → Protein structure databases with new web services for structural biology and biomedical research(2008)73 cited
- → PDBminer to Find and Annotate Protein Structures for Computational Analysis(2023)4 cited
- → Effect of low-complexity regions on protein structure determination(2007)14 cited
- → PSSARD: Protein sequence-structure analysis relational database(2005)4 cited
- → Functional Linkages Can Reveal Protein Complexes for Structure Determination(2007)4 cited