PMID- 16021622 OWN - NLM STAT- MEDLINE DCOM- 20060313 LR - 20061115 IS - 1097-0134 (Electronic) IS - 0887-3585 (Linking) VI - 60 IP - 4 DP - 2005 Sep 1 TI - Structural analysis of a set of proteins resulting from a bacterial genomics project. PG - 787-96 AB - The targets of the Structural GenomiX (SGX) bacterial genomics project were proteins conserved in multiple prokaryotic organisms with no obvious sequence homolog in the Protein Data Bank of known structures. The outcome of this work was 80 structures, covering 60 unique sequences and 49 different genes. Experimental phase determination from proteins incorporating Se-Met was carried out for 45 structures with most of the remainder solved by molecular replacement using members of the experimentally phased set as search models. An automated tool was developed to deposit these structures in the Protein Data Bank, along with the associated X-ray diffraction data (including refined experimental phases) and experimentally confirmed sequences. BLAST comparisons of the SGX structures with structures that had appeared in the Protein Data Bank over the intervening 3.5 years since the SGX target list had been compiled identified homologs for 49 of the 60 unique sequences represented by the SGX structures. This result indicates that, for bacterial structures that are relatively easy to express, purify, and crystallize, the structural coverage of gene space is proceeding rapidly. More distant sequence-structure relationships between the SGX and PDB structures were investigated using PDB-BLAST and Combinatorial Extension (CE). Only one structure, SufD, has a truly unique topology compared to all folds in the PDB. CI - Copyright 2005 Wiley-Liss, Inc. FAU - Badger, J AU - Badger J AD - Structural GenomiX Inc., San Diego, California, USA. jbadger@active-sight.com FAU - Sauder, J M AU - Sauder JM FAU - Adams, J M AU - Adams JM FAU - Antonysamy, S AU - Antonysamy S FAU - Bain, K AU - Bain K FAU - Bergseid, M G AU - Bergseid MG FAU - Buchanan, S G AU - Buchanan SG FAU - Buchanan, M D AU - Buchanan MD FAU - Batiyenko, Y AU - Batiyenko Y FAU - Christopher, J A AU - Christopher JA FAU - Emtage, S AU - Emtage S FAU - Eroshkina, A AU - Eroshkina A FAU - Feil, I AU - Feil I FAU - Furlong, E B AU - Furlong EB FAU - Gajiwala, K S AU - Gajiwala KS FAU - Gao, X AU - Gao X FAU - He, D AU - He D FAU - Hendle, J AU - Hendle J FAU - Huber, A AU - Huber A FAU - Hoda, K AU - Hoda K FAU - Kearins, P AU - Kearins P FAU - Kissinger, C AU - Kissinger C FAU - Laubert, B AU - Laubert B FAU - Lewis, H A AU - Lewis HA FAU - Lin, J AU - Lin J FAU - Loomis, K AU - Loomis K FAU - Lorimer, D AU - Lorimer D FAU - Louie, G AU - Louie G FAU - Maletic, M AU - Maletic M FAU - Marsh, C D AU - Marsh CD FAU - Miller, I AU - Miller I FAU - Molinari, J AU - Molinari J FAU - Muller-Dieckmann, H J AU - Muller-Dieckmann HJ FAU - Newman, J M AU - Newman JM FAU - Noland, B W AU - Noland BW FAU - Pagarigan, B AU - Pagarigan B FAU - Park, F AU - Park F FAU - Peat, T S AU - Peat TS FAU - Post, K W AU - Post KW FAU - Radojicic, S AU - Radojicic S FAU - Ramos, A AU - Ramos A FAU - Romero, R AU - Romero R FAU - Rutter, M E AU - Rutter ME FAU - Sanderson, W E AU - Sanderson WE FAU - Schwinn, K D AU - Schwinn KD FAU - Tresser, J AU - Tresser J FAU - Winhoven, J AU - Winhoven J FAU - Wright, T A AU - Wright TA FAU - Wu, L AU - Wu L FAU - Xu, J AU - Xu J FAU - Harris, T J R AU - Harris TJ LA - eng PT - Journal Article PT - Research Support, Non-U.S. Gov't PT - Research Support, U.S. Gov't, Non-P.H.S. PL - United States TA - Proteins JT - Proteins JID - 8700181 RN - 0 (Enzymes) RN - 0 (Escherichia coli Proteins) SB - IM MH - Databases, Protein MH - Enzymes/chemistry/genetics MH - Escherichia coli/*genetics MH - Escherichia coli Proteins/*chemistry/genetics MH - *Genome, Bacterial MH - *Genomics MH - Models, Molecular MH - Protein Conformation MH - Regression Analysis MH - X-Ray Diffraction EDAT- 2005/07/16 09:00 MHDA- 2006/03/15 09:00 CRDT- 2005/07/16 09:00 PHST- 2005/07/16 09:00 [pubmed] PHST- 2006/03/15 09:00 [medline] PHST- 2005/07/16 09:00 [entrez] AID - 10.1002/prot.20541 [doi] PST - ppublish SO - Proteins. 2005 Sep 1;60(4):787-96. doi: 10.1002/prot.20541.