Literature DB >> 11151005

Protein structural domains: analysis of the 3Dee domains database.

U Dengler1, A S Siddiqui, G J Barton.   

Abstract

The 3Dee database of domain definitions was developed as a comprehensive collection of domain definitions for all three-dimensional structures in the Protein Data Bank (PDB). The database includes definitions for complex, multiple-segment and multiple-chain domains as well as simple sequential domains, organized in a structural hierarchy. Two different snapshots of the 3Dee database were analyzed at September 1996 and November 1999. For the November 1999 release, 7,995 PDB entries contained 13,767 protein chains and gave rise to 18,896 domains. The domain sequences clustered into 1,715 domain sequence families, which were further clustered into a conservative 1,199 domain structure families (families with similar folds). The proportion of different domain structure families per domain sequence family increases from 84% for domains 1-100 residues long to 100% for domains greater than 600 residues. This is in keeping with the idea that longer chains will have more alternative folds available to them. Of the representative domains from the domain sequence families, 49% are in the range of 51-150 residues, whereas 64% of the representative chains over 200 residues have more than 1 domain. Of the representative chains, 8.5% are part of multichain domains. The largest multichain domain in the database has 14 chains and 1,400 residues, whereas the largest single-chain domain has 907 residues. The largest number of domains found in a protein is 13. The analysis shows that over the history of the PDB, new domain folds have been discovered at a slower rate than by random selection of all known folds. Between 1992 and 1997, a constant 1 in 11 new domains deposited in the PDB has shown no sequence similarity to a previously known domain sequence family, and only 1 in 15 new domain structures has had a fold that has not been seen previously. A comparison of the September 1996 release of 3Dee to the Structural Classification of Proteins (SCOP) showed that the domain definitions agreed for 80% of the representative protein chains. However, 3Dee provided explicit domain boundaries for more proteins. 3Dee is accessible on the World Wide Web at http://barton.ebi.ac.uk/servers/3Dee.html.

Mesh:

Substances:

Year:  2001        PMID: 11151005

Source DB:  PubMed          Journal:  Proteins        ISSN: 0887-3585


  10 in total

1.  A consensus view of fold space: combining SCOP, CATH, and the Dali Domain Dictionary.

Authors:  Ryan Day; David A C Beck; Roger S Armen; Valerie Daggett
Journal:  Protein Sci       Date:  2003-10       Impact factor: 6.725

2.  Estimating the accuracy of protein structures using residual dipolar couplings.

Authors:  Katya Simon; Jun Xu; Chinpal Kim; Nikolai R Skrynnikov
Journal:  J Biomol NMR       Date:  2005-10       Impact factor: 2.835

3.  A method for finding candidate conformations for molecular replacement using relative rotation between domains of a known structure.

Authors:  Jay I Jeong; Eaton E Lattman; Gregory S Chirikjian
Journal:  Acta Crystallogr D Biol Crystallogr       Date:  2006-03-18

4.  PASS2: a semi-automated database of protein alignments organised as structural superfamilies.

Authors:  V Mallika; Anirban Bhaduri; R Sowdhamini
Journal:  Nucleic Acids Res       Date:  2002-01-01       Impact factor: 16.971

5.  Universality in protein residue networks.

Authors:  Ernesto Estrada
Journal:  Biophys J       Date:  2010-03-03       Impact factor: 4.033

6.  The CATH hierarchy revisited-structural divergence in domain superfamilies and the continuity of fold space.

Authors:  Alison Cuff; Oliver C Redfern; Lesley Greene; Ian Sillitoe; Tony Lewis; Mark Dibley; Adam Reid; Frances Pearl; Tim Dallman; Annabel Todd; Richard Garratt; Janet Thornton; Christine Orengo
Journal:  Structure       Date:  2009-08-12       Impact factor: 5.006

7.  DIAL: a web-based server for the automatic identification of structural domains in proteins.

Authors:  Ganesan Pugalenthi; Govindaraju Archunan; Ramanathan Sowdhamini
Journal:  Nucleic Acids Res       Date:  2005-07-01       Impact factor: 16.971

8.  The path to enlightenment: making sense of genomic and proteomic information.

Authors:  Martin H Maurer
Journal:  Genomics Proteomics Bioinformatics       Date:  2004-05       Impact factor: 7.691

9.  ELISA: structure-function inferences based on statistically significant and evolutionarily inspired observations.

Authors:  Boris E Shakhnovich; John M Harvey; Steve Comeau; David Lorenz; Charles DeLisi; Eugene Shakhnovich
Journal:  BMC Bioinformatics       Date:  2003-09-02       Impact factor: 3.169

10.  OXBench: a benchmark for evaluation of protein multiple sequence alignment accuracy.

Authors:  G P S Raghava; Stephen M J Searle; Patrick C Audley; Jonathan D Barber; Geoffrey J Barton
Journal:  BMC Bioinformatics       Date:  2003-10-10       Impact factor: 3.169

  10 in total

北京卡尤迪生物科技股份有限公司 © 2022-2023.