Abstract
The SCOP (Structural Classification of Proteins) database is a comprehensive ordering of all proteins of known structure, according to their evolutionary and structural relationships. Protein domains in SCOP are grouped into species and hierarchically classified into families, superfamilies, folds and classes. Recently, we introduced a new set of features with the aim of standardizing access to the database, and providing a solid basis to manage the increasing number of experimental structures expected from structural genomics projects. These features include: a new set of identifiers, which uniquely identify each entry in the hierarchy; a compact representation of protein domain classification; a new set of parseable files, which fully describe all domains in SCOP and the hierarchy itself. These new features are reflected in the ASTRAL compendium. The SCOP search engine has also been updated, and a set of links to external resources added at the level of domain entries. SCOP can be accessed at http://scop.mrc-lmb.cam.ac.uk/scop.
MeSH Terms
Animals
Databases, Protein
Evolution, Molecular
Genome
Information Storage and Retrieval
Internet
Protein Structure, Tertiary
Proteins/chemistry,classification,genetics
Authors & Affiliations
5 authors, click to expand affiliations / ORCID
Lo Conte Loredana
MRC Laboratory of Molecular Biology, Hills Road, Cambridge CB2 2QH, UK.
[email protected]
Brenner Steven E
Hubbard Tim J P
Chothia Cyrus
Murzin Alexey G
References (18)
18 references, click to expand
-
Crystal structure of the N-terminal domain of MukB: a protein involved in chromosome partitioning.
Structure. 1999 Oct 15;7(10):1181-7
PMID: 10545328
-
Hidden Markov models for detecting remote protein homologies.
Bioinformatics. 1998;14(10):846-56
PMID: 9927713
-
Homology a personal view on some of the problems.
Trends Genet. 2000 May;16(5):227-31
PMID: 10782117
-
Structural biology of Rad50 ATPase: ATP-driven conformational control in DNA double-strand break repair and the ABC-ATPase superfamily.
Cell. 2000 Jun 23;101(7):789-800
PMID: 10892749
-
Structural analysis of the neuronal SNARE protein syntaxin-1A.
Biochemistry. 2000 Jul 25;39(29):8470-9
PMID: 10913252
-
Crystal structure of the SMC head domain: an ABC ATPase with 900 residues antiparallel coiled-coil inserted.
J Mol Biol. 2001 Feb 9;306(1):25-35
PMID: 11178891
-
PartsList: a web-based system for dynamically ranking protein folds based on disparate attributes, including whole-genome expression and interaction information.
Nucleic Acids Res. 2001 Apr 15;29(8):1750-64
PMID: 11292848
-
A tour of structural genomics.
Nat Rev Genet. 2001 Oct;2(10):801-9
PMID: 11584296
-
Assignment of homology to genome sequences using a library of hidden Markov models that represent all proteins of known structure.
J Mol Biol. 2001 Nov 2;313(4):903-19
PMID: 11697912
-
The Protein Data Bank: unifying the archive.
Nucleic Acids Res. 2002 Jan 1;30(1):245-8
PMID: 11752306
-
ASTRAL compendium enhancements.
Nucleic Acids Res. 2002 Jan 1;30(1):260-3
PMID: 11752310
-
The Pfam protein families database.
Nucleic Acids Res. 2002 Jan 1;30(1):276-80
PMID: 11752314
-
The Protein Data Bank: a computer-based archival file for macromolecular structures.
J Mol Biol. 1977 May 25;112(3):535-42
PMID: 875032
-
SCOP: a structural classification of proteins database for the investigation of sequences and structures.
J Mol Biol. 1995 Apr 7;247(4):536-40
PMID: 7723011
-
Understanding protein structure: using scop for fold interpretation.
Methods Enzymol. 1996;266:635-43
PMID: 8743710
-
Three-dimensional structure of an evolutionarily conserved N-terminal domain of syntaxin 1A.
Cell. 1998 Sep 18;94(6):841-9
PMID: 9753330
-
The PRESAGE database for structural genomics.
Nucleic Acids Res. 1999 Jan 1;27(1):251-3
PMID: 9847193
-
Three-dimensional structure of the neuronal-Sec1-syntaxin 1a complex.
Nature. 2000 Mar 23;404(6776):355-62
PMID: 10746715