Home LiteratureArticle Details
PMID: 17135202 Published · ppublish English Journal Article Research Support, N.I.H., Intramural

CDD: a conserved domain database for interactive domain family analysis.

Nucleic acids research ·Vol. 35 ·No. Database issue ·2007-01-00 ·Pages D237-40

Marchler-Bauer A, Anderson JB, Derbyshire MK, DeWeese-Scott C, Gonzales NR, Gwadz M, Hao L, He S, Hurwitz DI, Jackson JD, Ke Z, Krylov D, Lanczycki CJ, Liebert CA, Liu C, Lu F, Lu S, Marchler GH, Mullokandov M, Song JS, Thanki N, Yamashita RA, Yin JJ, Zhang D, Bryant SH

Abstract

The conserved domain database (CDD) is part of NCBI's Entrez database system and serves as a primary resource for the annotation of conserved domain footprints on protein sequences in Entrez. Entrez's global query interface can be accessed at http://www.ncbi.nlm.nih.gov/Entrez and will search CDD and many other databases. Domain annotation for proteins in Entrez has been pre-computed and is readily available in the form of 'Conserved Domain' links. Novel protein sequences can be scanned against CDD using the CD-Search service; this service searches databases of CDD-derived profile models with protein sequence queries using BLAST heuristics, at http://www.ncbi.nlm.nih.gov/Structure/cdd/wrpsb.cgi. Protein query sequences submitted to NCBI's protein BLAST search service are scanned for conserved domain signatures by default. The CDD collection contains models imported from Pfam, SMART and COG, as well as domain models curated at NCBI. NCBI curated models are organized into hierarchies of domains related by common descent. Here we report on the status of the curation effort and present a novel helper application, CDTree, which enables users of the CDD resource to examine curated hierarchies. More importantly, CDD and CDTree used in concert, serve as a powerful tool in protein classification, as they allow users to analyze protein sequences in the context of domain family hierarchies.

MeSH Terms
Amino Acid Sequence Animals Conserved Sequence Databases, Protein Internet Phylogeny Protein Structure, Tertiary/genetics Proteins/classification Sequence Analysis, Protein User-Computer Interface
Chemicals
Proteins
Authors & Affiliations
25 authors, click to expand affiliations / ORCID
Marchler-Bauer Aron
National Center for Biotechnology Information, National Library of Medicine, National Institutes of Health Building 38 A, Room 8N805, 8600 Rockville Pike, Bethesda, MD 20894, USA. [email protected]
Anderson John B
Derbyshire Myra K
DeWeese-Scott Carol
Gonzales Noreen R
Gwadz Marc
Hao Luning
He Siqian
Hurwitz David I
Jackson John D
Ke Zhaoxi
Krylov Dmitri
Lanczycki Christopher J
Liebert Cynthia A
Liu Chunlei
Lu Fu
Lu Shennan
Marchler Gabriele H
Mullokandov Mikhail
Song James S
Thanki Narmada
Yamashita Roxanne A
Yin Jodie J
Zhang Dachuan
Bryant Stephen H
References (10)
10 references, click to expand
  1. Cn3D: sequence and structure views for Entrez.
    Trends Biochem Sci. 2000 Jun;25(6):300-2 PMID: 10838572
  2. CDART: protein homology by domain architecture.
    Genome Res. 2002 Oct;12(10):1619-23 PMID: 12368255
  3. The COG database: an updated version includes eukaryotes.
    BMC Bioinformatics. 2003 Sep 11;4:41 PMID: 12969510
  4. CD-Search: protein domain annotations on the fly.
    Nucleic Acids Res. 2004 Jul 1;32(Web Server issue):W327-31 PMID: 15215404
  5. Gapped BLAST and PSI-BLAST: a new generation of protein database search programs.
    Nucleic Acids Res. 1997 Sep 1;25(17):3389-402 PMID: 9254694
  6. CDD: a Conserved Domain Database for protein classification.
    Nucleic Acids Res. 2005 Jan 1;33(Database issue):D192-6 PMID: 15608175
  7. Database resources of the National Center for Biotechnology Information.
    Nucleic Acids Res. 2006 Jan 1;34(Database issue):D173-80 PMID: 16381840
  8. Pfam: clans, web tools and services.
    Nucleic Acids Res. 2006 Jan 1;34(Database issue):D247-51 PMID: 16381856
  9. SMART 5: domains in the context of genomes and networks.
    Nucleic Acids Res. 2006 Jan 1;34(Database issue):D257-60 PMID: 16381859
  10. Functional classification using phylogenomic inference.
    PLoS Comput Biol. 2006 Jun 30;2(6):e77 PMID: 16846248
Article Info
Journal
Nucleic acids research
Abbr.
Nucleic Acids Res
ISSN
1362-4962
Published
2007-01-00
Epub
2006-00-29
Pages
D237-40
Language
English
Region
England
NLM ID
0411011
PMCID
PMC1751546
Subset
IM
Grants
Intramural NIH HHS · United States
Analysis Services
Analysis Services

Contact

No. 2 Wenbo Road, Zhangqiu District, Jinan, Shandong

Qilu Normal University · Genelibs Bioinformatics Lab

750 Shunhua Rd, Jinan

2F, Bldg F, University Science Park

Tel: 0531-88819269

WeChat Official Account

Follow our WeChat subscription account for real-time updates and the latest in medical and biological research.


Business Email

E-mail: [email protected]