Home LiteratureArticle Details
PMID: 3449852 Published · ppublish English Journal Article

PseqIP: a nonredundant and exhaustive protein sequence data bank generated from 4 major existing collections.

Proteins ·Vol. 1 ·No. 1 ·1986-09-00 ·Pages 60-5

Claverie JM, Bricault L

Abstract

Four major protein sequence data collections (NBRF-PIR, PSD-Kyoto, PGtrans, and NEWAT) have been merged into a single nonredundant data bank called PseqIP. The data bank entries were automatically matched by a heuristic computer program relying on the fast computation of the number of tetrapeptides shared by two sequences. PseqIP 1.0 includes 6,068 different protein sequences for a total of 1,357,067 residues, representing most of the available sequence information to date. During the course of this work, we found about 600 occurrences of a protein sequence recorded with a one-amino-acid variation in at least two different data banks. A flat file (ASCII computer-readable format) version of PseqIP 1.0, well-suited for exhaustive homology searches and statistical sequence analysis, is available from our laboratory.

MeSH Terms
Algorithms Amino Acid Sequence Information Systems Molecular Sequence Data Proteins Software
Chemicals
Proteins
Authors & Affiliations
2 authors, click to expand affiliations / ORCID
Claverie J M
Computer Science Unit, Institut Pasteur, Paris, France.
Bricault L
Article Info
Journal
Proteins
Abbr.
Proteins
ISSN
0887-3585
Published
1986-09-00
Pages
60-5
Language
English
Region
United States
NLM ID
8700181
Subset
IM
Analysis Services
Analysis Services

Contact

No. 2 Wenbo Road, Zhangqiu District, Jinan, Shandong

Qilu Normal University · Genelibs Bioinformatics Lab

750 Shunhua Rd, Jinan

2F, Bldg F, University Science Park

Tel: 0531-88819269

WeChat Official Account

Follow our WeChat subscription account for real-time updates and the latest in medical and biological research.


Business Email

E-mail: [email protected]