Home LiteratureArticle Details
PMID: 3447154 Published · ppublish English Journal Article Research Support, Non-U.S. Gov't Research Support, U.S. Gov't, P.H.S.

A standardized format for sequence data exchange.

Protein sequences & data analysis ·Vol. 1 ·No. 1 ·1987-00-00 ·Pages 27-39

George DG, Mewes HW, Kihara H

Abstract

At present there is no agreement upon a standard format for the presentation of sequence data; each of the major sequence databases has adopted their own format. As a result, efforts to pool these data and to develop software to manipulate the data have been hampered. A significant amount of software development time must be invested to handle the incompatibilities among these formats before software to solve biologically interesting problems can be implemented. In principle, the development of a standard format by the database distributors would be the best solution. However, because the databases have invested years of effort in the development of procedures specifically tailored to their own format, they are reluctant to change. Insisting that they convert to a new format would place an extreme burden on the already overtaxed resources of these groups. Furthermore, for certain specialized applications it is more efficient to present the data in nonstandard formats. An alternative solution is presented here. Rather than develop a single standard format for all sequence data, a standardized exchange format has been developed. This format was designed to serve as a common interface between the major formats currently in use. Data can be easily converted to and from it without significant loss of information. This alleviates difficulties inherent in dealing with multiple formats while preserving the local formats of the various databases.

MeSH Terms
Amino Acid Sequence Base Sequence Information Systems/standards Molecular Sequence Data
Authors & Affiliations
3 authors, click to expand affiliations / ORCID
George D G
Protein Identification Resource, Georgetown University Medical Center, Washington, DC 20007.
Mewes H W
Kihara H
Article Info
Journal
Protein sequences & data analysis
Abbr.
Protein Seq Data Anal
ISSN
0931-9506
Published
1987-00-00
Pages
27-39
Language
English
Region
Germany
NLM ID
8800894
Subset
IM
Grants
NCRR NIH HHS · RR01821 · United States
External Links
PubMed source
Analysis Services
Analysis Services

Contact

No. 2 Wenbo Road, Zhangqiu District, Jinan, Shandong

Qilu Normal University · Genelibs Bioinformatics Lab

750 Shunhua Rd, Jinan

2F, Bldg F, University Science Park

Tel: 0531-88819269

WeChat Official Account

Follow our WeChat subscription account for real-time updates and the latest in medical and biological research.


Business Email

E-mail: [email protected]