Home LiteratureArticle Details
PMID: 23228854 Published · ppublish English Journal Article Research Support, Non-U.S. Gov't

Identifying differentially spliced genes from two groups of RNA-seq samples.

Gene ·Vol. 518 ·No. 1 ·2013-04-10 ·Pages 164-70

Wang W, Qin Z, Feng Z, Wang X, Zhang X

Abstract

Recent study revealed that most human genes have alternative splicing and can produce multiple isoforms of transcripts. Differences in the relative abundance of the isoforms of a gene can have significant biological consequences. Identifying genes that are differentially spliced between two groups of RNA-sequencing samples is an important basic task in the study of transcriptomes with next-generation sequencing technology. We use the negative binomial (NB) distribution to model sequencing reads on exons, and propose a NB-statistic to detect differentially spliced genes between two groups of samples by comparing read counts on all exons. The method opens a new exon-based approach instead of isoform-based approach for the task. It does not require information about isoform composition, nor need the estimation of isoform expression. Experiments on simulated data and real RNA-seq data of human kidney and liver samples illustrated the method's good performance and applicability. It can also detect previously unknown alternative splicing events, and highlight exons that are most likely differentially spliced between the compared samples. We developed an NB-statistic method that can detect differentially spliced genes between two groups of samples without using a prior knowledge on the annotation of alternative splicing. It does not need to infer isoform structure or to estimate isoform expression. It is a useful method designed for comparing two groups of RNA-seq samples. Besides identifying differentially spliced genes, the method can highlight on the exons that contribute the most to the differential splicing. We developed a software tool called DSGseq for the presented method available at http://bioinfo.au.tsinghua.edu.cn/software/DSGseq.

MeSH Terms
Alternative Splicing Computer Simulation Exons Humans Kidney/physiology Liver/physiology Models, Genetic Models, Statistical ROC Curve Sequence Analysis, RNA/methods,statistics & numerical data Software Transcriptome
Authors & Affiliations
5 authors, click to expand affiliations / ORCID
Wang Weichen
MOE Key Laboratory of Bioinformatics, Bioinformatics Division and Center for Synthetic and Systems Biology, TNLIST, Department of Automation, Tsinghua University, Beijing 100084, China. [email protected]
Qin Zhiyi
Feng Zhixing
Wang Xi
Zhang Xuegong
Article Info
Journal
Gene
Abbr.
Gene
ISSN
1879-0038
Published
2013-04-10
Epub
2012-00-08
Pages
164-70
Language
English
Region
Netherlands
NLM ID
7706761
Subset
IM
Analysis Services
Analysis Services

Contact

No. 2 Wenbo Road, Zhangqiu District, Jinan, Shandong

Qilu Normal University · Genelibs Bioinformatics Lab

750 Shunhua Rd, Jinan

2F, Bldg F, University Science Park

Tel: 0531-88819269

WeChat Official Account

Follow our WeChat subscription account for real-time updates and the latest in medical and biological research.


Business Email

E-mail: [email protected]