Unknown

Dataset Information

0

SparseIso: a novel Bayesian approach to identify alternatively spliced isoforms from RNA-seq data.


ABSTRACT: Motivation:Recent advances in high-throughput RNA sequencing (RNA-seq) technologies have made it possible to reconstruct the full transcriptome of various types of cells. It is important to accurately assemble transcripts or identify isoforms for an improved understanding of molecular mechanisms in biological systems. Results:We have developed a novel Bayesian method, SparseIso, to reliably identify spliced isoforms from RNA-seq data. A spike-and-slab prior is incorporated into the Bayesian model to enforce the sparsity for isoform identification, effectively alleviating the problem of overfitting. A Gibbs sampling procedure is further developed to simultaneously identify and quantify transcripts from RNA-seq data. With the sampling approach, SparseIso estimates the joint distribution of all candidate transcripts, resulting in a significantly improved performance in detecting lowly expressed transcripts and multiple expressed isoforms of genes. Both simulation study and real data analysis have demonstrated that the proposed SparseIso method significantly outperforms existing methods for improved transcript assembly and isoform identification. Availability and implementation:The SparseIso package is available at http://github.com/henryxushi/SparseIso. Contact:xuan@vt.edu. Supplementary information:Supplementary data are available at Bioinformatics online.

SUBMITTER: Shi X 

PROVIDER: S-EPMC5870564 | biostudies-literature | 2018 Jan

REPOSITORIES: biostudies-literature

altmetric image

Publications

SparseIso: a novel Bayesian approach to identify alternatively spliced isoforms from RNA-seq data.

Shi Xu X   Wang Xiao X   Wang Tian-Li TL   Hilakivi-Clarke Leena L   Clarke Robert R   Xuan Jianhua J  

Bioinformatics (Oxford, England) 20180101 1


<h4>Motivation</h4>Recent advances in high-throughput RNA sequencing (RNA-seq) technologies have made it possible to reconstruct the full transcriptome of various types of cells. It is important to accurately assemble transcripts or identify isoforms for an improved understanding of molecular mechanisms in biological systems.<h4>Results</h4>We have developed a novel Bayesian method, SparseIso, to reliably identify spliced isoforms from RNA-seq data. A spike-and-slab prior is incorporated into th  ...[more]

Similar Datasets

| S-EPMC3820534 | biostudies-literature
| S-EPMC3330053 | biostudies-literature
| S-EPMC2868064 | biostudies-literature
| S-EPMC1616956 | biostudies-literature
| S-EPMC2364638 | biostudies-literature
| S-EPMC2790411 | biostudies-literature
| S-EPMC2839007 | biostudies-literature
| S-EPMC2651755 | biostudies-literature
| S-EPMC5636827 | biostudies-literature