PlumX Metrics
Embed PlumX Metrics

Improved rat genome gene prediction by integration of ESTs with RNA-Seq information

Bioinformatics, ISSN: 1460-2059, Vol: 31, Issue: 1, Page: 25-32
2015
  • 5
    Citations
  • 0
    Usage
  • 38
    Captures
  • 0
    Mentions
  • 0
    Social Media
Metric Options:   Counts1 Year3 Year

Metrics Details

Article Description

Motivation: RNA-Seq (also called whole-transcriptome sequencing) is an emerging technology that uses the capabilities of next-generation sequencing to detect and quantify entire transcripts. One of its important applications is the improvement of existing genome annotations. RNA-Seq provides rapid, comprehensive and cost-effective tools for the discovery of novel genes and transcripts compared with expressed sequence tag (EST), which is instrumental in gene discovery and gene sequence determination. The rat is widely used as a laboratory disease model, but has a less well-annotated genome as compared with humans and mice. In this study, we incorporated deep RNA-Seq data from three rat tissues - bone marrow, brain and kidney - with EST data to improve the annotation of the rat genome. Results: Our analysis identified 32 197 transcripts, including 13 461 known transcripts, 13 934 novel isoforms and 4802 new genes, which almost doubled the numbers of transcripts in the current public rat genome database (rn5). Comparisons of our predicted protein-coding gene sets with those in public datasets suggest that RNA-Seq significantly improves genome annotation and identifies novel genes and isoforms in the rat. Importantly, the large majority of novel genes and isoforms are supported by direct evidence of RNA-Seq experiments. These predicted genes were integrated into the Rat Genome Database (RGD) and can serve as an important resource for functional studies in the research community. Availability and implementation: The predicted genes are available at http://rgd.mcw.edu. Contact: or pliu@mcw.edu or yanlu76@zju.edu.cn Supplementary information: Supplementary data are available at Bioinformatics online.

Provide Feedback

Have ideas for a new metric? Would you like to see something else here?Let us know