Incorporating the human gene annotations in different databases significantly improved transcriptomic and genetic analyses

(Downloading may take up to 30 seconds. If the slide opens in your browser, select File -> Save As to save it.)

Click on image to view larger version.

FIGURE 1.
FIGURE 1.

Human transcriptomes in RefSeq, Ensembl, AceView, EnsAce, and AceEns. (A) The number of human genes and transcripts in each transcriptome. The counts for RefSeq are its unique genes and transcripts in UCSC for which the duplicated genes were removed. The Ensembl human transcriptome is the sum of its protein-coding and non-protein-coding transcripts (release 67 of GRCh37, corresponding to GENCODE 12). AceView transcripts that contained unknown bases of “N” were not taken into account. (B) Chromosome distribution of AceView human genes located in the intergenic and intronic regions of the Ensembl annotation. (C) Chromosome distribution of Ensembl human genes within the intergenic and intronic regions of AceView annotation. (D) Categories of the 18,083 Ensembl transcripts located in the AceView intergenic and intronic regions.

This Article

  1. RNA 19: 479-489