The Ensembl gene annotation system

Database (Oxford). 2016 Jun 23:2016:baw093. doi: 10.1093/database/baw093. Print 2016.

Abstract

The Ensembl gene annotation system has been used to annotate over 70 different vertebrate species across a wide range of genome projects. Furthermore, it generates the automatic alignment-based annotation for the human and mouse GENCODE gene sets. The system is based on the alignment of biological sequences, including cDNAs, proteins and RNA-seq reads, to the target genome in order to construct candidate transcript models. Careful assessment and filtering of these candidate transcripts ultimately leads to the final gene set, which is made available on the Ensembl website. Here, we describe the annotation process in detail.Database URL: http://www.ensembl.org/index.html.

MeSH terms

  • Animals
  • Databases, Nucleic Acid*
  • Databases, Protein*
  • Humans
  • Internet*
  • Mice
  • Molecular Sequence Annotation / methods*