Your browser doesn't support javascript.
loading
A unified gene catalog for the laboratory mouse reference genome.
Zhu, Y; Richardson, J E; Hale, P; Baldarelli, R M; Reed, D J; Recla, J M; Sinclair, R; Reddy, T B K; Bult, C J.
Afiliación
  • Zhu Y; The Jackson Laboratory, RL13, 600 Main Street, Bar Harbor, ME, 04609, USA.
Mamm Genome ; 26(7-8): 295-304, 2015 Aug.
Article en En | MEDLINE | ID: mdl-26084703
We report here a semi-automated process by which mouse genome feature predictions and curated annotations (i.e., genes, pseudogenes, functional RNAs, etc.) from Ensembl, NCBI and Vertebrate Genome Annotation database (Vega) are reconciled with the genome features in the Mouse Genome Informatics (MGI) database (http://www.informatics.jax.org) into a comprehensive and non-redundant catalog. Our gene unification method employs an algorithm (fjoin--feature join) for efficient detection of genome coordinate overlaps among features represented in two annotation data sets. Following the analysis with fjoin, genome features are binned into six possible categories (1:1, 1:0, 0:1, 1:n, n:1, n:m) based on coordinate overlaps. These categories are subsequently prioritized for assessment of annotation equivalencies and differences. The version of the unified catalog reported here contains more than 59,000 entries, including 22,599 protein-coding coding genes, 12,455 pseudogenes, and 24,007 other feature types (e.g., microRNAs, lincRNAs, etc.). More than 23,000 of the entries in the MGI gene catalog have equivalent gene models in the annotation files obtained from NCBI, Vega, and Ensembl. 12,719 of the features are unique to NCBI relative to Ensembl/Vega; 11,957 are unique to Ensembl/Vega relative to NCBI, and 3095 are unique to MGI. More than 4000 genome features fall into categories that require manual inspection to resolve structural differences in the gene models from different annotation sources. Using the MGI unified gene catalog, researchers can easily generate a comprehensive report of mouse genome features from a single source and compare the details of gene and transcript structure using MGI's mouse genome browser.
Asunto(s)

Texto completo: 1 Colección: 01-internacional Base de datos: MEDLINE Asunto principal: Programas Informáticos / Genoma / Genómica / Bases de Datos Genéticas Tipo de estudio: Prognostic_studies Límite: Animals Idioma: En Revista: Mamm Genome Asunto de la revista: GENETICA Año: 2015 Tipo del documento: Article País de afiliación: Estados Unidos Pais de publicación: Estados Unidos

Texto completo: 1 Colección: 01-internacional Base de datos: MEDLINE Asunto principal: Programas Informáticos / Genoma / Genómica / Bases de Datos Genéticas Tipo de estudio: Prognostic_studies Límite: Animals Idioma: En Revista: Mamm Genome Asunto de la revista: GENETICA Año: 2015 Tipo del documento: Article País de afiliación: Estados Unidos Pais de publicación: Estados Unidos