Quick Search:

uzh logo
Browse by:
bullet
bullet
bullet
bullet

Zurich Open Repository and Archive 

Permanent URL to this publication: http://dx.doi.org/10.5167/uzh-28688

Muller, J; Szklarczyk, D; Julien, P; Letunic, I; Roth, A; Kuhn, M; Powell, S; von Mering, C; Doerks, T; Jensen, L J; Bork, P (2010). eggNOG v2.0: extending the evolutionary genealogy of genes with enhanced non-supervised orthologous groups, species and functional annotations. Nucleic Acids Research, 38 (Da:D190-D195.

[img]
Preview
PDF
2MB

Abstract

The identification of orthologous relationships forms the basis for most comparative genomics studies. Here, we present the second version of the eggNOG database, which contains orthologous groups (OGs) constructed through identification of reciprocal best BLAST matches and triangular linkage clustering. We applied this procedure to 630 complete genomes (529 bacteria, 46 archaea and 55 eukaryotes), which is a 2-fold increase relative to the previous version. The pipeline yielded 224,847 OGs, including 9724 extended versions of the original COG and KOG. We computed OGs for different levels of the tree of life; in addition to the species groups included in our first release (i.e. fungi, metazoa, insects, vertebrates and mammals), we have now constructed OGs for archaea, fishes, rodents and primates. We automatically annotate the non-supervised orthologous groups (NOGs) with functional descriptions, protein domains, and functional categories as defined initially for the COG/KOG database. In-depth analysis is facilitated by precomputed high-quality multiple sequence alignments and maximum-likelihood trees for each of the available OGs. Altogether, eggNOG covers 2,242 035 proteins (built from 2,590,259 proteins) and provides a broad functional description for at least 1,966,709 (88%) of them. Users can access the complete set of orthologous groups via a web interface at: http://eggnog.embl.de.

Item Type:Journal Article, refereed, original work
Communities & Collections:07 Faculty of Science > Institute of Molecular Life Sciences
08 University Research Priority Programs > Systems Biology / Functional Genomics
DDC:570 Life sciences; biology
Language:English
Date:2010
Deposited On:21 Mar 2010 09:48
Last Modified:27 Nov 2013 18:38
Publisher:Oxford University Press
ISSN:0305-1048
Publisher DOI:10.1093/nar/gkp951
PubMed ID:19900971
Citations:Web of Science®. Times Cited: 104
Google Scholar™
Scopus®. Citation Count: 109

Users (please log in): suggest update or correction for this item

Repository Staff Only: item control page