PLoS Computational Biology (Jan 2012)

Resolving the ortholog conjecture: orthologs tend to be weakly, but significantly, more similar in function than paralogs.

  • Adrian M Altenhoff,
  • Romain A Studer,
  • Marc Robinson-Rechavi,
  • Christophe Dessimoz

DOI
https://doi.org/10.1371/journal.pcbi.1002514
Journal volume & issue
Vol. 8, no. 5
p. e1002514

Abstract

Read online

The function of most proteins is not determined experimentally, but is extrapolated from homologs. According to the "ortholog conjecture", or standard model of phylogenomics, protein function changes rapidly after duplication, leading to paralogs with different functions, while orthologs retain the ancestral function. We report here that a comparison of experimentally supported functional annotations among homologs from 13 genomes mostly supports this model. We show that to analyze GO annotation effectively, several confounding factors need to be controlled: authorship bias, variation of GO term frequency among species, variation of background similarity among species pairs, and propagated annotation bias. After controlling for these biases, we observe that orthologs have generally more similar functional annotations than paralogs. This is especially strong for sub-cellular localization. We observe only a weak decrease in functional similarity with increasing sequence divergence. These findings hold over a large diversity of species; notably orthologs from model organisms such as E. coli, yeast or mouse have conserved function with human proteins.