UCL Discovery
UCL home » Library Services » Electronic resources » UCL Discovery

Orthologous Matrix (OMA) algorithm 2.0: more robust to asymmetric evolutionary rates and more scalable hierarchical orthologous group inference

Train, C-M; Glover, NM; Gonnet, GH; Altenhoff, AM; Dessimoz, C; (2017) Orthologous Matrix (OMA) algorithm 2.0: more robust to asymmetric evolutionary rates and more scalable hierarchical orthologous group inference. Bioinformatics , 33 (14) i75-i82. 10.1093/bioinformatics/btx229. Green open access

[thumbnail of Dessimoz_Orthologous Matrix (OMA) algorithm 2.0_VoR.pdf]
Preview
Text
Dessimoz_Orthologous Matrix (OMA) algorithm 2.0_VoR.pdf - Published Version

Download (1MB) | Preview

Abstract

MOTIVATION: Accurate orthology inference is a fundamental step in many phylogenetics and comparative analysis. Many methods have been proposed, including OMA (Orthologous MAtrix). Yet substantial challenges remain, in particular in coping with fragmented genes or genes evolving at different rates after duplication, and in scaling to large datasets. With more and more genomes available, it is necessary to improve the scalability and robustness of orthology inference methods. RESULTS: We present improvements in the OMA algorithm: (i) refining the pairwise orthology inference step to account for same-species paralogs evolving at different rates, and (ii) minimizing errors in the pairwise orthology verification step by testing the consistency of pairwise distance estimates, which can be problematic in the presence of fragmentary sequences. In addition we introduce a more scalable procedure for hierarchical orthologous group (HOG) clustering, which are several orders of magnitude faster on large datasets. Using the Quest for Orthologs consortium orthology benchmark service, we show that these changes translate into substantial improvement on multiple empirical datasets. AVAILABILITY AND IMPLEMENTATION: This new OMA 2.0 algorithm is used in the OMA database (http://omabrowser.org) from the March 2017 release onwards, and can be run on custom genomes using OMA standalone version 2.0 and above (http://omabrowser.org/standalone).

Type: Article
Title: Orthologous Matrix (OMA) algorithm 2.0: more robust to asymmetric evolutionary rates and more scalable hierarchical orthologous group inference
Open access status: An open access version is available from UCL Discovery
DOI: 10.1093/bioinformatics/btx229
Publisher version: http://doi.org/10.1093/bioinformatics/btx229
Language: English
Additional information: © The Author 2017. Published by Oxford University Press. All rights reserved. For Permissions, please e-mail: journals.permissions@oup.com. This version is the author accepted manuscript. For information on re-use, please refer to the publisher’s terms and conditions.
Keywords: Science & Technology, Life Sciences & Biomedicine, Technology, Physical Sciences, Biochemical Research Methods, Biotechnology & Applied Microbiology, Computer Science, Interdisciplinary Applications, Mathematical & Computational Biology, Statistics & Probability, Biochemistry & Molecular Biology, Computer Science, Mathematics, PROTEIN FAMILIES, GENE, IDENTIFICATION, DATABASE, QUEST, HITS, TREE
UCL classification: UCL
UCL > Provost and Vice Provost Offices > School of Life and Medical Sciences
UCL > Provost and Vice Provost Offices > School of Life and Medical Sciences > Faculty of Life Sciences
UCL > Provost and Vice Provost Offices > School of Life and Medical Sciences > Faculty of Life Sciences > Div of Biosciences
UCL > Provost and Vice Provost Offices > School of Life and Medical Sciences > Faculty of Life Sciences > Div of Biosciences > Genetics, Evolution and Environment
URI: https://discovery.ucl.ac.uk/id/eprint/1566831
Downloads since deposit
69Downloads
Download activity - last month
Download activity - last 12 months
Downloads by country - last 12 months

Archive Staff Only

View Item View Item