Inferring hierarchical orthologous groups from orthologous gene pairs
- PMID: 23342000
- PMCID: PMC3544860
- DOI: 10.1371/journal.pone.0053786
Inferring hierarchical orthologous groups from orthologous gene pairs
Abstract
Hierarchical orthologous groups are defined as sets of genes that have descended from a single common ancestor within a taxonomic range of interest. Identifying such groups is useful in a wide range of contexts, including inference of gene function, study of gene evolution dynamics and comparative genomics. Hierarchical orthologous groups can be derived from reconciled gene/species trees but, this being a computationally costly procedure, many phylogenomic databases work on the basis of pairwise gene comparisons instead ("graph-based" approach). To our knowledge, there is only one published algorithm for graph-based hierarchical group inference, but both its theoretical justification and performance in practice are as of yet largely uncharacterised. We establish a formal correspondence between the orthology graph and hierarchical orthologous groups. Based on that, we devise GETHOGs ("Graph-based Efficient Technique for Hierarchical Orthologous Groups"), a novel algorithm to infer hierarchical groups directly from the orthology graph, thus without needing gene tree inference nor gene/species tree reconciliation. GETHOGs is shown to correctly reconstruct hierarchical orthologous groups when applied to perfect input, and several extensions with stringency parameters are provided to deal with imperfect input data. We demonstrate its competitiveness using both simulated and empirical data. GETHOGs is implemented as a part of the freely-available OMA standalone package (http://omabrowser.org/standalone). Furthermore, hierarchical groups inferred by GETHOGs ("OMA HOGs") on >1,000 genomes can be interactively queried via the OMA browser (http://omabrowser.org).
Conflict of interest statement
Figures
Similar articles
-
Orthologous Matrix (OMA) algorithm 2.0: more robust to asymmetric evolutionary rates and more scalable hierarchical orthologous group inference.Bioinformatics. 2017 Jul 15;33(14):i75-i82. doi: 10.1093/bioinformatics/btx229. Bioinformatics. 2017. PMID: 28881964 Free PMC article.
-
Identifying orthologs with OMA: A primer.F1000Res. 2020 Jan 17;9:27. doi: 10.12688/f1000research.21508.1. eCollection 2020. F1000Res. 2020. PMID: 32089838 Free PMC article.
-
Inferring Orthology and Paralogy.Methods Mol Biol. 2019;1910:149-175. doi: 10.1007/978-1-4939-9074-0_5. Methods Mol Biol. 2019. PMID: 31278664
-
Inferring orthology and paralogy.Methods Mol Biol. 2012;855:259-79. doi: 10.1007/978-1-61779-582-4_9. Methods Mol Biol. 2012. PMID: 22407712 Review.
-
The quest for orthologs: finding the corresponding gene across genomes.Trends Genet. 2008 Nov;24(11):539-51. doi: 10.1016/j.tig.2008.08.009. Epub 2008 Sep 24. Trends Genet. 2008. PMID: 18819722 Review.
Cited by
-
Inferring Interaction Networks from Transcriptomic Data: Methods and Applications.Methods Mol Biol. 2024;2812:11-37. doi: 10.1007/978-1-0716-3886-6_2. Methods Mol Biol. 2024. PMID: 39068355
-
Protein-Coding Gene Families in Prokaryote Genome Comparisons.Methods Mol Biol. 2024;2802:33-55. doi: 10.1007/978-1-0716-3838-5_2. Methods Mol Biol. 2024. PMID: 38819555
-
Quality assessment of gene repertoire annotations with OMArk.Nat Biotechnol. 2024 Feb 21. doi: 10.1038/s41587-024-02147-w. Online ahead of print. Nat Biotechnol. 2024. PMID: 38383603
-
De novo genome assembly and comparative genomics for the colonial ascidian Botrylloides violaceus.G3 (Bethesda). 2023 Sep 30;13(10):jkad181. doi: 10.1093/g3journal/jkad181. G3 (Bethesda). 2023. PMID: 37555394 Free PMC article.
-
Nano-Strainer: A workflow for the identification of single-copy nuclear loci for plant systematic studies, using target capture kits and Oxford Nanopore long reads.Ecol Evol. 2023 Jul 18;13(7):e10190. doi: 10.1002/ece3.10190. eCollection 2023 Jul. Ecol Evol. 2023. PMID: 37475726 Free PMC article.
References
-
- Fitch WM (1970) Distinguishing homologous from analogous proteins. Syst Zool 19: 99–113. - PubMed
-
- Altenhoff AM, Dessimoz C (2012) Inferring orthology and paralogy. In: Anisimova M, editor, Evolutionary Genomics, Clifton, NJ, USA: Methods in Molecular Biology. 259–279. - PubMed
-
- Wall DP, Fraser HB, Hirsh AE (2003) Detecting putative orthologs. Bioinformatics 19: 1710–1711. - PubMed
Publication types
MeSH terms
Grants and funding
LinkOut - more resources
Full Text Sources
Other Literature Sources