Efficient inference of homologs in large eukaryotic pan-proteomes

Autor:	Siavash Sheikhizadeh Anari, Dick de Ridder, M. Eric Schranz, Sandra Smit
Jazyk:	angličtina
Rok vydání:	2018
Předmět:	Pan-genome Protein similarity Homologous genes Orthology k-mer Computer applications to medicine. Medical informatics R858-859.7 Biology (General) QH301-705.5
Zdroj:	BMC Bioinformatics, Vol 19, Iss 1, Pp 1-11 (2018)
Druh dokumentu:	article
ISSN:	1471-2105
DOI:	10.1186/s12859-018-2362-4
Popis:	Abstract Background Identification of homologous genes is fundamental to comparative genomics, functional genomics and phylogenomics. Extensive public homology databases are of great value for investigating homology but need to be continually updated to incorporate new sequences. As new sequences are rapidly being generated, there is a need for efficient standalone tools to detect homologs in novel data. Results To address this, we present a fast method for detecting homology groups across a large number of individuals and/or species. We adopted a k-mer based approach which considerably reduces the number of pairwise protein alignments without sacrificing sensitivity. We demonstrate accuracy, scalability, efficiency and applicability of the presented method for detecting homology in large proteomes of bacteria, fungi, plants and Metazoa. Conclusions We clearly observed the trade-off between recall and precision in our homology inference. Favoring recall or precision strongly depends on the application. The clustering behavior of our program can be optimized for particular applications by altering a few key parameters. The program is available for public use at https://github.com/sheikhizadeh/pantools as an extension to our pan-genomic analysis tool, PanTools.
Databáze:	Directory of Open Access Journals
Externí odkaz:	https://doaj.org/article/d81c481b08054cdfa979ca5635f2eb31 Zobrazit plný text záznamu View record in DOAJ Plný text ve formátu PDF Plný text ve formátu HTML
Nepřihlášeným uživatelům se plný text nezobrazuje	K zobrazení výsledku je třeba se přihlásit.