SignalP 6.0 achieves signal peptide prediction across all types using protein language models

Autor: Alexander Rosenberg Johansen, Ole Winther, Søren Brunak, Teufel F, Konstantinos D. Tsirigos, Henrik Nielsen, Gislason Mh, Armenteros Jja, Pihl Si, von Heijne G
Rok vydání: 2021
Předmět:
Popis: Signal peptides (SPs) are short amino acid sequences that control protein secretion and translocation in all living organisms. As experimental characterization of SPs is costly, prediction algorithms are applied to predict them from sequence data. However, existing methods are unable to detect all known types of SPs. We introduce SignalP 6.0, the first model capable of detecting all five SP types. Additionally, the model accurately identifies the positions of regions within SPs, revealing the defining biochemical properties that underlie the function of SPs in vivo. Results show that SignalP 6.0 has improved prediction performance, and is the first model to be applicable to metagenomic data.SignalP 6.0 is available at https://services.healthtech.dtu.dk/service.php?SignalP-6.0
Databáze: OpenAIRE