SignalP 6.0 achieves signal peptide prediction across all types using protein language models

2021 
Abstract Signal peptides (SPs) are short amino acid sequences that control protein secretion and translocation in all living organisms. As experimental characterization of SPs is costly, prediction algorithms are applied to predict them from sequence data. However, existing methods are unable to detect all known types of SPs. We introduce SignalP 6.0, the first model capable of detecting all five SP types. Additionally, the model accurately identifies the positions of regions within SPs, revealing the defining biochemical properties that underlie the function of SPs in vivo. Results show that SignalP 6.0 has improved prediction performance, and is the first model to be applicable to metagenomic data. SignalP 6.0 is available at https://services.healthtech.dtu.dk/service.php?SignalP-6.0
    • Correction
    • Source
    • Cite
    • Save
    • Machine Reading By IdeaReader
    35
    References
    2
    Citations
    NaN
    KQI
    []