Загрузка...

Probabilistic variable-length segmentation of protein sequences for discriminative motif discovery (DiMotif) and sequence embedding (ProtVecX)

In this paper, we present peptide-pair encoding (PPE), a general-purpose probabilistic segmentation of protein sequences into commonly occurring variable-length sub-sequences. The idea of PPE segmentation is inspired by the byte-pair encoding (BPE) text compression algorithm, which has recently gain...

Полное описание

Сохранить в:
Библиографические подробности
Опубликовано в: :Sci Rep
Главные авторы: Asgari, Ehsaneddin, McHardy, Alice C., Mofrad, Mohammad R. K.
Формат: Artigo
Язык:Inglês
Опубликовано: Nature Publishing Group UK 2019
Предметы:
Online-ссылка:https://ncbi.nlm.nih.gov/pmc/articles/PMC6401088/
https://ncbi.nlm.nih.gov/pubmed/30837494
https://ncbi.nlm.nih.govhttp://dx.doi.org/10.1038/s41598-019-38746-w
Метки: Добавить метку
Нет меток, Требуется 1-ая метка записи!