首页    期刊浏览 2024年11月28日 星期四
登录注册

文章基本信息

  • 标题:De novo peptide sequencing by deep learning
  • 本地全文:下载
  • 作者:Ngoc Hieu Tran ; Xianglilan Zhang ; Lei Xin
  • 期刊名称:Proceedings of the National Academy of Sciences
  • 印刷版ISSN:0027-8424
  • 电子版ISSN:1091-6490
  • 出版年度:2017
  • 卷号:114
  • 期号:31
  • 页码:8247-8252
  • DOI:10.1073/pnas.1705691114
  • 语种:English
  • 出版社:The National Academy of Sciences of the United States of America
  • 摘要:De novo peptide sequencing from tandem MS data is the key technology in proteomics for the characterization of proteins, especially for new sequences, such as mAbs. In this study, we propose a deep neural network model, DeepNovo, for de novo peptide sequencing. DeepNovo architecture combines recent advances in convolutional neural networks and recurrent neural networks to learn features of tandem mass spectra, fragment ions, and sequence patterns of peptides. The networks are further integrated with local dynamic programming to solve the complex optimization task of de novo sequencing. We evaluated the method on a wide variety of species and found that DeepNovo considerably outperformed state of the art methods, achieving 7.7–22.9% higher accuracy at the amino acid level and 38.1–64.0% higher accuracy at the peptide level. We further used DeepNovo to automatically reconstruct the complete sequences of antibody light and heavy chains of mouse, achieving 97.5–100% coverage and 97.2–99.5% accuracy, without assisting databases. Moreover, DeepNovo is retrainable to adapt to any sources of data and provides a complete end-to-end training and prediction solution to the de novo sequencing problem. Not only does our study extend the deep learning revolution to a new field, but it also shows an innovative approach in solving optimization problems by using deep learning and dynamic programming.
  • 关键词:deep learning ; MS ; de novo sequencing
国家哲学社会科学文献中心版权所有