Complete de Novo Assembly of Monoclonal Antibody Sequences

Ngoc Hieu Tran, M. Ziaur Rahman, Lin He, Lei Xin, Baozhen Shan, Ming Li


64 被引用数 (Scopus)


De novo protein sequencing is one of the key problems in mass spectrometry-based proteomics, especially for novel proteins such as monoclonal antibodies for which genome information is often limited or not available. However, due to limitations in peptides fragmentation and coverage, as well as ambiguities in spectra interpretation, complete de novo assembly of unknown protein sequences still remains challenging. To address this problem, we propose an integrated system, ALPS, which for the first time can automatically assemble full-length monoclonal antibody sequences. Our system integrates de novo sequencing peptides, their quality scores and error-correction information from databases into a weighted de Bruijn graph to assemble protein sequences. We evaluated ALPS performance on two antibody data sets, each including a heavy chain and a light chain. The results show that ALPS was able to assemble three complete monoclonal antibody sequences of length 216-441 AA, at 100% coverage, and 96.64-100% accuracy.

ジャーナルScientific reports
出版ステータスPublished - 8月 26 2016

ASJC Scopus subject areas

  • 一般


「Complete de Novo Assembly of Monoclonal Antibody Sequences」の研究トピックを掘り下げます。これらがまとまってユニークなフィンガープリントを構成します。