Predicting antibody binders and generating synthetic antibodies using deep learning

Yoong Wearn Lim; Adam S. Adler; David S. Johnson

doi:10.1080/19420862.2022.2069075

mAbs (Dec 2022)

Predicting antibody binders and generating synthetic antibodies using deep learning

Yoong Wearn Lim,
Adam S. Adler,
David S. Johnson

Affiliations

Yoong Wearn Lim: GigaGen Inc. (A Grifols Company), South San Francisco, CA, USA
Adam S. Adler: GigaGen Inc. (A Grifols Company), South San Francisco, CA, USA
David S. Johnson: GigaGen Inc. (A Grifols Company), South San Francisco, CA, USA

DOI: https://doi.org/10.1080/19420862.2022.2069075
Journal volume & issue: Vol. 14, no. 1

Abstract

Read online

The antibody drug field has continually sought improvements to methods for candidate discovery and engineering. Historically, most such methods have been laboratory-based, but informatics methods have recently started to make an impact. Deep learning, a subfield of machine learning, is rapidly gaining prominence in the biomedical research. Recent advances in microfluidics technologies and next-generation sequencing have not only revolutionized therapeutic antibody discovery, but also contributed to a vast amount of antibody repertoire sequencing data, providing opportunities for deep learning-based applications. Previously, we used microfluidics, yeast display, and deep sequencing to generate a panel of binder and non-binder antibody sequences to the cancer immunotherapy targets PD-1 and CTLA-4. Here we encoded the antibody light and heavy chain complementarity-determining regions (CDR3s) into antibody images, then built and trained convolutional neural network models to classify binders and non-binders. To improve model interpretability, we performed in silico mutagenesis to identify CDR3 residues that were important for binder classification. We further built generative deep learning models using generative adversarial network models to produce synthetic antibodies against PD-1 and CTLA-4. Our models generated variable length CDR3 sequences that resemble real sequences. Overall, our study demonstrates that deep learning methods can be leveraged to mine and learn patterns in antibody sequences, offering insights into antibody engineering, optimization, and discovery.

Published in mAbs

ISSN: 1942-0862 (Print); 1942-0870 (Online)
Publisher: Taylor & Francis Group
Country of publisher: United Kingdom
LCC subjects: Medicine: Therapeutics. Pharmacology; Medicine: Internal medicine: Specialties of internal medicine: Immunologic diseases. Allergy
Website: https://www.tandfonline.com/journals/kmab

About the journal

Abstract

Keywords