Cargando…

Decoding the protein–ligand interactions using parallel graph neural networks

Protein–ligand interactions (PLIs) are essential for biochemical functionality and their identification is crucial for estimating biophysical properties for rational therapeutic design. Currently, experimental characterization of these properties is the most accurate method, however, this is very ti...

Descripción completa

Detalles Bibliográficos
Autores principales:	Knutson, Carter, Bontha, Mridula, Bilbrey, Jenna A., Kumar, Neeraj
Formato:	Online Artículo Texto
Lenguaje:	English
Publicado:	Nature Publishing Group UK 2022
Materias:	Article
Acceso en línea:	https://www.ncbi.nlm.nih.gov/pmc/articles/PMC9086424/ https://www.ncbi.nlm.nih.gov/pubmed/35538084 http://dx.doi.org/10.1038/s41598-022-10418-2

_version_	1784703997116940288
author	Knutson, Carter Bontha, Mridula Bilbrey, Jenna A. Kumar, Neeraj
author_facet	Knutson, Carter Bontha, Mridula Bilbrey, Jenna A. Kumar, Neeraj
author_sort	Knutson, Carter
collection	PubMed
description	Protein–ligand interactions (PLIs) are essential for biochemical functionality and their identification is crucial for estimating biophysical properties for rational therapeutic design. Currently, experimental characterization of these properties is the most accurate method, however, this is very time-consuming and labor-intensive. A number of computational methods have been developed in this context but most of the existing PLI prediction heavily depends on 2D protein sequence data. Here, we present a novel parallel graph neural network (GNN) to integrate knowledge representation and reasoning for PLI prediction to perform deep learning guided by expert knowledge and informed by 3D structural data. We develop two distinct GNN architectures: [Formula: see text] is the base implementation that employs distinct featurization to enhance domain-awareness, while [Formula: see text] is a novel implementation that can predict with no prior knowledge of the intermolecular interactions. The comprehensive evaluation demonstrated that GNN can successfully capture the binary interactions between ligand and protein’s 3D structure with 0.979 test accuracy for [Formula: see text] and 0.958 for [Formula: see text] for predicting activity of a protein–ligand complex. These models are further adapted for regression tasks to predict experimental binding affinities and [Formula: see text] crucial for compound’s potency and efficacy. We achieve a Pearson correlation coefficient of 0.66 and 0.65 on experimental affinity and 0.50 and 0.51 on [Formula: see text] with [Formula: see text] and [Formula: see text] , respectively, outperforming similar 2D sequence based models. Our method can serve as an interpretable and explainable artificial intelligence (AI) tool for predicted activity, potency, and biophysical properties of lead candidates. To this end, we show the utility of [Formula: see text] on SARS-Cov-2 protein targets by screening a large compound library and comparing the prediction with the experimentally measured data.
format	Online Article Text
id	pubmed-9086424
institution	National Center for Biotechnology Information
language	English
publishDate	2022
publisher	Nature Publishing Group UK
record_format	MEDLINE/PubMed
spelling	pubmed-90864242022-05-10 Decoding the protein–ligand interactions using parallel graph neural networks Knutson, Carter Bontha, Mridula Bilbrey, Jenna A. Kumar, Neeraj Sci Rep Article Protein–ligand interactions (PLIs) are essential for biochemical functionality and their identification is crucial for estimating biophysical properties for rational therapeutic design. Currently, experimental characterization of these properties is the most accurate method, however, this is very time-consuming and labor-intensive. A number of computational methods have been developed in this context but most of the existing PLI prediction heavily depends on 2D protein sequence data. Here, we present a novel parallel graph neural network (GNN) to integrate knowledge representation and reasoning for PLI prediction to perform deep learning guided by expert knowledge and informed by 3D structural data. We develop two distinct GNN architectures: [Formula: see text] is the base implementation that employs distinct featurization to enhance domain-awareness, while [Formula: see text] is a novel implementation that can predict with no prior knowledge of the intermolecular interactions. The comprehensive evaluation demonstrated that GNN can successfully capture the binary interactions between ligand and protein’s 3D structure with 0.979 test accuracy for [Formula: see text] and 0.958 for [Formula: see text] for predicting activity of a protein–ligand complex. These models are further adapted for regression tasks to predict experimental binding affinities and [Formula: see text] crucial for compound’s potency and efficacy. We achieve a Pearson correlation coefficient of 0.66 and 0.65 on experimental affinity and 0.50 and 0.51 on [Formula: see text] with [Formula: see text] and [Formula: see text] , respectively, outperforming similar 2D sequence based models. Our method can serve as an interpretable and explainable artificial intelligence (AI) tool for predicted activity, potency, and biophysical properties of lead candidates. To this end, we show the utility of [Formula: see text] on SARS-Cov-2 protein targets by screening a large compound library and comparing the prediction with the experimentally measured data. Nature Publishing Group UK 2022-05-10 /pmc/articles/PMC9086424/ /pubmed/35538084 http://dx.doi.org/10.1038/s41598-022-10418-2 Text en © The Author(s) 2022 https://creativecommons.org/licenses/by/4.0/Open AccessThis article is licensed under a Creative Commons Attribution 4.0 International License, which permits use, sharing, adaptation, distribution and reproduction in any medium or format, as long as you give appropriate credit to the original author(s) and the source, provide a link to the Creative Commons licence, and indicate if changes were made. The images or other third party material in this article are included in the article's Creative Commons licence, unless indicated otherwise in a credit line to the material. If material is not included in the article's Creative Commons licence and your intended use is not permitted by statutory regulation or exceeds the permitted use, you will need to obtain permission directly from the copyright holder. To view a copy of this licence, visit http://creativecommons.org/licenses/by/4.0/ (https://creativecommons.org/licenses/by/4.0/) .
spellingShingle	Article Knutson, Carter Bontha, Mridula Bilbrey, Jenna A. Kumar, Neeraj Decoding the protein–ligand interactions using parallel graph neural networks
title	Decoding the protein–ligand interactions using parallel graph neural networks
title_full	Decoding the protein–ligand interactions using parallel graph neural networks
title_fullStr	Decoding the protein–ligand interactions using parallel graph neural networks
title_full_unstemmed	Decoding the protein–ligand interactions using parallel graph neural networks
title_short	Decoding the protein–ligand interactions using parallel graph neural networks
title_sort	decoding the protein–ligand interactions using parallel graph neural networks
topic	Article
url	https://www.ncbi.nlm.nih.gov/pmc/articles/PMC9086424/ https://www.ncbi.nlm.nih.gov/pubmed/35538084 http://dx.doi.org/10.1038/s41598-022-10418-2
work_keys_str_mv	AT knutsoncarter decodingtheproteinligandinteractionsusingparallelgraphneuralnetworks AT bonthamridula decodingtheproteinligandinteractionsusingparallelgraphneuralnetworks AT bilbreyjennaa decodingtheproteinligandinteractionsusingparallelgraphneuralnetworks AT kumarneeraj decodingtheproteinligandinteractionsusingparallelgraphneuralnetworks

Decoding the protein–ligand interactions using parallel graph neural networks

Ejemplares similares