Cargando…

Improved Deep Learning Based Method for Molecular Similarity Searching Using Stack of Deep Belief Networks

Virtual screening (VS) is a computational practice applied in drug discovery research. VS is popularly applied in a computer-based search for new lead molecules based on molecular similarity searching. In chemical databases similarity searching is used to identify molecules that have similarities to...

Descripción completa

Detalles Bibliográficos
Autores principales: Nasser, Maged, Salim, Naomie, Hamza, Hentabli, Saeed, Faisal, Rabiu, Idris
Formato: Online Artículo Texto
Lenguaje:English
Publicado: MDPI 2020
Materias:
Acceso en línea:https://www.ncbi.nlm.nih.gov/pmc/articles/PMC7795308/
https://www.ncbi.nlm.nih.gov/pubmed/33383976
http://dx.doi.org/10.3390/molecules26010128
_version_ 1783634414523121664
author Nasser, Maged
Salim, Naomie
Hamza, Hentabli
Saeed, Faisal
Rabiu, Idris
author_facet Nasser, Maged
Salim, Naomie
Hamza, Hentabli
Saeed, Faisal
Rabiu, Idris
author_sort Nasser, Maged
collection PubMed
description Virtual screening (VS) is a computational practice applied in drug discovery research. VS is popularly applied in a computer-based search for new lead molecules based on molecular similarity searching. In chemical databases similarity searching is used to identify molecules that have similarities to a user-defined reference structure and is evaluated by quantitative measures of intermolecular structural similarity. Among existing approaches, 2D fingerprints are widely used. The similarity of a reference structure and a database structure is measured by the computation of association coefficients. In most classical similarity approaches, it is assumed that the molecular features in both biological and non-biologically-related activity carry the same weight. However, based on the chemical structure, it has been found that some distinguishable features are more important than others. Hence, this difference should be taken consideration by placing more weight on each important fragment. The main aim of this research is to enhance the performance of similarity searching by using multiple descriptors. In this paper, a deep learning method known as deep belief networks (DBN) has been used to reweight the molecule features. Several descriptors have been used for the MDL Drug Data Report (MDDR) dataset each of which represents different important features. The proposed method has been implemented with each descriptor individually to select the important features based on a new weight, with a lower error rate, and merging together all new features from all descriptors to produce a new descriptor for similarity searching. Based on the extensive experiments conducted, the results show that the proposed method outperformed several existing benchmark similarity methods, including Bayesian inference networks (BIN), the Tanimoto similarity method (TAN), adapted similarity measure of text processing (ASMTP) and the quantum-based similarity method (SQB). The results of this proposed multi-descriptor-based on Stack of deep belief networks method (SDBN) demonstrated a higher accuracy compared to existing methods on structurally heterogeneous datasets.
format Online
Article
Text
id pubmed-7795308
institution National Center for Biotechnology Information
language English
publishDate 2020
publisher MDPI
record_format MEDLINE/PubMed
spelling pubmed-77953082021-01-10 Improved Deep Learning Based Method for Molecular Similarity Searching Using Stack of Deep Belief Networks Nasser, Maged Salim, Naomie Hamza, Hentabli Saeed, Faisal Rabiu, Idris Molecules Article Virtual screening (VS) is a computational practice applied in drug discovery research. VS is popularly applied in a computer-based search for new lead molecules based on molecular similarity searching. In chemical databases similarity searching is used to identify molecules that have similarities to a user-defined reference structure and is evaluated by quantitative measures of intermolecular structural similarity. Among existing approaches, 2D fingerprints are widely used. The similarity of a reference structure and a database structure is measured by the computation of association coefficients. In most classical similarity approaches, it is assumed that the molecular features in both biological and non-biologically-related activity carry the same weight. However, based on the chemical structure, it has been found that some distinguishable features are more important than others. Hence, this difference should be taken consideration by placing more weight on each important fragment. The main aim of this research is to enhance the performance of similarity searching by using multiple descriptors. In this paper, a deep learning method known as deep belief networks (DBN) has been used to reweight the molecule features. Several descriptors have been used for the MDL Drug Data Report (MDDR) dataset each of which represents different important features. The proposed method has been implemented with each descriptor individually to select the important features based on a new weight, with a lower error rate, and merging together all new features from all descriptors to produce a new descriptor for similarity searching. Based on the extensive experiments conducted, the results show that the proposed method outperformed several existing benchmark similarity methods, including Bayesian inference networks (BIN), the Tanimoto similarity method (TAN), adapted similarity measure of text processing (ASMTP) and the quantum-based similarity method (SQB). The results of this proposed multi-descriptor-based on Stack of deep belief networks method (SDBN) demonstrated a higher accuracy compared to existing methods on structurally heterogeneous datasets. MDPI 2020-12-29 /pmc/articles/PMC7795308/ /pubmed/33383976 http://dx.doi.org/10.3390/molecules26010128 Text en © 2020 by the authors. Licensee MDPI, Basel, Switzerland. This article is an open access article distributed under the terms and conditions of the Creative Commons Attribution (CC BY) license (http://creativecommons.org/licenses/by/4.0/).
spellingShingle Article
Nasser, Maged
Salim, Naomie
Hamza, Hentabli
Saeed, Faisal
Rabiu, Idris
Improved Deep Learning Based Method for Molecular Similarity Searching Using Stack of Deep Belief Networks
title Improved Deep Learning Based Method for Molecular Similarity Searching Using Stack of Deep Belief Networks
title_full Improved Deep Learning Based Method for Molecular Similarity Searching Using Stack of Deep Belief Networks
title_fullStr Improved Deep Learning Based Method for Molecular Similarity Searching Using Stack of Deep Belief Networks
title_full_unstemmed Improved Deep Learning Based Method for Molecular Similarity Searching Using Stack of Deep Belief Networks
title_short Improved Deep Learning Based Method for Molecular Similarity Searching Using Stack of Deep Belief Networks
title_sort improved deep learning based method for molecular similarity searching using stack of deep belief networks
topic Article
url https://www.ncbi.nlm.nih.gov/pmc/articles/PMC7795308/
https://www.ncbi.nlm.nih.gov/pubmed/33383976
http://dx.doi.org/10.3390/molecules26010128
work_keys_str_mv AT nassermaged improveddeeplearningbasedmethodformolecularsimilaritysearchingusingstackofdeepbeliefnetworks
AT salimnaomie improveddeeplearningbasedmethodformolecularsimilaritysearchingusingstackofdeepbeliefnetworks
AT hamzahentabli improveddeeplearningbasedmethodformolecularsimilaritysearchingusingstackofdeepbeliefnetworks
AT saeedfaisal improveddeeplearningbasedmethodformolecularsimilaritysearchingusingstackofdeepbeliefnetworks
AT rabiuidris improveddeeplearningbasedmethodformolecularsimilaritysearchingusingstackofdeepbeliefnetworks