Cargando…
Properties of Protein Drug Target Classes
Accurate identification of drug targets is a crucial part of any drug development program. We mined the human proteome to discover properties of proteins that may be important in determining their suitability for pharmaceutical modulation. Data was gathered concerning each protein’s sequence, post-t...
Autores principales: | , |
---|---|
Formato: | Online Artículo Texto |
Lenguaje: | English |
Publicado: |
Public Library of Science
2015
|
Materias: | |
Acceso en línea: | https://www.ncbi.nlm.nih.gov/pmc/articles/PMC4379170/ https://www.ncbi.nlm.nih.gov/pubmed/25822509 http://dx.doi.org/10.1371/journal.pone.0117955 |
_version_ | 1782364156476784640 |
---|---|
author | Bull, Simon C. Doig, Andrew J. |
author_facet | Bull, Simon C. Doig, Andrew J. |
author_sort | Bull, Simon C. |
collection | PubMed |
description | Accurate identification of drug targets is a crucial part of any drug development program. We mined the human proteome to discover properties of proteins that may be important in determining their suitability for pharmaceutical modulation. Data was gathered concerning each protein’s sequence, post-translational modifications, secondary structure, germline variants, expression profile and drug target status. The data was then analysed to determine features for which the target and non-target proteins had significantly different values. This analysis was repeated for subsets of the proteome consisting of all G-protein coupled receptors, ion channels, kinases and proteases, as well as proteins that are implicated in cancer. Machine learning was used to quantify the proteins in each dataset in terms of their potential to serve as a drug target. This was accomplished by first inducing a random forest that could distinguish between its targets and non-targets, and then using the random forest to quantify the drug target likeness of the non-targets. The properties that can best differentiate targets from non-targets were primarily those that are directly related to a protein’s sequence (e.g. secondary structure). Germline variants, expression levels and interactions between proteins had minimal discriminative power. Overall, the best indicators of drug target likeness were found to be the proteins’ hydrophobicities, in vivo half-lives, propensity for being membrane bound and the fraction of non-polar amino acids in their sequences. In terms of predicting potential targets, datasets of proteases, ion channels and cancer proteins were able to induce random forests that were highly capable of distinguishing between targets and non-targets. The non-target proteins predicted to be targets by these random forests comprise the set of the most suitable potential future drug targets, and should therefore be prioritised when building a drug development programme. |
format | Online Article Text |
id | pubmed-4379170 |
institution | National Center for Biotechnology Information |
language | English |
publishDate | 2015 |
publisher | Public Library of Science |
record_format | MEDLINE/PubMed |
spelling | pubmed-43791702015-04-09 Properties of Protein Drug Target Classes Bull, Simon C. Doig, Andrew J. PLoS One Research Article Accurate identification of drug targets is a crucial part of any drug development program. We mined the human proteome to discover properties of proteins that may be important in determining their suitability for pharmaceutical modulation. Data was gathered concerning each protein’s sequence, post-translational modifications, secondary structure, germline variants, expression profile and drug target status. The data was then analysed to determine features for which the target and non-target proteins had significantly different values. This analysis was repeated for subsets of the proteome consisting of all G-protein coupled receptors, ion channels, kinases and proteases, as well as proteins that are implicated in cancer. Machine learning was used to quantify the proteins in each dataset in terms of their potential to serve as a drug target. This was accomplished by first inducing a random forest that could distinguish between its targets and non-targets, and then using the random forest to quantify the drug target likeness of the non-targets. The properties that can best differentiate targets from non-targets were primarily those that are directly related to a protein’s sequence (e.g. secondary structure). Germline variants, expression levels and interactions between proteins had minimal discriminative power. Overall, the best indicators of drug target likeness were found to be the proteins’ hydrophobicities, in vivo half-lives, propensity for being membrane bound and the fraction of non-polar amino acids in their sequences. In terms of predicting potential targets, datasets of proteases, ion channels and cancer proteins were able to induce random forests that were highly capable of distinguishing between targets and non-targets. The non-target proteins predicted to be targets by these random forests comprise the set of the most suitable potential future drug targets, and should therefore be prioritised when building a drug development programme. Public Library of Science 2015-03-30 /pmc/articles/PMC4379170/ /pubmed/25822509 http://dx.doi.org/10.1371/journal.pone.0117955 Text en © 2015 Bull, Doig http://creativecommons.org/licenses/by/4.0/ This is an open-access article distributed under the terms of the Creative Commons Attribution License, which permits unrestricted use, distribution, and reproduction in any medium, provided the original author and source are properly credited. |
spellingShingle | Research Article Bull, Simon C. Doig, Andrew J. Properties of Protein Drug Target Classes |
title | Properties of Protein Drug Target Classes |
title_full | Properties of Protein Drug Target Classes |
title_fullStr | Properties of Protein Drug Target Classes |
title_full_unstemmed | Properties of Protein Drug Target Classes |
title_short | Properties of Protein Drug Target Classes |
title_sort | properties of protein drug target classes |
topic | Research Article |
url | https://www.ncbi.nlm.nih.gov/pmc/articles/PMC4379170/ https://www.ncbi.nlm.nih.gov/pubmed/25822509 http://dx.doi.org/10.1371/journal.pone.0117955 |
work_keys_str_mv | AT bullsimonc propertiesofproteindrugtargetclasses AT doigandrewj propertiesofproteindrugtargetclasses |