Unknown

Dataset Information

0

Lipophilicity prediction of peptides and peptide derivatives by consensus machine learning.


ABSTRACT: Lipophilicity prediction is routinely applied to small molecules and presents a working alternative to experimental log?P or log?D determination. For compounds outside the domain of classical medicinal chemistry these predictions lack accuracy, advocating the development of bespoke in silico approaches. Peptides and their derivatives and mimetics fill the structural gap between small synthetic drugs and genetically engineered macromolecules. Here, we present a data-driven machine learning method for peptide log?D 7.4 prediction. A model for estimating the lipophilicity of short linear peptides consisting of natural amino acids was developed. In a prospective test, we obtained accurate predictions for a set of newly synthesized linear tri- to hexapeptides. Further model development focused on more complex peptide mimetics from the AstraZeneca compound collection. The results obtained demonstrate the applicability of the new prediction model to peptides and peptide derivatives in a log?D 7.4 range of approximately -3 to 5, with superior accuracy to established lipophilicity models for small molecules.

SUBMITTER: Fuchs JA 

PROVIDER: S-EPMC6151477 | biostudies-literature | 2018 Sep

REPOSITORIES: biostudies-literature

altmetric image

Publications

Lipophilicity prediction of peptides and peptide derivatives by consensus machine learning.

Fuchs Jens-Alexander JA   Grisoni Francesca F   Kossenjans Michael M   Hiss Jan A JA   Schneider Gisbert G  

MedChemComm 20180822 9


Lipophilicity prediction is routinely applied to small molecules and presents a working alternative to experimental log <i>P</i> or log <i>D</i> determination. For compounds outside the domain of classical medicinal chemistry these predictions lack accuracy, advocating the development of bespoke <i>in silico</i> approaches. Peptides and their derivatives and mimetics fill the structural gap between small synthetic drugs and genetically engineered macromolecules. Here, we present a data-driven ma  ...[more]

Similar Datasets

| S-EPMC3226394 | biostudies-literature
2013-01-01 | E-GEOD-29210 | biostudies-arrayexpress
| S-EPMC5652333 | biostudies-literature
| S-EPMC8597836 | biostudies-literature
| S-EPMC10635286 | biostudies-literature
2022-05-26 | MTBLS2841 | MetaboLights
| S-EPMC5885771 | biostudies-literature
| S-EPMC9850743 | biostudies-literature
2013-01-01 | GSE29210 | GEO
| S-EPMC9421197 | biostudies-literature