Building 2D classification models and 3D CoMSIA models on small-molecule inhibitors of both wild-type and T790M/L858R double-mutant EGFR

Epidermal growth factor receptor (EGFR) has received widespread attention because it is an important target for anticancer drug design. Mutations in the EGFR, especially the T790M/L858R double mutation, have made cancer treatment more difficult. We herein built the structure–activity relationship mo...

Ausführliche Beschreibung

Gespeichert in:
Bibliographische Detailangaben
Veröffentlicht in:Molecular diversity 2022-06, Vol.26 (3), p.1715-1730
Hauptverfasser: Huo, Donghui, Wang, Hongzhao, Qin, Zijian, Tian, Yujia, Yan, Aixia
Format: Artikel
Sprache:eng
Schlagworte:
Online-Zugang:Volltext
Tags: Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
Beschreibung
Zusammenfassung:Epidermal growth factor receptor (EGFR) has received widespread attention because it is an important target for anticancer drug design. Mutations in the EGFR, especially the T790M/L858R double mutation, have made cancer treatment more difficult. We herein built the structure–activity relationship models of small-molecule inhibitors on wild-type and T790M/L858R double-mutant EGFR with a whole dataset of 379 compounds. For 2D classification models, we used ECFP4 fingerprints to build support vector machine and random forest models and used SMILES to build self-attention recurrent neural network models. Each of all six models resulted in an accuracy of above 0.87 and the Matthews correlation coefficient value of above 0.76 on the test set, respectively. We concluded that inhibitors containing anilinoquinoline and methoxy or fluoro phenyl are highly active against wild EGFR. Substructures such as anilinopyrimidine, acrylamide, amino phenyl, methoxy phenyl, and thienopyrimidinyl amide appeared more in highly active inhibitors against double-mutant EGFR. We also used self-organizing map to cluster the inhibitors into six subsets based on ECFP4 fingerprints and analyzed the activity characteristics of different scaffolds in each subset. Among them, three datasets, which are based on pteridin, anilinopyrimidine, and anilinoquinoline scaffold, were selected to build 3D comparative molecular similarity analysis models individually. Models with the leave-one-out coefficient of determination (q 2 ) above 0.65 were selected, and five descriptor types (steric, electrostatic, hydrophobic, donor, and acceptor) were used to study the effects of side chains of inhibitors on the activity against wild-type and mutant-type EGFR. Graphic abstract
ISSN:1381-1991
1573-501X
DOI:10.1007/s11030-021-10300-9