{"title":"Deciphering Cathepsin K inhibitors: a combined QSAR, docking and MD simulation based machine learning approaches for drug design.","authors":"S Ilyas, J Lee, Y Hwang, Y Choi, D Lee","doi":"10.1080/1062936X.2024.2405626","DOIUrl":null,"url":null,"abstract":"<p><p>Cathepsin K (CatK), a lysosomal cysteine protease, contributes to skeletal abnormalities, heart diseases, lung inflammation, and central nervous system and immune disorders. Currently, CatK inhibitors are associated with severe adverse effects, therefore limiting their clinical utility. This study focuses on exploring quantitative structure-activity relationships (QSAR) on a dataset of CatK inhibitors (1804) compiled from the ChEMBL database to predict the inhibitory activities. After data cleaning and pre-processing, a total of 1568 structures were selected for exploratory data analysis which revealed physicochemical properties, distributions and statistical significance between the two groups of inhibitors. PubChem fingerprinting with 11 different machine-learning classification models was computed. The comparative analysis showed the ET model performed well with accuracy values for the training set (0.999), cross-validation (0.970) and test set (0.977) in line with OECD guidelines. Moreover, to gain structural insights on the origin of CatK inhibition, 15 diverse molecules were selected for molecular docking. The CatK inhibitors (1 and 2) exhibited strong binding energies of -8.3 and -7.2 kcal/mol, respectively. MD simulation (300 ns) showed strong structural stability, flexibility and interactions in selected complexes. This synergy between QSAR, docking, MD simulation and machine learning models strengthen our evidence for developing novel and resilient CatK inhibitors.</p>","PeriodicalId":21446,"journal":{"name":"SAR and QSAR in Environmental Research","volume":" ","pages":"771-793"},"PeriodicalIF":2.3000,"publicationDate":"2024-09-01","publicationTypes":"Journal Article","fieldsOfStudy":null,"isOpenAccess":false,"openAccessPdf":"","citationCount":"0","resultStr":null,"platform":"Semanticscholar","paperid":null,"PeriodicalName":"SAR and QSAR in Environmental Research","FirstCategoryId":"93","ListUrlMain":"https://doi.org/10.1080/1062936X.2024.2405626","RegionNum":3,"RegionCategory":"环境科学与生态学","ArticlePicture":[],"TitleCN":null,"AbstractTextCN":null,"PMCID":null,"EPubDate":"2024/10/9 0:00:00","PubModel":"Epub","JCR":"Q3","JCRName":"CHEMISTRY, MULTIDISCIPLINARY","Score":null,"Total":0}
引用次数: 0
Abstract
Cathepsin K (CatK), a lysosomal cysteine protease, contributes to skeletal abnormalities, heart diseases, lung inflammation, and central nervous system and immune disorders. Currently, CatK inhibitors are associated with severe adverse effects, therefore limiting their clinical utility. This study focuses on exploring quantitative structure-activity relationships (QSAR) on a dataset of CatK inhibitors (1804) compiled from the ChEMBL database to predict the inhibitory activities. After data cleaning and pre-processing, a total of 1568 structures were selected for exploratory data analysis which revealed physicochemical properties, distributions and statistical significance between the two groups of inhibitors. PubChem fingerprinting with 11 different machine-learning classification models was computed. The comparative analysis showed the ET model performed well with accuracy values for the training set (0.999), cross-validation (0.970) and test set (0.977) in line with OECD guidelines. Moreover, to gain structural insights on the origin of CatK inhibition, 15 diverse molecules were selected for molecular docking. The CatK inhibitors (1 and 2) exhibited strong binding energies of -8.3 and -7.2 kcal/mol, respectively. MD simulation (300 ns) showed strong structural stability, flexibility and interactions in selected complexes. This synergy between QSAR, docking, MD simulation and machine learning models strengthen our evidence for developing novel and resilient CatK inhibitors.
期刊介绍:
SAR and QSAR in Environmental Research is an international journal welcoming papers on the fundamental and practical aspects of the structure-activity and structure-property relationships in the fields of environmental science, agrochemistry, toxicology, pharmacology and applied chemistry. A unique aspect of the journal is the focus on emerging techniques for the building of SAR and QSAR models in these widely varying fields. The scope of the journal includes, but is not limited to, the topics of topological and physicochemical descriptors, mathematical, statistical and graphical methods for data analysis, computer methods and programs, original applications and comparative studies. In addition to primary scientific papers, the journal contains reviews of books and software and news of conferences. Special issues on topics of current and widespread interest to the SAR and QSAR community will be published from time to time.