Unknown

Dataset Information

0

An Explainable Supervised Machine Learning Model for Predicting Respiratory Toxicity of Chemicals Using Optimal Molecular Descriptors.


ABSTRACT: Respiratory toxicity is a serious public health concern caused by the adverse effects of drugs or chemicals, so the pharmaceutical and chemical industries demand reliable and precise computational tools to assess the respiratory toxicity of compounds. The purpose of this study is to develop quantitative structure-activity relationship models for a large dataset of chemical compounds associated with respiratory system toxicity. First, several feature selection techniques are explored to find the optimal subset of molecular descriptors for efficient modeling. Then, eight different machine learning algorithms are utilized to construct respiratory toxicity prediction models. The support vector machine classifier outperforms all other optimized models in 10-fold cross-validation. Additionally, it outperforms the prior study by 2% in prediction accuracy and 4% in MCC. The best SVM model achieves a prediction accuracy of 86.2% and a MCC of 0.722 on the test set. The proposed SVM model predictions are explained using the SHapley Additive exPlanations approach, which prioritizes the relevance of key modeling descriptors influencing the prediction of respiratory toxicity. Thus, our proposed model would be incredibly beneficial in the early stages of drug development for predicting and understanding potential respiratory toxic compounds.

SUBMITTER: Jaganathan K 

PROVIDER: S-EPMC9028223 | biostudies-literature | 2022 Apr

REPOSITORIES: biostudies-literature

altmetric image

Publications

An Explainable Supervised Machine Learning Model for Predicting Respiratory Toxicity of Chemicals Using Optimal Molecular Descriptors.

Jaganathan Keerthana K   Tayara Hilal H   Chong Kil To KT  

Pharmaceutics 20220411 4


Respiratory toxicity is a serious public health concern caused by the adverse effects of drugs or chemicals, so the pharmaceutical and chemical industries demand reliable and precise computational tools to assess the respiratory toxicity of compounds. The purpose of this study is to develop quantitative structure-activity relationship models for a large dataset of chemical compounds associated with respiratory system toxicity. First, several feature selection techniques are explored to find the  ...[more]

Similar Datasets

| S-EPMC11251627 | biostudies-literature
| S-EPMC11005042 | biostudies-literature
| S-EPMC10042997 | biostudies-literature
| S-EPMC11534633 | biostudies-literature
| S-EPMC11522690 | biostudies-literature
| S-EPMC11391437 | biostudies-literature
| S-EPMC8151683 | biostudies-literature
| S-EPMC5472585 | biostudies-literature
| S-EPMC8317304 | biostudies-literature
| S-EPMC7645727 | biostudies-literature