On QSAR-based cardiotoxicity modeling with the expressiveness-enhanced graph learning model and dual-threshold scheme-Reference-Cited by-同舟云学术

On QSAR-based cardiotoxicity modeling with the expressiveness-enhanced graph learning model and dual-threshold scheme

Published:2023-05-09 Issue: Volume:14 Page:
ISSN:1664-042X
Container-title:Frontiers in Physiology
language:
Short-container-title:Front. Physiol.

Author:

Wang Huijia,Zhu Guangxian,Izu Leighton T.,Chen-Izu Ye,Ono Naoaki,Altaf-Ul-Amin MD,Kanaya Shigehiko,Huang Ming

Abstract

Introduction: Given the direct association with malignant ventricular arrhythmias, cardiotoxicity is a major concern in drug design. In the past decades, computational models based on the quantitative structure–activity relationship have been proposed to screen out cardiotoxic compounds and have shown promising results. The combination of molecular fingerprint and the machine learning model shows stable performance for a wide spectrum of problems; however, not long after the advent of the graph neural network (GNN) deep learning model and its variant (e.g., graph transformer), it has become the principal way of quantitative structure–activity relationship-based modeling for its high flexibility in feature extraction and decision rule generation. Despite all these progresses, the expressiveness (the ability of a program to identify non-isomorphic graph structures) of the GNN model is bounded by the WL isomorphism test, and a suitable thresholding scheme that relates directly to the sensitivity and credibility of a model is still an open question.Methods: In this research, we further improved the expressiveness of the GNN model by introducing the substructure-aware bias by the graph subgraph transformer network model. Moreover, to propose the most appropriate thresholding scheme, a comprehensive comparison of the thresholding schemes was conducted.Results: Based on these improvements, the best model attains performance with 90.4% precision, 90.4% recall, and 90.5% F1-score with a dual-threshold scheme (active: <1μM; non-active: >30μM). The improved pipeline (graph subgraph transformer network model and thresholding scheme) also shows its advantages in terms of the activity cliff problem and model interpretability.

Funder

Japan Society for the Promotion of Science

Publisher

Frontiers Media SA

Subject

Physiology (medical),Physiology

Reference56 articles.

1. Pharmacogenomics and acquired long qt syndrome;Aerssens;Future Med.,2005

2. Cardiotoxicity of anticancer drugs: The need for cardio-oncology and cardio-oncological prevention;Albini;J. Natl. Cancer Inst.,2010

3. The properties of known drugs. 1. molecular frameworks;Bemis;J. Med. Chem.,1996

Cited by 2 articles. 订阅此论文施引文献订阅此论文施引文献，注册后可以免费订阅5篇论文的施引文献，订阅后可以查看论文全部施引文献

1. Exploring the artificial intelligence and machine learning models in the context of drug design difficulties and future potential for the pharmaceutical sectors;Methods;2023-11

2. Using the Correlation Intensity Index to Build a Model of Cardiotoxicity of Piperidine Derivatives;Molecules;2023-09-12