
arXiv: 2205.15364
Abstract As a necessary process of modern drug development, finding a drug compound that can selectively bind to a specific protein is highly challenging and costly. Exploring drug‐target interaction strength in terms of drug‐target affinity (DTA) is an emerging and effective research approach for drug development. However, it is challenging to model drug‐target interactions in a deep learning manner, and few studies provide interpretable analysis of models. This paper proposes a DTA prediction method (mutual transformer‐drug target affinity [MT‐DTA]) with interactive learning and an autoencoder mechanism. The proposed MT‐DTA builds a variational autoencoders system with a cascade structure of the attention model and convolutional neural networks. It not only enhances the ability to capture the characteristic information of a single molecular sequence but also establishes the characteristic expression relationship for each substructure in a single molecular sequence. On this basis, a molecular information interaction module is constructed, which adds information interaction paths between molecular sequence pairs and complements the expression of correlations between molecular substructures. The performance of the proposed model was verified on two public benchmark datasets, KIBA and Davis, and the results confirm that the proposed model structure is effective in predicting DTA. Additionally, attention transformer models with different configurations can improve the feature expression of drug/protein molecules. The model performs better in correctly predicting interaction strengths compared with state‐of‐the‐art baselines. In addition, the diversity of drug/protein molecules can be better expressed than existing methods such as SeqGAN and Co‐VAE to generate more effective new drugs. The DTA value prediction module fuses the drug‐target pair interaction information to output the predicted value of DTA. Additionally, this paper theoretically proves that the proposed method maximises evidence lower bound for the joint distribution of the DTA prediction model, which enhances the consistency of the probability distribution between actual and predicted values. The source code of proposed method is available at https://github.com/Lamouryz/Code/tree/main/MT‐DTA .
FOS: Computer and information sciences, Computer Science - Machine Learning, medical applications, deep learning, Biomolecules (q-bio.BM), Machine Learning (cs.LG), QA76.75-76.765, Quantitative Biology - Biomolecules, FOS: Biological sciences, Computational linguistics. Natural language processing, Computer software, P98-98.5
FOS: Computer and information sciences, Computer Science - Machine Learning, medical applications, deep learning, Biomolecules (q-bio.BM), Machine Learning (cs.LG), QA76.75-76.765, Quantitative Biology - Biomolecules, FOS: Biological sciences, Computational linguistics. Natural language processing, Computer software, P98-98.5
| selected citations These citations are derived from selected sources. This is an alternative to the "Influence" indicator, which also reflects the overall/total impact of an article in the research community at large, based on the underlying citation network (diachronically). | 63 | |
| popularity This indicator reflects the "current" impact/attention (the "hype") of an article in the research community at large, based on the underlying citation network. | Top 1% | |
| influence This indicator reflects the overall/total impact of an article in the research community at large, based on the underlying citation network (diachronically). | Top 10% | |
| impulse This indicator reflects the initial momentum of an article directly after its publication, based on the underlying citation network. | Top 1% |
