EDBT 2026 Demo / reviewers in the wild / expert
Zhisheng Wei
dblp:286/1824
· DBLP profile ↗
6ranked-venue papers
0as first author
6since 2021 · last 2026
0000-0002-9141-9418ORCID · corroborated
Domains — the database's venue-derived domains; a paper can count in several
Applied, interdisciplinary, general and emerging computing · 6 · 6 since 2021
| Year | Publication | Venue | Position |
|---|---|---|---|
| 2026 | SMENET: A Multi-View Semantic Model for Multi-Level Enzyme Function PredictionabstractComprehending biological reproduction and cellular metabolism is facilitated by the Enzyme Commission, which matches protein sequences to the biochemical reactions they catalyse through EC numbers. In recent years, several methods have been proposed for predicting enzyme function. However, these methods still encounter challenges. Firstly, traditional methods for manually designing enzyme features are complex and cumbersome, lacking an effective generalized method for embedding enzyme sequences. Secondly, the distribution gap between different enzymes is significant, which resulting in existing methods struggling to predict multilevel enzyme functions. Thirdly, traditional enzyme function prediction models only extract single view feature of enzyme, so there is still room for further improving the ability of these models to extract enzyme data. To address these challenges, a new multilevel enzyme function prediction model (SMENET) based on multi-view semantics is proposed. This method uses protein large language model to extract semantic information. Subsequently, this semantic information is fed into multiple information extraction network modules, followed by using Biologic Sematic Attention to integrate these views' information. Finally, a multi-view adaptive fusion network is designed to extract the best common representation between multiple semantic views. Extensive experiments were conducted on multiple datasets to validate the effectiveness of SMENET. Hanwen Zhou, Wei Zhang 0221, Zhaohong Deng, Guanjin Wang, Zhisheng Wei, Xiaoyong Pan, Hong-Bin Shen, Dongjun Yu, Jing Wu 0030 |
IEEE Trans. Comput. Biol. Bioinform. | 5 |
| 2025 | CATransUnetLBP: Accurate Prediction of Protein-Ligand Binding Pockets Using a Hybrid NetworkabstractThe development of intelligent methods capable of predicting protein-ligand binding sites has become a popular research field. Recently, deep learning based methods have been proposed as a promising solution for this task. However, some limitations still exist. For example, the network structure is not optimized for predicting protein binding pockets, which limits the model's capabilities. To address the aforementioned challenges, a novel method called CATransUnetLPB is proposed, in which a new network structure named CATransUnet is designed. The proposed CATransUnet combines CNN and Transformer models to accurately segment binding pocket regions from protein 3D structures. It outperforms existing representative methods on three test sets, demonstrating the effectiveness of optimizing the deep network model for detecting protein ligand binding pockets. Furthermore, we conduct thorough analysis on applying data augmentation to protein data structure and confirm that such technique can enhance the model's generalization ability, thereby ensuring good performance on new protein structures. Moreover, experiments show that the predicted binding pockets from our model can complement the results obtained from other methods. This suggests that integrating our method with existing approaches could further improve the prediction of protein-ligand binding pockets. Cheng Cai, Zhaohong Deng, Andong Li, Yun Zuo 0001, Haoran Chen 0003, Zhisheng Wei, Xiaoyong Pan, Hong-Bin Shen, Dongjun Yu |
IEEE Trans. Comput. Biol. Bioinform. | 6 |
| 2025 | DMMAFS: Protein Function Prediction Based on Multi-Modal Multi-Attention Fusion FeaturesabstractIntelligent prediction of protein function is more efficient and less resource-consuming and has achieved significant progress in recent years. However, most of the current methods are performed solely based on the sequence information of proteins. These methods overlook information of other modalities that the proteins themselves possess, which makes it difficult to achieve the desired predicted results. Furthermore, a few existing methods based on multiple modal information fuse them in a simple splicing manner and fail to fully exploit the complementary relation between different modalities. To address the above-mentioned challenges, we propose Multi-modal Multi-attention fusion Features (DMMAFS), a method based on deep learning, to predict protein function. On the one hand, DMMAFS gains the semantic information embedded in the sequence itself through the self-attention learning of the sequence. On the other hand, DMMAFS employs the 3D structural information of proteins to compensate for the sequence information. Particularly, a S-C cross-modal cross-attention fusion network module is proposed that not only optimizes the weights of the semantic information but also efficiently fuses the sequence features with the structural information, thus avoiding the simple splicing of different modal features. Our experimental results demonstrate that the proposed DMMAFS outperforms the state-of-the-art methods in protein function prediction. Liangwen He, Zhaohong Deng, Fuping Hu, Yun Zuo 0001, Haoran Chen 0003, Xiaoyong Pan, Zhisheng Wei, Hong-Bin Shen, Dongjun Yu, Jing Wu 0030 |
IEEE Trans. Comput. Biol. Bioinform. | 9 |
| 2025 | SEFP: Structure-Based Enzyme Function PredictionabstractTraditional biological experimental methods to determine enzyme properties are time-consuming and costly, leading to an increasing interest in computational models for enzyme function prediction. However, the existing computational methods are insufficient and inefficient to exploit enzyme structure. In this work, we introduce SEFP, a novel method leveraging enzyme point clouds for enzyme function prediction. The structure encoder of SEFP uses a tailored enzyme point cloud network to analyze the three-dimensional arrangement of atoms within the enzyme, integrating hierarchical residue global features through a residue feature adapter to extract detailed enzyme point features. Additionally, the Bio-BCS residue feature encoder extracts enzyme residue features with channel and spatial weights using a specially designed attention mechanism. Finally, SEFP fuses point and residue features to generate the final prediction results. Comparative evaluations show that SEFP outperforms various recent computational methods, demonstrating superior performance. On the RSCB enzyme structure dataset, SEFP achieves an f1-score of 95.85, outperforming two representative structure-based methods, EnzyNet and DeepFri. On the HECNet dataset, SEFP maintains its superiority over all comparison sequence-based methods, yielding an f1-score of 94.29. Ablation studies are conducted to confirm the effectiveness of individual modules within SEFP. These findings underscore the potential of SEFP for reliable and precise enzyme function prediction, offering advancements in bioinformatics and computational biology. Guanqing Yu, Zhaohong Deng, Chenxi Luo, Cheng Cai, Wei Zhang 0221, Fuping Hu, Kup-Sze Choi, Zhisheng Wei, Jing Wu 0030 |
IEEE Trans. Comput. Biol. Bioinform. | 9 |
| 2023 | MMSMAPlus: a multi-view multi-scale multi-attention embedding model for protein function predictionabstractProtein is the most important component in organisms and plays an indispensable role in life activities. In recent years, a large number of intelligent methods have been proposed to predict protein function. These methods obtain different types of protein information, including sequence, structure and interaction network. Among them, protein sequences have gained significant attention where methods are investigated to extract the information from different views of features. However, how to fully exploit the views for effective protein sequence analysis remains a challenge. In this regard, we propose a multi-view, multi-scale and multi-attention deep neural model (MMSMA) for protein function prediction. First, MMSMA extracts multi-view features from protein sequences, including one-hot encoding features, evolutionary information features, deep semantic features and overlapping property features based on physiochemistry. Second, a specific multi-scale multi-attention deep network model (MSMA) is built for each view to realize the deep feature learning and preliminary classification. In MSMA, both multi-scale local patterns and long-range dependence from protein sequences can be captured. Third, a multi-view adaptive decision mechanism is developed to make a comprehensive decision based on the classification results of all the views. To further improve the prediction performance, an extended version of MMSMA, MMSMAPlus, is proposed to integrate homology-based protein prediction under the framework of multi-view deep neural model. Experimental results show that the MMSMAPlus has promising performance and is significantly superior to the state-of-the-art methods. The source code can be found at https://github.com/wzy-2020/MMSMAPlus. Zhaohong Deng, Wei Zhang 0221, Qiongdan Lou, Kup-Sze Choi, Zhisheng Wei, Jing Wu 0030 |
Briefings Bioinform. | 6 |
| 2022 | circRNA-binding protein site prediction based on multi-view deep learning, subspace learning and multi-view classifierabstractCircular RNAs (circRNAs) generally bind to RNA-binding proteins (RBPs) to play an important role in the regulation of autoimmune diseases. Thus, it is crucial to study the binding sites of RBPs on circRNAs. Although many methods, including traditional machine learning and deep learning, have been developed to predict the interactions between RNAs and RBPs, and most of them are focused on linear RNAs. At present, few studies have been done on the binding relationships between circRNAs and RBPs. Thus, in-depth research is urgently needed. In the existing circRNA-RBP binding site prediction methods, circRNA sequences are the main research subjects, but the relevant characteristics of circRNAs have not been fully exploited, such as the structure and composition information of circRNA sequences. Some methods have extracted different views to construct recognition models, but how to efficiently use the multi-view data to construct recognition models is still not well studied. Considering the above problems, this paper proposes a multi-view classification method called DMSK based on multi-view deep learning, subspace learning and multi-view classifier for the identification of circRNA-RBP interaction sites. In the DMSK method, first, we converted circRNA sequences into pseudo-amino acid sequences and pseudo-dipeptide components for extracting high-dimensional sequence features and component features of circRNAs, respectively. Then, the structure prediction method RNAfold was used to predict the secondary structure of the RNA sequences, and the sequence embedding model was used to extract the context-dependent features. Next, we fed the above four views' raw features to a hybrid network, which is composed of a convolutional neural network and a long short-term memory network, to obtain the deep features of circRNAs. Furthermore, we used view-weighted generalized canonical correlation analysis to extract four views' common features by subspace learning. Finally, the learned subspace common features and multi-view deep features were fed to train the downstream multi-view TSK fuzzy system to construct a fuzzy rule and fuzzy inference-based multi-view classifier. The trained classifier was used to predict the specific positions of the RBP binding sites on the circRNAs. The experiments show that the prediction performance of the proposed method DMSK has been improved compared with the existing methods. The code and dataset of this study are available at https://github.com/Rebecca3150/DMSK. Zhaohong Deng, Xiaoyong Pan, Zhisheng Wei, Hong-Bin Shen, Kup-Sze Choi, Shitong Wang 0001, Jing Wu 0030 |
Briefings Bioinform. | 5 |