Hanhan Cong

dblp:84/5161 · DBLP profile ↗
← Back
7ranked-venue papers
1as first author
5since 2021 · last 2023
—ORCID · none

Domains — the database's venue-derived domains; a paper can count in several

Applied, interdisciplinary, general and emerging computing · 7 · 1 first-author · 5 since 2021
YearPublicationVenuePosition
2023 Protein-protein interaction site prediction by model ensembling with hybrid feature and self-attention
abstract
BACKGROUND: Protein-protein interactions (PPIs) are crucial in various biological functions and cellular processes. Thus, many computational approaches have been proposed to predict PPI sites. Although significant progress has been made, these methods still have limitations in encoding the characteristics of each amino acid in sequences. Many feature extraction methods rely on the sliding window technique, which simply merges all the features of residues into a vector. The importance of some key residues may be weakened in the feature vector, leading to poor performance. RESULTS: We propose a novel sequence-based method for PPI sites prediction. The new network model, PPINet, contains multiple feature processing paths. For a residue, the PPINet extracts the features of the targeted residue and its context separately. These two types of features are processed by two paths in the network and combined to form a protein representation, where the two types of features are of relatively equal importance. The model ensembling technique is applied to make use of more features. The base models are trained with different features and then ensembled via stacking. In addition, a data balancing strategy is presented, by which our model can get significant improvement on highly unbalanced data. CONCLUSION: The proposed method is evaluated on a fused dataset constructed from Dset186, Dset_72, and PDBset_164, as well as the public Dset_448 dataset. Compared with current state-of-the-art methods, the performance of our method is better than the others. In the most important metrics, such as AUPRC and recall, it surpasses the second-best programmer on the latter dataset by 6.9% and 4.7%, respectively. We also demonstrated that the improvement is essentially due to using the ensemble model, especially, the hybrid feature. We share our code for reproducibility and future research at https://github.com/CandiceCong/StackingPPINet .
Hanhan Cong, Hong Liu 0013, Cheng Liang 0001, Yuehui Chen
BMC Bioinform.1
2022 i6mA-word2vec: A Newly Model Which Used Distributed Features for Predicting DNA N6-Methyladenine Sites in Genomes
Wenzhen Fu, Yixin Zhong, Jiazi Chen, Hanhan Cong
ICIC (2)6
2022 Classification of S-succinylation Sites of Cysteine by Neural Network
Tong Meng, Yuehui Chen, Jiazi Chen, Hanhan Cong
ICIC (2)6
2022 SeqVec-GAT: A Golgi Classification Model Based on Multi-headed Graph Attention Network
Jianan Sui, Yuehui Chen, Jiazi Chen, Hanhan Cong
ICIC (2)6
2022 Predicting Protein-DNA Binding Sites by Fine-Tuning BERT
Yuehui Chen, Jiazi Chen, Hanhan Cong
ICIC (2)6
2016 Prediction of Phosphorylation Sites Using PSO-ANNs
Ruizhi Han, Dong Wang 0021, Yuehui Chen, Wenzheng Bao, Hanhan Cong
ICIC (1)6
2016 Feature Combination Methods for Prediction of Subcellular Locations of Proteins with Both Single and Multiple Sites
Dong Wang 0021, Yuehui Chen, Shanping Qiao, Yaou Zhao, Hanhan Cong
ICIC (1)6