Yun Zhang 0019

dblp:02/6428-19 · DBLP profile ↗
← Back
22ranked-venue papers
7as first author
17since 2021 · last 2026
0000-0001-8716-4179ORCID · conflict

Domains — the database's venue-derived domains; a paper can count in several

Artificial intelligence and machine learning · 16 · 7 first-author · 12 since 2021Databases, data management, data science and information retrieval · 3 · 1 first-author · 1 since 2021Applied, interdisciplinary, general and emerging computing · 3 · 3 since 2021Graphics, computer vision, multimedia, augmented reality and games · 1 · 1 since 2021
YearPublicationVenuePosition
2026 ECG-GTMD: A CVD diagnosis model based on graph-transformer with multi-domain fusion of spatiotemporal and band features
Lihan Xia, Yun Zhang 0019, Yongguo Liu
Inf. Sci.2
2026 Adaptive latent disease state learning for multimodal Alzheimer's disease biomarker detection with missing modalities
Zhi Chen 0014, Fengli Zhang, Yun Zhang 0019, Jiajing Zhu, Qiaoqin Li, Yongguo Liu
Pattern Recognit.3
2025 FRGEM: Feature integration pre-training based Gaussian embedding model for Chinese word representation
Yun Zhang 0019, Yongguo Liu, Jiajing Zhu, Zhi Chen 0014, Fengli Zhang
Expert Syst. Appl.1
2025 Inner-character and Inner-word Features Based Representation Learning for Chinese Word Embedding
abstract
Chinese word embedding is a significant task in natural language processing (NLP). Most researchers explored Chinese word embedding according to radical, component, stroke n -gram and character features. Besides these features, Chinese characters still have structure and pinyin characteristics. In this article, we propose ensemble ssp2vec and connective ssp2vec to utilize inner-character features (stroke, structure, and pinyin) for learning Chinese word embeddings. Then we design hierarchical ssp2vec to forecast the contexts according to the combination of inner-character (stroke, structure, and pinyin) and inner-word features (character) of Chinese words to explore different feature combination ways for learning feature relevance and comprehending word semantics, where feature substring is proposed to learn the relevancy of stroke, structure, and pinyin. Experimental results for word analogy, word similarity, text classification, and named entity recognition tasks demonstrate that the proposed methods outperform most state-of-the-art models.
Yun Zhang 0019, Yongguo Liu, Jiajing Zhu, Zhi Chen 0014, Shuangqing Zhai, Xindong Wu 0001
ACM Trans. Asian Low Resour. Lang. Inf. Process.1
2025 Enhanced Multimodal Low-Rank Embedding-Based Feature Selection Model for Multimodal Alzheimer's Disease Diagnosis
abstract
Identification of Alzheimer's disease (AD) with multimodal neuroimaging data has been receiving increasing attention. However, the presence of numerous redundant features and corrupted neuroimages within multimodal datasets poses significant challenges for existing methods. In this paper, we propose a feature selection method named Enhanced Multimodal Low-rank Embedding (EMLE) for multimodal AD diagnosis. Unlike previous methods utilizing convex relaxations of the -norm, EMLE exploits an -norm regularized projection matrix to obtain an embedding representation and select informative features jointly for each modality. The -norm, employing an upper-bounded nonconvex Minimax Concave Penalty (MCP) function to characterize sparsity, offers a superior approximation for the -norm compared to other convex relaxations. Next, a similarity graph is learned based on the self-expressiveness property to increase the robustness to corrupted data. As the approximation coefficient vectors of samples from the same class should be highly correlated, an MCP function introduced norm, i.e., matrix -norm, is applied to constrain the rank of the graph. Furthermore, recognizing that diverse modalities should share an underlying structure related to AD, we establish a consensus graph for all modalities to unveil intrinsic structures across multiple modalities. Finally, we fuse the embedding representations of all modalities into the label space to incorporate supervisory information. The results of extensive experiments on the Alzheimer's Disease Neuroimaging Initiative datasets verify the discriminability of the features selected by EMLE.
Zhi Chen 0014, Yongguo Liu, Yun Zhang 0019, Jiajing Zhu, Qiaoqin Li, Xindong Wu 0001
IEEE Trans. Medical Imaging3
2024 HIFINet: Examination-Diagnosis-Treatment Hierarchical Feedback Interaction Network for Medication Recommendation
abstract
Abstract Medication combination recommendation is critical in clinic, since accurately predicting therapeutic drug can provide essential decision support to physicians. However, current approaches do not consider the multilevel structure of electronic health record (EHR) data or the hierarchical dependencies between multiple visits, leading to suboptimal recommendations. To address these limitations, we propose a novel hierarchical feedback interaction network (HIFINet) to utilize an examination-diagnosis-treatment hierarchical network for modeling the inherent multilevel structure of EHR data. The feedback long short-term memory network called FeLSTM, which is the basic unit of our hierarchical network, performs hierarchical interactions and leverages change information as feedback to propagate forward among different levels. Additionally, HIFINet contains four modules. First, an embedding module is designed to learn the health information representation of patients. Second, a three-layer time-series learning module is employed to capture temporal dependencies within each sequence. Next, a differential feedback interaction module is developed to capture the difference features between visits. Finally, an attention fusion module is used to learn a comprehensive representation of the patient’s health information and to recommend next multiple treatment medications. HIFINet is compared with state-of-the-art approaches on a real-world dataset. The results indicate that HIFINet outperforms other approaches, offering more accurate recommendations.
Hengjie Zheng, Yongguo Liu, Shangming Yang, Yun Zhang 0019, Jiajing Zhu, Zhi Chen 0014
Neural Process. Lett.4
2024 Shared Manifold Regularized Joint Feature Selection for Joint Classification and Regression in Alzheimer's Disease Diagnosis
abstract
In Alzheimer’s disease (AD) diagnosis, joint feature selection for predicting disease labels (classification) and estimating cognitive scores (regression) with neuroimaging data has received increasing attention. In this paper, we propose a model named Shared Manifold regularized Joint Feature Selection (SMJFS) that performs classification and regression in a unified framework for AD diagnosis. For classification, unlike the existing works that build least squares regression models which are insufficient in the ability of extracting discriminative information for classification, we design an objective function that integrates linear discriminant analysis and subspace sparsity regularization for acquiring an informative feature subset. Furthermore, the local data relationships are learned according to the samples’ transformed distances to exploit the local data structure adaptively. For regression, in contrast to previous works that overlook the correlations among cognitive scores, we learn a latent score space to capture the correlations and employ the latent space to design a regression model with ℓ2,1-norm regularization, facilitating the feature selection in regression task. Moreover, the missing cognitive scores can be recovered in the latent space for increasing the number of available training samples. Meanwhile, to capture the correlations between the two tasks and describe the local relationships between samples, we construct an adaptive shared graph to guide the subspace learning in classification and the latent cognitive score learning in regression simultaneously. An efficient iterative optimization algorithm is proposed to solve the optimization problem. Extensive experiments on three datasets validate the discriminability of the features selected by SMJFS.
Zhi Chen 0014, Yongguo Liu, Yun Zhang 0019, Jiajing Zhu, Qiaoqin Li, Xindong Wu 0001
IEEE Trans. Image Process.3
2023 A Weakly Supervised Deep Learning Model for Alzheimer's Disease Prognosis Using MRI and Incomplete Labels
Zhi Chen 0014, Yongguo Liu, Yun Zhang 0019, Jiajing Zhu, Qiaoqin Li
ICONIP (3)3
2023 DAEM: Deep attributed embedding based multi-task learning for predicting adverse drug-drug interaction
Jiajing Zhu, Yongguo Liu, Yun Zhang 0019, Zhi Chen 0014, Kun She 0001, Rongsheng Tong
Expert Syst. Appl.3
2023 Orthogonal latent space learning with feature weighting and graph learning for multimodal Alzheimer's disease diagnosis
Zhi Chen 0014, Yongguo Liu, Yun Zhang 0019, Qiaoqin Li
Medical Image Anal.3
2023 A Word-Concept Heterogeneous Graph Convolutional Network for Short Text Classification
Shigang Yang, Yongguo Liu, Yun Zhang 0019, Jiajing Zhu
Neural Process. Lett.3
2022 Exploring Chinese word embedding with similar context and reinforcement learning
Yun Zhang 0019, Yongguo Liu, Shuangqing Zhai
Neural Comput. Appl.1
2022 Multi-Attribute Discriminative Representation Learning for Prediction of Adverse Drug-Drug Interaction
abstract
Adverse drug-drug interaction (ADDI) is a significant life-threatening issue, posing a leading cause of hospitalizations and deaths in healthcare systems. This paper proposes a unified Multi-Attribute Discriminative Representation Learning (MADRL) model for ADDI prediction. Unlike the existing works that equally treat features of each attribute without discrimination and do not consider the underlying relationship among drugs, we first develop a regularized optimization problem based on CUR matrix decomposition for joint representative drug and discriminative feature selection such that the selected drugs and features can well approximate the original feature spaces and the critical factors discriminative to ADDIs can be properly explored. Different from the existing models that ignore the consistent and unique properties among attributes, a Generative Adversarial Network (GAN) framework is then designed to capture the inter-attribute shared and intra-attribute specific representations of adverse drug pairs for exploiting their consensus and complementary information in ADDI prediction. Meanwhile, MADRL is compatible with any kind of attributes and capable of exploring their respective effects on ADDI prediction. An iterative algorithm based on the alternating direction method of multipliers is developed for optimization. Experiments on publicly available dataset demonstrate the effectiveness of MADRL when compared with eleven baselines and its six variants.
Jiajing Zhu, Yongguo Liu, Yun Zhang 0019, Zhi Chen 0014, Xindong Wu 0001
IEEE Trans. Pattern Anal. Mach. Intell.3
2022 Adaptive Regularized Multiattribute Fuzzy Distance Learning for Predicting Adverse Drug-Drug Interaction
abstract
Adverse drug–drug interaction (ADDI) causes harmful injuries and accidental deaths in patients, posing as a significant life-threatening issue in public health. Early prediction of ADDIs has become an increasingly concerning task for the safety of pharmacotherapy during clinical treatments. In this article, we propose an adaptive regularized multiattribute fuzzy distance (MAFD) learning model for ADDI prediction. Unlike the existing works that only focus on whether an adverse interaction occurs or not for a specific drug pair and do not consider their implicit medication risks, MAFD employs fuzzy distance learning by designing a fuzzy membership matrix to model the adverse distance with a fuzziness level for exploring the medication risks of adverse drug pairs. Meanwhile, for each attribute, we develop two projection matrices to respectively map its original feature and adverse interaction spaces into a common space for eliminating noisy information and capturing their compact and informative representations. Besides, adaptive regularization is explicitly designed to investigate the underlying characteristics of different attributes in ADDI modeling and neighborhood structure preservation is seamlessly integrated to benefit the prediction results. The optimization problem is solved by an iterative algorithm based on the alternating direction method of multipliers with detailed convergence proofs. Experiments on real-world dataset demonstrate the effectiveness of MAFD when compared with ten baselines and its five variants.
Jiajing Zhu, Yongguo Liu, Yun Zhang 0019, Zhi Chen 0014, Xindong Wu 0001
IEEE Trans. Fuzzy Syst.3
2021 Time-frequency deep metric learning for multivariate time series classification
Zhi Chen 0014, Yongguo Liu, Jiajing Zhu, Yun Zhang 0019, Rongjiang Jin, Xia He, Lidian Chen
Neurocomputing4
2021 FSPRM: A Feature Subsequence Based Probability Representation Model for Chinese Word Embedding
abstract
Chinese word embedding models capture Chinese semantics based on the character feature of Chinese words and the internal features of Chinese characters such as radical, component, stroke, structure and pinyin. However, some features are overlapping and most methods do not consider their relevance. Meanwhile, they express words as point vectors that cannot better capture different aspect semantics of Chinese words. In this paper, we propose a Feature Subsequence based Probability Representation Model (FSPRM) for learning Chinese word embeddings, in which we first integrate the morphological and phonetic features (stroke, structure and pinyin) of Chinese characters and learn their relevance by designing a feature subsequence to capture relatively comprehensive semantics of Chinese words, then feature probability distribution is proposed for capturing different aspect meanings of Chinese words based on the three internal features and probability representation by estimating its mean as the sum of feature subsequences. Chinese words with similar features may have similar semantics, then we map Chinese words to feature probability distributions and design a similarity-based objective for predicting the contextual words of the target word to learn their semantics. Extensive experiments on word analogy, word similarity, text classification and named entity recognition tasks demonstrate that the proposed method outperforms most state-of-the-art approaches.
Yun Zhang 0019, Yongguo Liu, Jiajing Zhu, Xindong Wu 0001
IEEE ACM Trans. Audio Speech Lang. Process.1
2021 Attribute Supervised Probabilistic Dependent Matrix Tri-Factorization Model for the Prediction of Adverse Drug-Drug Interaction
abstract
Adverse drug-drug interaction (ADDI) becomes a significant threat to public health. Despite the detection of ADDIs is experimentally implemented in the early development phase of drug design, many potential ADDIs are still clinically explored by accidents, leading to a large number of morbidity and mortality. Several computational models are designed for ADDI prediction. However, they take no consideration of drug dependency, although many drugs usually produce synergistic effects and own highly mutual dependency in treatments, which contains underlying information about ADDIs and benefits ADDI prediction. In this paper, we design a dependent network to model the drug dependency and propose an attribute supervised learning model Probabilistic Dependent Matrix Tri-Factorization (PDMTF) for ADDI prediction. In particular, PDMTF incorporates two drug attributes, molecular structure and side effect, and their correlation to model the adverse interactions among drugs. The dependent network is represented by a dependent matrix, which is first formulated by the row precision matrix of the predicted attribute matrices and then regularized by the molecular structure similarities among drugs. Meanwhile, an efficient alternating algorithm is designed for solving the optimization problem of PDMTF. Experiments demonstrate the superior performance of the proposed model when compared with eight baselines and its two variants.
Jiajing Zhu, Yongguo Liu, Yun Zhang 0019
IEEE J. Biomed. Health Informatics3
2020 LILPA: A label importance based label propagation algorithm for community detection with application to core drug discovery
Yun Zhang 0019, Yongguo Liu, Qiaoqin Li, Rongjiang Jin, Chuanbiao Wen
Neurocomputing1
2020 A no self-edge stochastic block model and a heuristic algorithm for balanced anti-community detection in networks
Jiajing Zhu, Yongguo Liu, Zhi Chen 0014, Yun Zhang 0019, Shangming Yang, Changhong Yang, Wen Yang 0007, Xindong Wu 0001
Inf. Sci.5
2020 GLLPA: A Graph Layout based Label Propagation Algorithm for community detection
Yun Zhang 0019, Yongguo Liu, Rongjiang Jin, Lidian Chen, Xindong Wu 0001
Knowl. Based Syst.1
2019 Learning Chinese Word Embeddings from Stroke, Structure and Pinyin of Characters
abstract
Chinese word embeddings have recently attracted much attention in natural language processing (NLP). Existing researches learn Chinese word embeddings based on characters, radicals, components and stroke n-gram. Besides abovementioned features, Chinese characters also own structure and pinyin features. In this paper, we design feature substring, a super set of radicals, components and stroke n-gram with structure and pinyin information, to integrate stroke, structure and pinyin features of Chinese characters and capture the semantics of Chinese words. Based on the feature substring, we propose a novel method ssp2vec to predict the contextual words based on the feature substrings of the target words for learning Chinese word embeddings. It is based on our observation that exploiting the morphological information (stroke and structure) and the phonetic information (pinyin) is crucial for capturing the meanings of Chinese words. Meanwhile, the phonetic information (pinyin) can assist the model to distinguish Chinese words. Experimental results on word analogy, word similarity, text classification and named entity recognition tasks show that the proposed method obtains better results than state-of-the-art approaches.
Yun Zhang 0019, Yongguo Liu, Jiajing Zhu, Ziqiang Zheng, Shuangqing Zhai
CIKM1
2019 IHPreten: A novel supervised learning framework with attribute regularization for prediction of incompatible herb pair in traditional Chinese medicine
Jiajing Zhu, Yongguo Liu, Yun Zhang 0019, Zhi Chen 0014, Qiaoqin Li, Shangming Yang, Shuangqing Zhai, Chuanbiao Wen
Neurocomputing3