EDBT 2026 Demo / reviewers in the wild / expert
Xiaoyi Lv
dblp:253/2860
· DBLP profile ↗
47ranked-venue papers
0as first author
46since 2021 · last 2026
—ORCID · conflict
Domains — the database's venue-derived domains; a paper can count in several
Artificial intelligence and machine learning · 27 · 26 since 2021Graphics, computer vision, multimedia, augmented reality and games · 9 · 9 since 2021Applied, interdisciplinary, general and emerging computing · 6 · 6 since 2021Databases, data management, data science and information retrieval · 4 · 4 since 2021Systems, architecture and hardware · 1 · 1 since 2021Human-computer interaction and ubiquitous computing · 1 · 1 since 2021
| Year | Publication | Venue | Position |
|---|---|---|---|
| 2026 | Rigorous Cross-Cohort Evaluation and Compatibility Screening for Microbiome-Based Colorectal Cancer Prediction
Xiaona Zhu, Xuecong Tian, Haiqing Sun, Xiaoyi Lv |
ICIC (27) | 5 |
| 2026 | MARD-Mol: a hybrid autoregressive-diffusion paradigm for coarse-grained molecular modelingabstractMOTIVATION: Deep generative models have transformed drug molecule generation. However, molecules exhibit complex hierarchical structures, requiring models to simultaneously balance macroscopic topological coherence and microscopic chemical self-consistency. Although autoregressive (AR) and discrete diffusion paradigms are highly complementary, integrating their advantages within a unified architecture remains severely limited by traditional "atom-by-atom" fine-grained modeling. RESULTS: We propose MARD-Mol, a hybrid AR-diffusion framework based on motif-inspired units. By elevating the representation granularity from atoms to motif-inspired units and introducing a dual-stream hierarchical attention mechanism, it couples inter-unit AR global scaffold planning with intra-unit discrete diffusion generation. To support goal-directed drug discovery, we reformulate property optimization into an iterative "diagnose-and-repair" process, enabling targeted optimization of defective motifs while preserving the global scaffold. Extensive experiments demonstrate that MARD-Mol achieves an 86.0% Quality score in de novo generation and exhibits superior performance in fragment-constrained and multi-objective optimization, establishing a new paradigm for high-quality drug design. AVAILABILITY AND IMPLEMENTATION: The source code and datasets used in this study are available at GitHub: https://github.com/CSUBioGroup/MARD-Mol. Sizhe Zhang, Xiaoyi Lv |
Bioinform. | 4 |
| 2026 | A trinity-branch parallel fusion and supervised enhancement network: A multimodal celiac disease diagnosis network based on transformer and dual-tower supervision
Xuguang Zhou, Xiaoyi Lv |
Eng. Appl. Artif. Intell. | 6 |
| 2026 | The research on the diagnostic technology for aortic dissection and acute myocardial infarction based on Raman and infrared spectroscopy combined with multimodal deep learning
Guangyao Ma, Xuguang Zhou, Ting Tian, Xiaoyi Lv, Yining Yang |
Eng. Appl. Artif. Intell. | 9 |
| 2026 | Mixture-of-experts-based hierarchical dynamic multimodal fusion network for dermatological diagnosis
Min Li 0093, Enguang Zuo, Xiaoyi Lv, Shumei Bao, Chengwei Rao, Chen Chen 0078 |
Neurocomputing | 5 |
| 2026 | FSC-MAE: Feature structure coordinated mask autoencoder
Enguang Zuo, Chen Chen 0078, Xiaoyi Lv, Ruishuang Sun, Yinhong Li, Hongbing Ma |
Neurocomputing | 5 |
| 2026 | GHTMDA: A self-supervised heterogeneous graph hierarchical contrastive learning model for efficient metabolite-disease associations prediction
Binglu Hu, Xuecong Tian, Xiaoyi Lv |
Inf. Process. Manag. | 5 |
| 2026 | A Multi-instance Learning Network with Prototype-instance Adversarial Contrastive for Cervix Pathology Grading
Furong Luo, Binlin Ma, Shuxian Liu, Xiaoyi Lv, Pan Huang 0001 |
Medical Image Anal. | 5 |
| 2026 | PatchFusionMLP: A scalable multi-resolution MLP framework for time series prediction
Xinyu Bi, Xiaoyi Lv, Junyu Zhu, Hongbing Ma, Enguang Zuo |
Pattern Recognit. | 4 |
| 2026 | Infrared and visible image fusion model based on source image interaction
Kuizhuang Liu, Chengwei Rao, Xiaoyi Lv |
Pattern Recognit. | 6 |
| 2026 | A single-cell feature selection method based on subspace and minimum redundancy, applicable to multi-omics
Xiaoyi Lv, Yuanzhi He, Jin Gu |
Pattern Recognit. | 3 |
| 2025 | MBGNet: Mamba-Based Boundary-Guided Multimodal Medical Image Segmentation Network
Ke Xu 0005, Guangjian Liu, Enguang Zuo, Xiaoyi Lv |
CVM (1) | 7 |
| 2025 | SCGRL: Graph representation learning based on edge structure contrastive self-supervised frameworkabstractIn recent years, significant advancements have been made in contrastive self-supervised learning for graphs. However, most of the existing methods start from the feature level and ignore the structural information. In this work, we propose a graph representation learning based on edge structure contrastive self-supervised framework (SCGRL), which leverages a novel edge-structure-based "masked edges vs. complementary edges" instance pairs to fully utilize the topological information of the graph, and attempts to reconstruct the original graph using the visible graph structure. In the feature processing, normal coded features are constrained with the coded features without gradient updating to enhance the encoder’s prediction ability for the masked representation. In addition, boundary losses are designed to ensure that the model can accurately distinguish between different instance pairs. We conduct extensive experiments on various benchmark datasets to demonstrate that SCGRL outperforms the state-of-the-art in different downstream tasks, especially link prediction. Ruishuang Sun, Ruiting Wang, Enguang Zuo, Junyu Zhu, Chen Chen 0078, Xiaoyi Lv |
ICME | 7 |
| 2025 | Frequency-Spatial Domain Fusion for Graph Anomaly DetectionabstractGraph anomaly detection (GAD) remains a major challenge in artificial intelligence security applications. Existing graph neural networks (GNN) face two key issues: (1) the neighborhood smoothing in spatial domain methods can mask the distribution differences between normal and abnormal graphs, filtering out important high-frequency signals; (2) their excessive reliance on local structural patterns fails to capture global frequency responses, limiting the effectiveness of anomaly detection. To address these issues, we propose FSGAD (Frequency-Spatial Graph Anomaly Detection), an unsupervised framework that combines spectral analysis and contrastive learning. Our main innovations include spectral fingerprint extraction, which extracts node-level spectral features through graph Fourier transforms, capturing global distribution patterns to accurately detect abnormal nodes; and neighborhood contrastive learning and dual-channel feature reconstruction, which use spatial and frequency domain information for precise anomaly pattern detection. Our experimental results on eight benchmark datasets show that FSGAD surpasses existing methods with an AUC improvement of 2.9% and a time consumption reduction of 21%. Our code and data are available at https://github.com/Senbao/FSGAD. Senbao Hou, Enguang Zuo, Ruiting Wang, Xiaoyi Lv |
IJCNN | 5 |
| 2025 | Clinical Experience-inspired Multimodal Fusion Networks for Dermatological ClassificationabstractMultimodal fusion algorithms that integrate clinical images and metadata are important for dermatological classification. However, existing fusion algorithms are overly focused on model optimization at the technical level, neglecting the role of prior knowledge in the medical domain to guide the model. Therefore, this paper proposes a clinical experience-inspired multimodal fusion network (CEMF-Net), which embeds clinical experience into the multimodal fusion model and effectively improves the interpretability and clinical applicability of the model. Specifically, first, based on inspiration from a medical perspective, we have designed a clinical image amplification (CIA) block, through which the operation of doctors’ amplification observation is effectively simulated using bicubic interpolation. Meanwhile, considering the interdependence between the local detail analysis of the lesion area and the global localization demand, we design the metadata-guided lesion localization (MLL) block for precise localization operation. Secondly, an adaptive multi-scale fusion strategy is proposed at the decision-making level, which dynamically integrates the discriminative information by autonomously learning the prediction weight parameters of the features at different scales, to better simulate the clinician’s comprehensive trade-offs of the multi-scale visual features in the diagnosis process. Experimental results on the PAD-UFES-20 and Derm7pt datasets show that CEMF-Net outperforms existing representative dermatological classification algorithms in evaluation metrics such as accuracy, precision, and balanced accuracy. The experimental results demonstrate the effectiveness of embedding clinical perspectives into dermatological classification algorithms and provide new research ideas in the field of dermatological multimodal fusion. Enguang Zuo, Guangjian Liu, Xiaoyi Lv |
IJCNN | 8 |
| 2025 | Beyond Local Features: A Metadata-Driven Image Frequency Modulation Network for Skin Disease ClassificationabstractMetadata (e.g., age, genetic history, etc.) and clinical images provide a multidimensional perspective of patients with skin diseases. Effective integration of the complementary information from these two modalities is critical. However, existing fusion techniques often combine convolutional neural networks and transformers to learn local and global features. The conventional matrix multiplication operation in transformers not only fails to focus on the frequency components of different modalities, but also leads to high spatial and temporal complexity. Therefore, this study introduces Fourier transform to capture the global information of images through point product operations in the frequency domain, thus achieving efficient integration of multi-frequency features of images driven by metadata. Specifically, we designed a Metadata-Driven Image Frequency Modulation Network (MDFM Net). The network consists of multiple cascaded frequency domain multidimensional collaborative layers (FDMC layers), which can refine the image frequency domain features step by step. The FDMC layer is composed of parallel Channel-Wise Frequency Domain Fusion Block (CWF Block) and Patch-Wise Frequency Domain Fusion block (PWF Block), which are driven by metadata and dynamically adjust the multi-frequency features of the image from different dimensions. Extensive experiments were conducted on the PAD-UFES-20 and Derm7pt datasets, and our method achieved an accuracy of 84.9% and 80.2%, respectively, improving the SOTA methods by 2.9% and 2.6%. The code will be released at: https://github.com/wwy8/MDFM. Enguang Zuo, Xiaoyi Lv |
IJCNN | 5 |
| 2025 | Zero-shot Stance Detection with Sentiment Signals and Contrastive LearningabstractZero-shot stance detection aims to identify users’ stance without labeled data. This paradigm alleviates the data dependency problem of in-target stance detection and has attracted extensive research. In this article, we propose a novel zero-shot stance detection model consisting of three parts: topic mapping module, sentiment signal extraction module, and contrastive learning module. We introduce sentiment signals to improve the stance detection model performance and use contrastive learning to enhance the learning representation quality. We conducted experiments on the VAST dataset to validate the effectiveness of our proposed method. Guangzhen Liu, Xuehua Bi, Xiaoyi Lv |
IJCNN | 5 |
| 2025 | FreTime:Dual-Branch Frequency-Time Representation Learning for Time SeriesabstractTime series analysis plays a fundamental role in revealing data evolution, trends, and cyclical patterns. However, existing studies often fail to effectively address the dynamic dependencies between variables in multidimensional time series and the temporal evolution patterns within variables, thereby limiting the effectiveness of complex time series feature analysis. In this paper, we propose a dual-branch frequency-time interactive representation learning model (FreqTime) that captures the correlations between variables and the temporal dependencies within variables through a collaborative architecture in the time domain and frequency domain. The time domain branch uses an inverse Transformer architecture to model cross-variable interactions, while the frequency domain branch utilizes multi-scale gated convolutions to capture features and map them back to the time domain. Finally, global representations are obtained by interactively fusing the representations learned from the two branches in the time domain. Experiments demonstrate that FreqTime achieves state-of-the-art performance on long sequence prediction, classification, and anomaly detection tasks, and exhibits strong robustness in noisy environments. Junyu Zhu, Enguang Zuo, Ruishuang Sun, Ziwei Yan, Chen Chen 0078, Xiaoyi Lv |
SMC | 8 |
| 2025 | TDMFS: Tucker decomposition multimodal fusion model for pan-cancer survival prediction
Jinchao Chen, Enguang Zuo, Ziwei Yan, Xinya Chen, Xiaoyi Lv |
Artif. Intell. Medicine | 11 |
| 2025 | Disentangled global and local features of multi-source data variational autoencoder: An interpretable model for diagnosing IgAN via multi-source Raman spectral fusion techniques
Wei Shuai, Xuecong Tian, Enguang Zuo, Jin Gu, Chen Chen 0078, Xiaoyi Lv |
Artif. Intell. Medicine | 8 |
| 2025 | Predicting disease associations based on the higher order structure of ceRNA networksabstractCompetitive endogenous RNA (ceRNA) network regulation is an important posttranscriptional regulatory mechanism that plays an important role in physiological and pathological processes, and has been widely used in biomarker screening and regulatory factor studies of disease-related genes. However, existing studies have mainly focused on the association of a single type of RNA with disease, while studies targeting the application of ceRNA networks in disease prediction are still limited, so it is crucial to explore the potential of ceRNA networks in disease prediction. In this study, we propose CERDA-HOSR, a computational method for mining ceRNA network-disease associations based on higher order graph attention networks. The method uses higher order graph convolutional networks to aggregate neighborhood information to generate representations of different RNAs and diseases. Given the higher order complexity of biological networks and sample imbalance problem, traditional random negative sampling is difficult to effectively capture global information; for this reason, a higher order negative sampling strategy is designed to optimize the quality of negative samples by combining the network structure and higher order neighborhood relations to improve the generalization ability and prediction accuracy of the model. Finally, LightGBM calculates the ceRNA network-disease association probability based on the learned embedding. A large number of simulation experiments validate the superiority of CERDA-HOSR, and its practical application is further demonstrated by case studies of cardiovascular disease, acute myeloid leukemia, and papillary thyroid cancer. In addition, ablation experiments and exploratory analyses further enhance its robustness and provide an effective tool for disease prediction and biomarker screening. Zhaoliang Chai, Xuecong Tian, Xiaoyi Lv |
Briefings Bioinform. | 5 |
| 2025 | A single-cell RNA sequencing data imputation method based on non-negative matrix factorization and multi-kernel similarity network fusion
Jin Gu, Xinya Chen, Xiaoyi Lv |
Eng. Appl. Artif. Intell. | 8 |
| 2025 | High-order graph convolutional networks for circular Ribonucleic Acid and disease association prediction incorporating multiple biological relationships
Xiaoyi Lv, Jin Gu, Enguang Zuo, Chenjie Chang |
Eng. Appl. Artif. Intell. | 3 |
| 2025 | WIGNN: An adaptive graph-structured reasoning model for credit default predictionabstractIn credit default prediction, the main challenge is handling complex data structures and addressing data class imbalance . Given class imbalance and multi-dimensional data, general models find it difficult to fully explore the deep interdependencies within the data and the interaction effects between local and global. To overcome these challenges, this study proposes a Weighted Imbalanced Graph Neural Network (WIGNN) model that integrates adaptive graph structure inference with differential weight connectivity strategy, and the model solves the existing problems from the perspective of differential weight connectivity and graph balancing. Here, the weight connection uses the Gaussian kernel function to refine calculations and an adaptive percentile method to adjust sparsity , improving the understanding and efficiency of mining data connections. The weighted graph generated by this method can reflect the interaction between nodes and improve the model’s ability to analyse complex data structures. Based on this weighted graph, the graph imbalance module adopts a reinforcement learning-driven neighbour sampling strategy to adjust the sampling threshold automatically, optimizes the node embedding through message aggregation, and combines with a cost-sensitive matrix to improve classification accuracy and cost-effectiveness of the model on diverse credit datasets. We applied the WIGNN model to six real and class-imbalanced credit datasets, comparing it with 11 mainstream credit default prediction models. Evaluated using metrics Area Under the Curve (AUC), Geometric Mean (G-mean), and Accuracy. The results show that WIGNN significantly outperforms other models in handling class imbalance and graph sparsity , demonstrating its potential in financial credit applications. Zhipeng Yan, Hanwen Qu, Chen Chen 0078, Xiaoyi Lv, Enguang Zuo, Xulun Cai |
Eng. Appl. Artif. Intell. | 4 |
| 2025 | The MLSE-SCAM architecture combines with the improved DRSN-TIC model for Raman spectroscopy small-sample data learning
Enguang Zuo, Zhongcheng Gong, Xiaoyi Lv |
Expert Syst. Appl. | 10 |
| 2025 | RLCFE-Net: A reparameterization large convolutional kernel feature extraction network for weed detection in multiple scenarios
Zhenhong Jia, Baoquan Ge, Sensen Song, Congbing He, Jiajia Wang 0006, Xiaoyi Lv |
Expert Syst. Appl. | 9 |
| 2025 | TreeXformer: Extracting tabular feature-context information using tree-structured semantics
Yinhong Li, Hanwen Qu, Chen Chen 0078, Xiaoyi Lv, Enguang Zuo, Xulun Cai |
Inf. Process. Manag. | 4 |
| 2025 | DCFusion: Difference correlation-driven fusion mechanism of infrared and visible images
Min Li 0093, Enguang Zuo, Chaoxun Guo, Yunling Wang, Xiaoyi Lv, Chen Chen 0078 |
Pattern Recognit. | 8 |
| 2025 | Efficient time series adaptive representation learning via Dynamic Routing Sparse Attention
Enguang Zuo, Chen Chen 0078, Ziwei Yan, Xiaoyi Lv |
Pattern Recognit. | 7 |
| 2025 | Address Anomalies at Critical Crossroads for Graph Anomaly DetectionabstractGraph anomaly detection (GAD) on attributed networks aims to capture abnormal nodes whose attributes or structures differ significantly from most nodes. The existing GAD models amplify the representation differences between normal and abnormal nodes to identify anomalies via carefully designed feature extraction modules. However, these models ignore the bottlenecks encountered by abnormal nodes in message passing. In particular, when the anomalies occurs at critical crossroads, the information of multiple nodes is compressed into a fixed-length representation, and the resulting over-squashing weakens the abnormal information. To address this, we propose an unsupervisedSTructural optimization model guided by sIMilarity reconstruction (STIM). Specifically, we define redundant edges that cause over-squashing, design the Neighbor-Structure Optimization module to filter redundant edges through the edge-dropping strategy based on critical crossroads, and optimize the graph structure to alleviate over-squashing. In addition, to alleviate the over-smoothing caused by the high inter-class node similarity of the data itself and the edge-dropping strategy, we design the Neighbor-Similarity Reconstruction module based on similarity calculation, which guides the model to expand inter-class variation. Extensive experiments on benchmark datasets show that STIM can effectively optimize message passing and improve anomaly detection performance. The source code is available athttps://github.com/Junyi-Yan/STIM. Junyi Yan, Enguang Zuo, Ke Liang 0006, Meng Liu 0014, Miaomiao Li 0001, Xinwang Liu 0002, Xiaoyi Lv, Kai Lu 0001 |
IEEE Trans. Knowl. Data Eng. | 7 |
| 2024 | MDKFusion: Medical Domain Knowledge-Inspired Area Amplification Network for Multi-Sequence MRI Image Fusion in Ischemic StrokeabstractMulti-sequence MRI image fusion technology aids radiologists in quickly and accurately assessing ischemic lesions and their surrounding areas by combining DWI and FLAIR images to generate information-rich fusion images. Despite the rapid development of medical image fusion techniques, existing methods are predominantly focused on technical-level model optimization and fail to effectively integrate medical domain knowledge. This limitation reduces their clinical applicability and model interpretability. Inspired by radiologists' diagnostic pattern, which involves focusing on and enlarging lesion areas, we propose a medical domain knowledge-inspired area amplification network for multi-sequence MRI image fusion in ischemic stroke, named MDKFusion. Specifically, we design the Lesion Area Amplification (LAA) module, which uses bicubic interpolation for adaptive amplification and incorporates crosslevel and neighboring-level feature mapping with high-level feature co-guidance. This design emulates radiologists' practice of zooming in to examine lesions, thereby enhancing interpretability. Additionally, we employ the Feature Guidance Module (FGM) to achieve progressive guidance and feature integration. We further introduce the ℒSCDloss function to minimize pixel discrepancies between source and fused images, improving fusion quality. Compared to various mainstream fusion methods, MDKFusion achieves state-of-the-art (SOTA) performance across eight objective evaluation metrics. To confirm its practical value in clinical diagnosis, we invited five radiologists to perform a subjective evaluation of the fused images. Our code will be available at https://github.com/MinLila/MDKFusion. Min Li 0093, Pahati Tuxunjiang, Enguang Zuo, Xiaoyi Lv, Yunling Wang, Chen Chen 0078 |
BIBM | 5 |
| 2024 | SMAE: A Split Masked Graph AutoencoderabstractAutoencoders, as a generative self-supervised learning, have received more and more attention in recent years in image, video, and other media-related information processing. However, Graph AutoEncoder (GAE) has yet to achieve the capability demonstrated by contrastive learning in the task-centered on attribute networks. The main limitation lies in the fact that traditional autoencoder architectures require pretext tasks that align with downstream tasks, resulting in limited expressive power of the encoder. In this paper, we propose a novel separable-task generative self-supervised learning framework capable of providing high-quality representations, Split Masked AutoEncoder (SMAE), which unleashes the encoder’s ability to extract representations through an intelligent design. Our approach focuses on unlocking the potential of the encoder by introducing encoding transfer and feature replacement strategies, thereby enabling self-supervised pretext tasks to achieve atomic separation and fully unleash the encoder’s feature representation potential. We conducted extensive experiments on widely-used graph classification datasets, and the results demonstrate that SMAE outperforms state-of-the-art baselines in terms of graph classification accuracy and generation quality. Furthermore, our experimental findings show that prediction at the representation layer is more effective than original graph layer reconstruction in the field of masked graph autoencoders. Ruiting Wang, Enguang Zuo, Chen Chen 0078, Junyi Yan, Ziwei Yan, Xiaoyi Lv |
ICME | 8 |
| 2024 | A Survey of Zero-Shot Stance Detection
Guangzhen Liu, Xuehua Bi, Xiaoyi Lv |
NLPCC (5) | 5 |
| 2024 | Rethinking the Necessity of Learnable Modal Alignment for Medical Image Fusion
Min Li 0093, Enguang Zuo, Xiaoyi Lv, Chen Chen 0078 |
PRCV (5) | 4 |
| 2024 | DRA-CN: A Novel Dual-Resolution Attention Capsule Network for Histopathology Image Classification
Palidan Tursun, Xiaoyi Lv, Chen Chen 0006, Yunling Wang |
PRCV (14) | 4 |
| 2024 | A prospective study: Advances in chaotic characteristics of serum Raman spectroscopy in the field of assisted diagnosis of disease
Chen Chen 0078, Xuecong Tian, Enguang Zuo, Chenjie Chang, Min Li 0093, Xiaoyi Lv |
Expert Syst. Appl. | 10 |
| 2024 | CMACF: Transformer-based cross-modal attention cross-fusion model for systemic lupus erythematosus diagnosis combining Raman spectroscopy, FTIR spectroscopy, and metabolomics
Xuguang Zhou, Chen Chen 0078, Xiaoyi Lv, Enguang Zuo, Min Li 0093 |
Inf. Process. Manag. | 3 |
| 2024 | Self-contrastive Feature Guidance Based Multidimensional Collaborative Network of metadata and image features for skin disease classification
Min Li 0093, Enguang Zuo, Chen Chen 0078, Xiaoyi Lv |
Pattern Recognit. | 6 |
| 2024 | DSFusion: Infrared and visible image fusion method combining detail and scene information
Kuizhuang Liu, Min Li 0093, Chengwei Rao, Enguang Zuo, Yunling Wang, Ziwei Yan, Chen Chen 0078, Xiaoyi Lv |
Pattern Recognit. | 10 |
| 2024 | ASFFuse: Infrared and visible image fusion model based on adaptive selection feature maps
Kuizhuang Liu, Min Li 0093, Enguang Zuo, Chen Chen 0078, Yunling Wang, Xiaoyi Lv |
Pattern Recognit. | 8 |
| 2023 | Rethinking graph anomaly detection: A self-supervised Group Discrimination paradigm with Structure-AwareabstractStructural anomalies are the core problem in graph anomaly detection. However, the current mainstream self-supervised graph anomaly detection models do not directly model structural anomalies and their expensive time consumption limits the efficiency of graph anomaly detection. For this reason, we rethink graph anomaly detection and propose a self-supervised Group Discrimination paradigm with Structure-Aware (GDSA). Our model can be explicitly aware of the graph topology changes by multi-view structure disturbance. Moreover, GDSA transforms graph anomaly detection into discriminating the scalar summaries of positive and negative group nodes. The results of extensive experiments on four benchmark datasets show that GDSA outperforms current state-of-the-art methods, with the most significant AUC performance improvement of 28.7%. Notably, in scalability testing on a large-scale dataset, the training time and testing time of GDSA are 1181.0× and 5064.7× faster than the baseline, respectively, with 61.9% savings in memory usage. Junyi Yan, Enguang Zuo, Chen Chen 0078, Tianle Li, Xiaoyi Lv |
ICME | 7 |
| 2023 | A Masked Attention Network with Query Sparsity Measurement for Time Series Anomaly DetectionabstractTime series aomaly detection has been widely studied in recent years. Previous research focuses on point-wise features and pairwise associations for feature learning or designed anomaly scores based on prior knowledge. However, these methods cannot fully learn the intricate abnormal dynamic information and can only identify a limited class of anomalies. We propose a Masked Attention Network with Query Sparsity Measurement (MAN-QSM) to address the above challenges. This model uses two kinds of prior knowledge to fully exploit the differences between normal and abnormal points from two perspectives: pairwise association and sequence-level information. We designs the anomaly mask mechanism to collaborate with the training strategy to amplify the difference between normal and abnormal points. In experiments, we compare the model with classical methods, reconstruction-based models, autoregressive-based models, and state-of-the-art models, and the MAN-QSM achieves state-of-the-art results on SMD, PSM, and MSL datasets with an average of 16% reduction in error rate. Enguang Zuo, Chen Chen 0078, Junyi Yan, Tianle Li, Xiaoyi Lv |
ICME | 7 |
| 2023 | Accelerating BWA-MEM Read Mapping on GPUsabstractAdvancements in Next-Generation Sequencing (NGS) have significantly reduced the cost of generating DNA sequence data and increased the speed of data production. However, such high-throughput data production has increased the need for efficient data analysis programs. One of the most computationally demanding steps in analyzing sequencing data is mapping short reads produced by NGS to a reference DNA sequence, such as a human genome. The mapping program BWA-MEM and its newer version BWA-MEM2, optimized for CPUs, are some of the most popular choices for this task. In this study, we discuss the implementation of BWA-MEM on GPUs. This is a challenging task because many algorithms and data structures in BWA-MEM do not execute efficiently on the GPU architecture. This paper identifies major challenges in developing efficient GPU code on all major stages of the BWA-MEM program, including seeding, seed chaining, Smith-Waterman alignment, memory management, and I/O handling. We conduct comparison experiments against BWA-MEM and BWA-MEM2 running on a 64-thread CPU. The results show that our implementation achieved up to 3.2x speedup over BWA-MEM2 and up to 5.8x over BWA-MEM when using an NVIDIA A40. Using an NVIDIA A6000 and an NVIDIA A100, we achieved a wall-time speedup of up to 3.4x/3.8x over BWA-MEM2 and up to 6.1x/6.8x over BWA-MEM, respectively. In stage-wise comparison, the A40/A6000/A100 GPUs respectively achieved up to 3.7/3.8/4x, 2/2.3/2.5x, and 3.1/5/7.9x speedup on the three major stages of BWA-MEM: seeding and seed chaining, Smith-Waterman, and making SAM output. To the best of our knowledge, this is the first study that attempts to implement the entire BWA-MEM program on GPUs. Yi-Cheng Tu, Xiaoyi Lv |
ICS | 3 |
| 2023 | MLDF-Net: Metadata Based Multi-level Dynamic Fusion Network
Enguang Zuo, Chen Chen 0078, Yunling Wang, Xiaoyi Lv, Min Li 0093 |
PRCV (1) | 7 |
| 2023 | SUCOLA: Self-adaptive structure refinement unsupervised contrastive learning framework for food safety risk early warning
Enguang Zuo, Junyi Yan, Alimjan Aysa, Chen Chen 0078, Hongbing Ma, Xiaoyi Lv, Kurban Ubul |
Eng. Appl. Artif. Intell. | 7 |
| 2023 | Recognizing breast tumors based on mammograms combined with pre-trained neural networks
Yujie Bai, Min Li 0093, Xiaojing Gan, Chen Chen 0078, Xiaoyi Lv |
Multim. Tools Appl. | 7 |
| 2020 | Fully convolutional attention network for biomedical image segmentation
Junlong Cheng, Shengwei Tian, Long Yu 0001, Hongchun Lu, Xiaoyi Lv |
Artif. Intell. Medicine | 5 |