EDBT 2026 Demo / reviewers in the wild / expert
Zhaoyi Ye
dblp:319/7723
· DBLP profile ↗
10ranked-venue papers
2as first author
10since 2021 · last 2026
0000-0003-2193-9124ORCID · corroborated
Domains — the database's venue-derived domains; a paper can count in several
Applied, interdisciplinary, general and emerging computing · 7 · 2 first-author · 7 since 2021Artificial intelligence and machine learning · 3 · 3 since 2021Graphics, computer vision, multimedia, augmented reality and games · 1 · 1 first-author · 1 since 2021
| Year | Publication | Venue | Position |
|---|---|---|---|
| 2026 | Rethinking Multi-Center Semi-Supervised Breast Cancer Ultrasound Image Segmentation: An Intermediate-Domain PerspectiveabstractMulti-center breast ultrasound image segmentation aims to leverage limited labeled data from a single center to enhance model discriminability across unlabeled data from other centers. However, differences in equipment parameters, disease severity, and imaging conditions collectively contribute to significant cross-domain shifts in multi-center data. In a spirit of the golden mean, we argue that constructing an intermediate domain between the source and target domains can effectively improve model generalization. Therefore, we propose a Cross-domain Few-label Generalization (CFG) framework for multi-center breast ultrasound image segmentation. Specifically, we design the Intermediate Domain Generator (IDG) to generate intermediate domain samples that contain features from both the source and target domains bidirectionally, enabling the model to explicitly learn universal semantic representations. Additionally, we apply Swin Masked Autoencoder (MAE) to mask and reconstruct ultrasound images, simulating speckle noise encountered during clinical ultrasound acquisition, thereby increasing the diversity of intermediate domain samples. Furthermore, we integrate the Kolmogorov-Arnold Network (KAN) with UNet to construct KAN-UNet, integrating learnable spline functions directly onto the edges, enabling effective multi-scale perception of breast cancer lesion features. Experimental results show that even with limited labeled data from the source domain (BUSI-WHU), the CFG framework achieves a Kappa value of 77.17%, surpassing ten state-of-the-art methods and outperforming the second-best method by 0.78% across four multi-center ultrasound datasets (BUSI-WHU, BUSI, Dataset-B, and Dataset-C) collected from different medical centers. Zhaoyi Ye, Du Wang, Sheng Liu 0016, Liye Mei |
IEEE J. Biomed. Health Informatics | 1 |
| 2025 | FViM: Frequency Vision Mamba for Label-Free Cell Death Pathway Prediction in Lung Cancer Chemotherapy
Zhaoyi Ye, Shubin Wei, Liye Mei, Yueyun Weng, Qing Geng, Du Wang |
MICCAI (11) | 1 |
| 2025 | Visual fidelity and full-scale interaction driven network for infrared and visible image fusion
Liye Mei, Xinglong Hu, Zhaoyi Ye, Zhiwei Ye |
Pattern Recognit. | 3 |
| 2025 | DDRL: Domain Distribution Reconstruction Learning for Binary Change Detection in Remote Sensing ImagesabstractChange detection (CD) aims to identify and locate changes in the same observed surface coverage area across bitemporal images. This technique has widespread applications in urban planning, land use, and disaster damage extraction. Deep learning-based CD methods typically use learnable encoders to map bitemporal images to a common domain distribution space, allowing for the discrimination and localization of change and invariant features. However, due to differences in imaging mechanisms, seasons, and shooting angles, a large number of pseudochanges may easily appear, affecting the accurate recognition of the domain distribution space. In addition, binary CD focuses solely on whether scene targets have changed, resulting in change labels that encompass a variety of different objects, thus increasing the significance of intraclass differences. To address the aforementioned issues, we propose a domain distribution reconstruction learning (DDRL) framework for binary CD, which effectively mitigates the problem of pseudochanges by detecting abnormal feature domain distributions. Specifically, DDRL first extracts multiscale features from bitemporal images using a Siamese cross-window self-attention module, achieving feature domain transformation from the original space. Subsequently, it employs a graph attention enhanced (GAE) module to improve the low-level domain distribution, enabling it to focus on change regions. In addition, DDRL utilizes a cross-domain feature contrastive learning (CFCL) module for reconstructive learning of high-level fused features. This process ensures that intraclass features are compact, while interclass features are dispersed within the high-level domain distribution, thereby significantly improving the domain distribution representation to discriminate pseudochanges. Experimental results show that the proposed DDRL performs excellently across multiple public datasets, surpassing mainstream methods and significantly improving CD performance. The source code will be made available athttps://github.com/yzygit1230/DDRL. Wei Yang 0043, Zhaoyi Ye, Liye Mei, Yongxiang Yao, Yansheng Li 0001 |
IEEE Trans. Geosci. Remote. Sens. | 2 |
| 2025 | EMGANet: Edge-Aware Multi-Scale Group-Mix Attention Network for Breast Cancer Ultrasound Image SegmentationabstractBreast cancer is one of the most prevalent diseases for women worldwide. Early and accurate ultrasound image segmentation plays a crucial role in reducing mortality. Although deep learning methods have demonstrated remarkable segmentation potential, they still struggle with challenges in ultrasound images, including blurred boundaries and speckle noise. To generate accurate ultrasound image segmentation, this paper proposes the Edge-Aware Multi-Scale Group-Mix Attention Network (EMGANet), which generates accurate segmentation by integrating deep and edge features. The Multi-Scale Group Mix Attention block effectively aggregates both sparse global and local features, ensuring the extraction of valuable information. The subsequent Edge Feature Enhancement block then focuses on cancer boundaries, enhancing the segmentation accuracy. Therefore, EMGANet effectively tackles unclear boundaries and noise in ultrasound images. We conduct experiments on two public datasets (Dataset-B, BUSI) and one private dataset which contains 927 samples from Renmin Hospital of Wuhan University (BUSI-WHU). EMGANet demonstrates superior segmentation performance, achieving an overall accuracy (OA) of 98.56%, a mean IoU (mIoU) of 90.32%, and an ASSD of 6.1 pixels on the BUSI-WHU dataset. Additionally, EMGANet performs well on two public datasets, with a mIoU of 88.2% and an ASSD of 9.2 pixels on Dataset-B, and a mIoU of 81.37% and an ASSD of 18.27 pixels on the BUSI dataset. EMGANet achieves a state-of-the-art segmentation performance of about 2% in mIoU across three datasets. In summary, the proposed EMGANet significantly improves breast cancer segmentation through Edge-Aware and Group-Mix Attention mechanisms, showing great potential for clinical applications. Yazhao Mao, Jingwen Deng, Zhaoyi Ye, Lan Dong, Jinxuan Hou, Sheng Liu 0016, Du Wang, Shengrong Sun, Liye Mei |
IEEE J. Biomed. Health Informatics | 4 |
| 2025 | MRRM: Advanced Biomarker Alignment in Multi-Staining Pathology Images via Multi-Scale Ring Rotation-Invariant MatchingabstractPathology image matching is crucial for assisting pathologists in the comprehensive diagnosis of cancerous areas. However, variations in image rotation and staining caused by inherent slide imaging techniques increase the burden on pathologists, complicating the examination of cancer across different pathology slides. To address this challenge, we introduce multi-scale ring rotation-invariant matching (MRRM), which improves image matching efficiency using ring topology, assisting pathologists in robustly aligning biomarker information across various pathology images. Specifically, by employing multi-scale rings as convolution kernels, we accurately locate keypoints from the differencing of the ring pyramid, which not only enhances the likelihood of successful pathology image matching but also supports our feature descriptor in achieving advantageous performance in rotation-invariance. Experiments show that with manually annotated golden landmarks as the standard in 81 cases, exhibiting significantly superior matching accuracy (130.93 $\,\mu \mathrm{m}$) and a success rate of 93.83% compared to other methods, particularly in cases with rotated pathology images. This meets the routine diagnostic requirements of pathologists for cancer diagnosis. Taobo Hu, Zhengxiong Li, Mengping Long, Zhaoyi Ye, Yaxiaer Yalikun, Sheng Liu 0016, Yiqiang Liu, Du Wang, Jianghua Wu, Liye Mei |
IEEE J. Biomed. Health Informatics | 5 |
| 2024 | GTMFuse: Group-attention transformer-driven multiscale dense feature-enhanced network for infrared and visible image fusion
Liye Mei, Xinglong Hu, Zhaoyi Ye, Linfeng Tang, Xin Hao |
Knowl. Based Syst. | 3 |
| 2024 | Adjacent Self-Similarity 3-D Convolution for Multimodal Image RegistrationabstractSignificant challenges exist in the registration of multimodal images (MMIs) due to nonlinear radiation differences, variations in lighting, and interference from image noise. These issues often lead to unreliable similarity measurements and low accuracy in point matching during multimodal registration. To address these challenges, this letter introduces a novel MMI registration method based on adjacent self-similarity 3-D convolution (ASTC). The proposed method consists of three main steps: feature point extraction, where key points are uniformly extracted via the block-FAST method; ASTC salient feature construction, where a local adjacent self-similarity (ASS) model is employed to create multidimensional features; and feature structure enhancement, where a 3-D convolution is used for feature enhancement and finishing the process of image feature description. This letter evaluates the ASTC method against six sets of representative MMIs and compares it with six other algorithms. The results demonstrate that: 1) the ASTC algorithm effectively overcomes radiation distortion, intensity differences, and lighting differences in MMIs, leading to improved accuracy in point matching; and 2) the ASTC algorithm achieves higher matching efficiency and reduces time consumption, making it a practical choice for various data types. In summary, the proposed ASTC algorithm offers a robust solution for reliable registration of MMIs, addressing common challenges related to image differences and improving the overall accuracy of the process. The experimental data and code link used in this letter can be found athttps://github.com/yangwill81/ASTC. Wei Yang 0043, Liye Mei, Zhaoyi Ye, Ying Wang 0123, Xinglong Hu, Yongxiang Yao |
IEEE Geosci. Remote. Sens. Lett. | 3 |
| 2024 | SCD-SAM: Adapting Segment Anything Model for Semantic Change Detection in Remote Sensing ImageryabstractSemantic change detection (SCD) has gradually emerged as a prominent research focus in remote sensing image processing due to its critical role in earth observation applications. In view of its powerful semantic-driven feature extraction capability, the Segment Anything Model (SAM) has demonstrated its suitability across various visual scenes. However, it suffers from significant performance degradation when confronted with remote sensing images, especially those containing various ground objects that possess significant inter-class similarity and substantial intra-class variations. To address the above issues, we propose SCD-SAM, aiming to leverage the potent visual recognition capabilities of SAM for enhanced accuracy and robustness in SCD. Specifically, we introduce a contextual semantic change-aware dual encoder that combines MobileSAM and CNN to extract progressive semantic change features in parallel, and inject local features into the MobileSAM encoder through depth feature interaction to compensate for the Transformer’s limitations in perceiving local semantic details. Besides, in order to utilize the strong visual feature extraction capability of MobileSAM in remote sensing images, we propose a semantic adaptor that aggregates semantic-oriented information about changing objects. To better integrate the extracted contextual semantic information, we devise a progressive feature aggregation dual decoder that aggregates binary change features and semantic change features respectively, alleviating the semantic gap across different scales. The quantitative and visual results show that SCD-SAM outperforms the state-of-the-art SCD methods on publicly open SCD datasets (e.g., SECOND-CD and Landsat-CD). The code will be made available at https://github.com/yzygit1230/SCD-SAM. Liye Mei, Zhaoyi Ye, Hongzhu Wang, Ying Wang 0123, Wei Yang 0043, Yansheng Li 0001 |
IEEE Trans. Geosci. Remote. Sens. | 2 |
| 2021 | Scene Text Detection Based On Fusion NetworkabstractDue to the robustness resulted from scale transformation and unbalanced distribution of training samples in scene text detection task, a new fusion framework TSFnet is proposed in this paper. This framework is composed of Detection Stream, Judge Stream and Fusion Stream. In the Detection Stream, loss balance factor (LBF) is raised to improve the region proposal network (RPN). To predict the global text segmentation map, the algorithm combines regression strategy and case segmentation method. In the Judge Stream, a classification of the samples is proposed based on the Judge Map and the corresponding tags to calculate the overlap rate. As a support of Detection Stream, feature pyramid network is utilized in the algorithm to extract Judge Map and calculate LBF. In the Fusion Stream, a new fusion algorithm is raised. By fusing the output of the two streams, we can position the text area in the natural scene accurately. Finally, the algorithm is experimented on the standard data sets ICDAR 2015 and ICDAR2017-MLT. The test results show that the [Formula: see text] values are 87.8% and 67.57%, respectively, superior to the state-of-the art models. This proves that the algorithm can solve the robustness issues under the unbalance between scale transformation and training data. Xuezhuan Zhao, Lingling Li 0004, Lishen Pei, Zhaoyi Ye |
Int. J. Pattern Recognit. Artif. Intell. | 5 |