Keyun Zhao

dblp:330/8387 · DBLP profile ↗
← Back
8ranked-venue papers
0as first author
8since 2021 · last 2024
0000-0003-3497-9360ORCID · corroborated

Domains — the database's venue-derived domains; a paper can count in several

Applied, interdisciplinary, general and emerging computing · 7 · 7 since 2021Artificial intelligence and machine learning · 1 · 1 since 2021
YearPublicationVenuePosition
2024 GaMPF: A Full-Scale Gated Message Passing Framework Based on Collaborative Estimation for VHR Remote Sensing Image Change Detection
abstract
With the maturity and popularization of high-performance sensor technology, it is now possible to acquire huge amounts of very high-resolution (VHR) remote sensing images. The change detection (CD) for VHR images is currently receiving special attention for remote sensing earth observation applications, however, as a hot research field, it needs to be studied in depth to improve the detection accuracy of fine changes. To this end, a full-scale gated message passing framework (GaMPF) based on collaborative estimation for VHR remote sensing image change detection is proposed in this paper. On one hand, the key embedding representation is generated for each feature map by means of the collaborative estimation (CE) strategy; On the other hand, grounded in timing analysis, bitemporal features are sent selectively on dual paths according to the full-scale gated (FsG) mechanism. Specifically, this framework consists of the following four components: 1) Taking shared-weights Siamese network as an encoder to extract multi-scale features; 2) Generate a set of shared compact bases under the CE strategy and infer the key embedding representations on the basis of the shared bases for feature maps at the same level, considering the representations as the gated switches; 3) FsG mechanism is used as the mode of message passing between bitemporal images, which guides the information can be transmitted simultaneously on both within-and cross-temporal paths. 4) Creating a stepwise dense fusion module (DFM) as a decoder for predicting the change map. Experimental results show that the GaMPF proposed in this paper outperforms existing SOTA methods, and is particularly good at detecting edges and small objects. The source code will be released at https://github.com/zxylnnu/GaMPF.
Xiao-Yang Zhao 0003, Keyun Zhao, Siyao Li, Chuanming Song 0001, Xiang-Hai Wang 0001
IEEE Trans. Geosci. Remote. Sens.2
2023 MCT-Net: Multi-hierarchical cross transformer for hyperspectral and multispectral image fusion
Xiang-Hai Wang 0001, Xinying Wang 0005, Ruoxi Song, Xiao-Yang Zhao 0003, Keyun Zhao
Knowl. Based Syst.5
2023 GTMSiam: Gated Transmitting-Based Multiscale Siamese Network for Hyperspectral Image Change Detection
abstract
Hyperspectral image change detection (HSI-CD) is a technique that detects changes in land cover occurring in a specific area within a closed time. At present, most existing methods for HSI-CD employ exceedingly intricate network architectures, leading to a high model complexity that hampers the achievement of a favorable trade-off between change detection accuracy and timeliness. Furthermore, existing methods often confine the feature extraction process to a single scale rather than multiple diverse scales. However, employing a multiscale approach for feature extraction allows for capturing finer-grained features encompassing more intricate details, as well as coarser-grained features that aggregate local information over a larger range. On the other hand, most existing methods overemphasize the complexity of the feature extraction process and underestimate the importance of the conversion process from bi-temporal features to valuable change features. To this end, a gated transmitting based multiscale siamese network (GTMSiam) is proposed, which mainly contains the following two portions: 1) dual branches with the siamese structure, which capture spatial features of the HSIs at multiple scales while preserving rich spectral information. Moreover, the siamese design effectively reduces the network parameters, thereby alleviating the computational complexity of the model. 2) gated change information transmitting module (GTM), which utilizes gated neural units to transform bi-temporal image features into land cover change information, while progressively transmitting change information at different scales. This enables the network to leverage diverse scale change information for comprehensive discrimination of land object changes. Experimental results on three publicly available datasets demonstrate the superior performance of the proposed GTMSiam. Simultaneously, the complexity analysis experiment proves that the GTMSiam can give consideration to both detection performance and timeliness. The source code of this letter will be released at https://github.com/zkylnnu/GTMSiam.
Xiang-Hai Wang 0001, Keyun Zhao, Xiao-Yang Zhao 0003, Siyao Li
IEEE Geosci. Remote. Sens. Lett.2
2023 BiG-FSLF: A Cross Heterogeneous Domain Few-Shot Learning Framework Based on Bidirectional Generation for Hyperspectral Image Change Detection
abstract
In recent years, hyperspectral image change detection (HSI-CD) based on deep learning has achieved high detection accuracy, but these methods obtain excellent detection results usually rely on having sufficient labeled samples to train the network. However, the production of HSI label is difficult, costly and inefficient. In practical tasks, often only a limited number of labeled samples can be obtained due to the limitation of timeliness. To address this problem, a cross heterogeneous domain few-shot learning framework based on bidirectional generation (BiG-FSLF) is proposed for HSI-CD, which aims to solve the few-shot problem of HSI-CD by few-shot learning (FSL), and to assist HSI-FSL perform better by obtaining learnable changed information (i.e., empirical knowledge) from another remote sensing data. Specifically, a multitask generation encoder (MLGenE) is designed to take on both the tasks of FSL and domain adaptation to achieve HSI-CD under the condition of cross heterogeneous domain few-shot. First, we take any pair of image data in a very high resolution image (VHRI) CD dataset as the source domain and HSI is used as the target domain, using sufficient labeled samples in source domain and a small number of labeled samples in target domain for FSL. Meanwhile, a bidirectional generation domain adaptation (BiGDA) method based on generative adversarial strategy is proposed to achieve adaptive alignment of the two heterogeneous domains (source and target domains) feature distributions, to mitigate the impact of the domain shift problem inherent to cross domain data on FSL. Abundant experiments with only five training samples on the publicly available popular HSI-CD datasets confirm that the proposed method can show great detection performance. The source code of the proposed framework will be released at https://github.com/lsylnnu/BiG-FSLF.
Xiang-Hai Wang 0001, Siyao Li, Xiao-Yang Zhao 0003, Keyun Zhao
IEEE Trans. Geosci. Remote. Sens.4
2023 TriTF: A Triplet Transformer Framework Based on Parents and Brother Attention for Hyperspectral Image Change Detection
abstract
Hyperspectral image (HSI) change detection (CD) is a technique to accurately detect land cover changes by using HSIs with rich spatial-spectral information. In recent years, the HSI-CD methods based on convolutional neural networks (CNNs) have achieved great success because of their flexible and effective feature extraction ability. However, these methods often take the HSI patches as the input of the networks, which undoubtedly hinders the overall perception of the HSIs. Meanwhile, the valuable temporal information in HSIs is often underutilized. For this end, a triplet transformer framework (TriTF) based on parents-temporal attention and brother-spatial attention is proposed for HSI-CD. The proposed framework mainly contains the following three parts: 1) Transformer-based network backbone, which uses the self-attention to capture the correlation between arbitrarily two pixels in the same patch and extracts the global spatial correlation in the unit of encoded input patches; 2) parents-temporal attention (PTA) branch. Unlike the previous cross-temporal attention mechanisms of the “T1↔T2” mode which only consider the interaction between bi-temporal HSIs, this paper constructs a novel PTA of the “T1→T3←T2” mode which takes the difference-temporal image T3 as the core. The impact of bi-temporal HSIs on the land cover changes is more concerned in the PTA; 3) brother-spatial attention (BSA) branch. The most similar patch in the current training batch of each patch is defined as its brother patch. Furthermore, cross-spatial attention is applied to propagate the features of the brother patch to the current patch. Thus, the middle- and long-range dependencies can be utilized and the scope of feature propagation can be extended. In this paper, the experiments under low and high sampling rates are conducted and proved the outstanding change detection performance of the proposed TriTF when compared with abundant state-of-the-art (SOTA) CD algorithms. The source code of this paper will be released at https://github.com/zkylnnu/TriTF.
Xiang-Hai Wang 0001, Keyun Zhao, Xiao-Yang Zhao 0003, Siyao Li
IEEE Trans. Geosci. Remote. Sens.2
2023 GeSANet: Geospatial-Awareness Network for VHR Remote Sensing Image Change Detection
abstract
The characteristics of very high resolution (VHR) remote sensing images (RSIs) have higher spatial resolution inherently, and are easier to obtain globally compared with hyperspectral images (HSIs), making it possible to detect small-scale land cover changes in multiple applications. RSI change detection (RSI-CD) based on deep learning has been paid attention to and become a frontier research field in recent years, and is currently facing two challenging problems: The first is high dependence on registration between bi-temporal images caused by high spatial resolution; The other is high pseudo-change information response caused by low spectral resolution. In order to address the above-mentioned two problems, a novel RSI-CD framework called Geospatial-Awareness Network (GeSANet) based on the geospatial Position Matching Mechanism (PMM) with multi-level adjustment and the geo-spatial Content Reasoning Mechanism (CRM) with diverse pseudo-change information filtering is proposed. First of all, the PMM assigns independent two-dimensional offset coordinates to each position in the previous temporal image, afterwards, bilinear interpolation is employed to obtain the subpixel feature value after the offset, and the sparse results based on the difference are transmitted to the next level prediction to realize multi-level geospatial correction. The CRM extracts global features from the corrected sparse feature map in terms of dimensions, implementing effective discriminant feature extraction on basis of the original feature map in a stepwise refinement manner through the cross-dimension exchange mechanism, to filter out various pseudo-change information as well as maintain real change information. Comparison experiments with five recent SOTA methods are carried out on two popular datasets with diverse changes, the results show that the proposed method has good robustness and validity for multi-temporal RSI-CD. In particular, it has a strong comparative advantage in detecting small entity changes and edge details. The source code of the proposed framework can be downloaded from https://github.com/zxylnnu/GeSANet.
Xiao-Yang Zhao 0003, Keyun Zhao, Siyao Li, Xiang-Hai Wang 0001
IEEE Trans. Geosci. Remote. Sens.2
2022 CSDBF: Dual-Branch Framework Based on Temporal-Spatial Joint Graph Attention With Complement Strategy for Hyperspectral Image Change Detection
abstract
Hyperspectral image (HSI) change detection (CD) aims at obtaining internal components’ change information of land cover and land use. In recent years, the development of convolutional neural networks (CNNs) has greatly promoted the research progress in this field. However, the fixed small-size convolution kernels used by CNNs have severely limited the receptive field of information. Another defect of most CNN-based models is their strong dependence on samples, and they are not competent for tasks with a small number of samples. Besides, the traditional CNN-based models can only perform convolution to learn the spatial–spectral features in the Euclidean space, which is not conducive to capturing the geometric changes in land covers in the HSIs. Differently, the graph attention network (GAT) has come into prominence due to its ability to capture the holistic topology structure of images flexibly, and the attention coefficients can be used to effectively model the long-range correlations between land covers. The semi-supervised nature of GAT is also well-suited to handle HSI-CD tasks with limited samples. Nevertheless, the pixel-level topology structure often generates expensive computational costs. To this end, a dual-branch framework based on temporal–spatial joint graph attention (TSJGAT) with complement strategy (CSDBF) is proposed for HSI-CD, which extracts superpixel- and pixel-level features from bitemporal HSIs in parallel and enables them to complement each other. The proposed CSDBF mainly consists of two branches: superpixel-level feature extraction branch (S-branch) and pixel-level feature extraction branch (P-branch). In the S-branch, we introduce the idea of GAT into HSI-CD for the first time and propose a novel TSJGAT module. Thus, the temporal–spatial features of HSIs are propagated and aggregated on the nonlinear graph structure, which makes the changed regions more discriminable. In the P-branch, pixel-level features are obtained by CNNs to correct uncertain factors caused by superpixel segmentation in the S-branch, which is complementary to the S-branch and lays a foundation for more accurate CD. Abundant experiments show that compared with other pioneer methods, the proposed CSDBF can improve the Kappa coefficient by more than 1.9% and 2.5% on average in general sampling rate situations and a low sampling rate situation, respectively, which shows better robustness and better detection accuracy than most existing state-of-the-art methods. The source code of this article can be downloaded fromhttps://github.com/zkylnnu/CSDBF.
Xiang-Hai Wang 0001, Keyun Zhao, Xiao-Yang Zhao 0003, Siyao Li
IEEE Trans. Geosci. Remote. Sens.2
2022 FSL-Unet: Full-Scale Linked Unet With Spatial-Spectral Joint Perceptual Attention for Hyperspectral and Multispectral Image Fusion
abstract
The application of hyperspectral image (HSI) is more and more extensive, but the lower spatial resolution seriously affects its application effect. Using low-resolution hyperspectral image (LR-HSI) and high-resolution multispectral image (MSI) fusion technology to achieve super-resolution reconstruction of HSI has become a mainstream method. However, most of the existing fusion methods do not make full use of the large-scale range of remote sensing images, and neglect the preservation of spatial-spectral information in the fusion process. Considering that the spectral information in fused high-resolution hyperspectral image (HR-HSI) mainly depends on HSI, and the spatial information mainly depends on MSI, this paper proposes a full-scale linked Unet with spatial-spectral joint perceptual attention for hyperspectral and multispectral image fusion (FSL-Unet). The FSL-Unet consists of two modules, the first is spatial-spectral attention extraction module (SSAE), which is used to calculate the spectral attention of LR-HSI and the spatial attention of HR-MSI at different scales. The second is the full-scale link U-shaped fusion module (FLUF), which adopts a multi-level feature extraction strategy, using denser full-scale skip connections to explore feature information in a finer-grained range, enabling flexible combination of multi-scale and multi-path features. At the same time, we propose spatial-spectral joint peceptual attention (SSJPA) on the encoder side of FLUF. SSJPA can make full use of the attention maps computed by the SSAE, and then effectively embed spatial and spectral information into the fused image, enabling uninterrupted information transfer and aggregation. To demonstrate the effectiveness of FSL-Unet, we selected five public hyperspectral datasets for experiments. Compared with other eight state-of-the-art fusion methods, the experimental results show that the FSL-Unet achieves competitive results. The source code for FSL-Unet can be downloaded from https://github.com/wxy11-27/FSL-Unet.
Xiang-Hai Wang 0001, Xinying Wang 0005, Keyun Zhao, Xiao-Yang Zhao 0003, Chuanming Song 0001
IEEE Trans. Geosci. Remote. Sens.3