Zhenhong Du

dblp:152/4615 · DBLP profile ↗
← Back
22ranked-venue papers
3as first author
17since 2021 · last 2025
0000-0001-9449-0415ORCID · verified

Domains — the database's venue-derived domains; a paper can count in several

Applied, interdisciplinary, general and emerging computing · 11 · 11 since 2021Databases, data management, data science and information retrieval · 7 · 2 first-author · 4 since 2021Artificial intelligence and machine learning · 4 · 1 first-author · 2 since 2021
YearPublicationVenuePosition
2025 Using an attention-based architecture to incorporate context similarity into spatial non-stationarity estimation
abstract
Geographically weighted regression (GWR) facilitates spatial modeling by providing location-specific coefficients to capture spatial non-stationarity. GWR incorporates a distance decay effect, assigning greater weights to proximal observations under the assumption they exert more influence on the regression parameters. However, distant observations may share significant context similarities, such as socioeconomic or environmental factors, which can influence the regression model. This study introduces an attention-based architecture to address context similarity between samples. A deep learning model termed Context-Attention Geographically Weighted Regression (CatGWR) is proposed to integrate context similarity with distance-based proximity to enhance the estimation of spatial non-stationarity in spatial regression models. Such an integration results in contextualized spatial weights for CatGWR to identify the varying patterns of nonstationary relationships across different spatial locations and context conditions. Validation through simulation experiments and an empirical study on housing prices in Shenzhen, China, shows the superior predictive accuracy and robustness of CatGWR in modeling complex spatial interactions, especially under contextual influences, in which CatGWR improves the R2 of fit and prediction results by at least 6% compared to existing models. Future work will focus on optimizing bandwidth selection and exploring additional attention mechanisms to enhance model performance.
Sensen Wu, Jiale Ding, Ruoxu Wang, Ziyu Yin, Bo Huang 0001, Zhenhong Du
Int. J. Geogr. Inf. Sci.7
2025 Heterogeneous Contrastive Graph Fusion Network for Classification of Hyperspectral and LiDAR Data
abstract
In recent years, the rapid advancement of multi-sensory platforms has significantly increased the availability of multisource remote sensing data, facilitating its systematic application to various tasks. The joint classification of hyperspectral images (HSIs) and light detection and ranging (LiDAR) data remains a critical research topic, with a key challenge being the effective extraction and integration of complementary information from multi-source remote sensing data. However, existing graph convolutional networks (GCNs)-based methods often fail to account for the heterogeneous topological relationships between HSI and LiDAR. Moreover, the discriminative power of HSI and LiDAR features extracted by existing methods is insufficient. In addition, existing methods are unable to fully exploit the rich self-supervised information present in local neighborhood. To address these limitations, we propose a heterogeneous contrastive graph fusion network (HCGFN) for the joint classification of HSI and LiDAR data. First, we propose a branch enhancement module to enhance the discriminative power of HSI and LiDAR. Second, a contrastive learning module is introduced to effectively align HSI and LiDAR representations. Finally, we propose a dynamic heterogeneous graph structure learning module to model heterogeneous relationship and achieve efficient interaction and effective fusion between HSI and LiDAR. The extensive experimental results on three benchmark datasets indicate the effectiveness of the proposed HCGFN compared with other state-of-the-art methods. Specifically, under limited training samples, the proposed HCGFN outperformed state-of-the-art methods in overall accuracy by 5.10%, 2.46%, and 8.79% on datasets Trento, MUUFL, and Houston2013, respectively.
Haoyu Jing, Sensen Wu, Laifu Zhang, Fanen Meng, Zhenhong Du
IEEE Trans. Geosci. Remote. Sens.7
2025 LOGCAN++: Adaptive Local-Global Class-Aware Network for Semantic Segmentation of Remote Sensing Images
abstract
Remote sensing images are usually characterized by complex backgrounds, scale and orientation variations, and large intraclass variance. General semantic segmentation methods usually fail to fully investigate the above issues, and thus their performances on remote sensing image segmentation are limited. In this article, we propose our LOGCAN++, a semantic segmentation model customized for remote sensing images, which is made up of a global class-aware (GCA) module and several local class-aware (LCA) modules. The GCA module captures global representations for class-level context modeling to reduce the interference of background noise. The LCA module generates local class representations as intermediate perceptual elements to indirectly associate pixels with the global class representations, targeting dealing with the large intraclass variance problem. In particular, we introduce affine transformations in the LCA module for adaptive extraction of local class representations to effectively tolerate scale and orientation variations in remote sensing images. Extensive experiments on three benchmark datasets show that our LOGCAN++ outperforms current mainstream general and remote sensing semantic segmentation methods and achieves a better trade-off between speed and accuracy.
Rongrong Lian, Zhenkai Wu, Fan Yang 0100, Mengting Ma, Sensen Wu, Zhenhong Du, Wei Zhang 0243, Siyang Song
IEEE Trans. Geosci. Remote. Sens.8
2025 Optimized Attention-Enhanced Physics-Guided Neural Network for Satellite-Based Ocean Subsurface Temperature Predicting
Sensen Wu, Minlong Huang, Lian Feng, Chengfeng Le, Renyi Liu, Zhenhong Du
IEEE Trans. Geosci. Remote. Sens.11
2024 A neural network model to optimize the measure of spatial proximity in geographically weighted regression approach: a case study on house price in Wuhan
abstract
The estimation of spatial heterogeneity within real estate markets holds significant importance in house price modelling. However, employing a single or straightforward distance to measure spatial proximity is probably insufficient in complex urban areas, thereby resulting in an inadequate modelling of spatial heterogeneity. To address this issue, this paper incorporates multiple distance measures within a neural network framework to achieve an optimized measure of spatial proximity (OSP). Consequently, a geographically neural network weighted regression model with optimized measure of spatial proximity (osp-GNNWR) is devised for the purpose of spatially heterogeneous modeling. Trained as a unified model, osp-GNNWR obviates the need for separate pretraining of OSP. This enables OSP to delineate the modeled spatial process through a post hoc calculated value. Through simulation experiments and a real-world case study on house prices, the proposed model reaches more accurate descriptions of diverse spatial processes and exhibits better overall performance. The interpretable results of the case study in Wuhan demonstrate the efficacy of the osp-GNNWR model in addressing spatial heterogeneity within real estate markets, suggesting its potential for modelling and predicting complex geographical phenomena.
Jiale Ding, Wenying Cen, Sensen Wu, Bo Huang 0001, Zhenhong Du
Int. J. Geogr. Inf. Sci.7
2024 Aggregative and Contrastive Dual-View Graph Attention Network for Hyperspectral Image Classification
abstract
Graph convolutional networks (GCNs) have recently gained prominence in hyperspectral images (HSIs) classification tasks given their superior performance on non-Euclidean data. However, GCN-based methods are heavily reliant on complete graph structural information, which can cause the aggregation and transmission of information across nodes from differing classes, thereby compromising the classification performance. Furthermore, the scarcity of labeled pixels in HSIs often limits the representational capability of such methods. To address these issues, we propose an aggregative and contrastive dual-view graph attention network (ACoD-GAT) for HSI classification. Specifically, we present a progressive aggregation module, including a pixel clustering submodule and a node aggregation submodule to exploit semantic information at various levels. Besides, we integrate multiscale manipulation with a diffusion matrix to construct the dual view to further extract semantic information from both local and global perspectives. Moreover, we design an unsupervised contrastive loss function and a supervised contrastive loss function to facilitate contrastive learning on the dual view, improving the representational capabilities of ACoD-GAT with very few labeled samples. The extensive experimental results on four benchmark datasets demonstrate the superiority of the proposed ACoD-GAT compared with other state-of-the-art methods.
Haoyu Jing, Sensen Wu, Laifu Zhang, Fanen Meng, Tian Feng 0001, Zhenhong Du
IEEE Trans. Geosci. Remote. Sens.8
2024 A Conditional Diffusion Model With Fast Sampling Strategy for Remote Sensing Image Super-Resolution
abstract
Conventional deep learning-based methods for single remote sensing image super-resolution (SRSISR) have made remarkable progress. However, the super-resolution (SR) outputs of these methods are yet to become sufficiently satisfactory in visual quality. Recent diffusion model-based generative deep learning models are capable to enhance the visual quality of output images, but this capability is limited due to their sampling efficiency. In this article, we propose FastDiffSR, an SRSISR method based on a conditional diffusion model. Specifically, we devise a novel sampling strategy to reduce the number of sampling steps required by the diffusion model while ensuring the sampling quality. Meanwhile, the residual image is adopted to reduce computational costs, demonstrating that integrating channel attention and spatial attention begets a further improvement in the visual quality of output images. Compared to the state-of-the-art (SOTA) convolutional neural network (CNN)-based, GAN-based, and Transformer-based SR methods, our FastDiffSR improves the learned perceptual image patch similarity (LPIPS) by 0.1–0.2 and achieves better visual results in some real-world scenes. Compared with existing diffusion-based SR methods, our FastDiffSR achieves significant improvements in pixel-level evaluation metric peak signal-noise ratio (PSNR) while having smaller model parameters and obtaining better SR results on Vaihingen data with faster inference time by 2.8–28 times, showing excellent generalization ability and time efficiency. Our code will be open source athttps://github.com/Meng-333/FastDiffSR.
Fanen Meng, Haoyu Jing, Laifu Zhang, Yingchao Ren, Sensen Wu, Tian Feng 0001, Renyi Liu, Zhenhong Du
IEEE Trans. Geosci. Remote. Sens.10
2024 Single Remote Sensing Image Super-Resolution via a Generative Adversarial Network With Stratified Dense Sampling and Chain Training
abstract
Super-resolution (SR) methods have significantly contributed to the improvement of the spatial resolution of remote sensing (RS) images. The development of deep learning empowers novel methods to learn informative feature representation from massive low-resolution (LR) and high-resolution (HR) image pairs. Conventional RS image SR methods, however, may fail in large-scale ($\times 8$and$\times 9$) SR tasks. Specifically, a larger scale factor corresponds to less information in LR images, which is a considerable challenge to SR. To address the issue, we propose a novel method for single RS image SR (SRSISR) based on stratified dense sampling to effectively extract image features. Specifically, the proposed SR dense-sampling residual attention network (SRDSRAN) combines dense sampling and residual learning to improve multilevel feature fusion and gradient propagation and employs local and global attentions to learn important features and long-range interdependence in the channel and spatial dimensions. Meanwhile, we also devise a discriminator model using local and global attentions and with the loss function integrating${L}_{1}$pixel loss,${L}_{1}$perceptual loss, and relativistic adversarial loss to obtain the perceptually realistic images. Besides, we introduce a chain training to promote performance and expedite the training process for large-scale SR. Experimental results on UC Merced image and other multispectral data demonstrated that our SRDSRAN outperformed the current state-of-the-art methods quantitatively and in visual quality and obtained a higher classification accuracy in scene classification, proving its potential for applications with other downstream tasks. The code of SRADSGAN will be available athttps://github.com/Meng-333/SRADSGAN.
Fanen Meng, Sensen Wu, Zhe Zhang 0040, Tian Feng 0001, Renyi Liu, Zhenhong Du
IEEE Trans. Geosci. Remote. Sens.7
2024 Causality-Guided Stepwise Intervention and Reweighting for Remote Sensing Image Semantic Segmentation
abstract
Semantic segmentation is one of the most significant tasks in remote sensing (RS) image interpretation, which focuses on learning global and local information to infer the semantic label of each pixel. Previous studies devise encoder-decoder structured deep learning (DL) models to extract global and local features from RS images with the help of pretraining knowledge to predict semantic labels. However, due to the common heterogeneity between the data for pretraining and the data to be semantically segmented, these models fail to learn general features appropriate to RS datasets. In this article, we propose a novel formulation of the above problem from a causal perspective, where the learned features from pretrained models result from causality and spurious correlations, and only the former carries general information that remains invariant regardless of the exact task and dataset. Based on the above formulation, we propose stepwise intervention and reweighting (SIR). It can reduce the confounding bias introduced by the pretraining knowledge and improve the model’s ability to learn general features, making semantic segmentation of RS images benefit more from pretraining. Besides, we conduct a detailed theoretical analysis of our methods and conduct extensive experiments on two widely used public RS datasets. Experimental results demonstrate that applying SIR to encoder-decoder semantic segmentation models achieves performance improvements, proving the effectiveness and application values of the proposed method.
Baohong Li, Laifu Zhang, Kun Kuang 0001, Sensen Wu, Tian Feng 0001, Zhenhong Du
IEEE Trans. Geosci. Remote. Sens.8
2024 A Downscaling Framework for Urban Nighttime Light Based on Multifactor Geographically Neural Network Weighted Regression
abstract
Downscaling nighttime light (NTL) from satellite imagery presents valuable applications at a more detailed spatial scale, especially in the realms of urban expansion and socio-economic assessment. Nevertheless, due to the complexity of geographical conditions and uncertainties in the relationships among multiple factors, the precision of NTL downscaling often encounters constraints. In this work, an incorporated multifactor geographically neural network weighted regression (MF-GNNWR) NTL downscaling framework is proposed to solve the spatial nonstationarity in high-heterogeneous urban areas, which mainly uses geographically neural network weighted regression (GNNWR) combined with multiple factors including surface physical characteristics, socio-economic attributes, and human activities to improve the accuracy of NTL, particularly in urban regions with complicated land cover. The findings illustrate that the MF-GNNWR framework displays finer downscaling accuracy on different land cover, effectively enhancing data quality. Notably, our findings underscore the pronounced influence of socio-economic and human activity factors on NTL downscaling. Comparative analysis against several alternative downscaling methodologies reveals that the MF-GNNWR framework outperforms them, exhibiting a remarkable 23.10% improvement in the Pearson correlation coefficient (r) and achieving a root-mean-square error (RMSE) of$16.95~\text {{nW/c}{m}}^{2}{/\text {sr}}$, and after residual compensation, r continue s to increase by 1.5%, while RMSE decreases by$0.157~\text {nW/cm}^{2}{/\text {sr}}$. These findings highlight the efficacy of the proposed framework in downscaling NTL, underscoring its advantages and practical utility.
Laifu Zhang, Sensen Wu, Minggao Liang, Haoyu Jing, Fanen Meng, Zhenhong Du
IEEE Trans. Geosci. Remote. Sens.10
2022 Geographically convolutional neural network weighted regression: a method for modeling spatially non-stationary relationships based on a global spatial proximity grid
abstract
Geographically weighted regression (GWR) is a classical method of modeling spatially non-stationary relationships. The geographically neural network weighted regression (GNNWR) model solves the problem of the inaccurate construction of spatial weight kernels using a spatially weighted neural network. However, when the spatial distribution of observations is uneven, the spatial proximity expression in the input of GWR and GNNWR models does not fully represent the impact of the whole research space on the estimating point. Therefore, we established a global spatial proximity grid (GSPG) to express the spatial proximity of each estimating point and proposed a spatially weighted convolutional neural network (SWCNN) to extract the relationship between the GSPG and spatial weights. Finally, we proposed a geographically convolutional neural network weighted regression (GCNNWR) model combining SWCNN and ordinary linear regression (OLR) model to estimate spatial non-stationarity. We used two case studies of simulated data and real environment data to demonstrate the advancements of the GCNNWR model. The GCNNWR model achieved higher estimation accuracy and greater predictive power than the OLR, GWR, multi-scale GWR (MGWR), and GNNWR models. Moreover, the GCNNWR model maintained its better stability and accuracy in estimating spatially non-stationary relationships when the distribution of observations was uneven.
Sensen Wu, Hongye Zhou, Feng Zhang 0009, Bo Huang 0001, Zhenhong Du
Int. J. Geogr. Inf. Sci.7
2022 A neural network framework for fine-grained tropical cyclone intensity prediction
Zhe Zhang 0040, Xuying Yang, Lingfei Shi, Zhenhong Du, Feng Zhang 0009, Renyi Liu
Knowl. Based Syst.5
2022 A neural network with spatiotemporal encoding module for tropical cyclone intensity estimation from infrared satellite image
Zhe Zhang 0040, Xuying Yang, Zhenhong Du
Knowl. Based Syst.6
2022 Single-Image Super-Resolution for Remote Sensing Images Using a Deep Generative Adversarial Network With Local and Global Attention Mechanisms
abstract
Super-resolution (SR) technology is an important way to improve spatial resolution under the condition of sensor hardware limitations. With the development of deep learning (DL), some DL-based SR models have achieved state-of-the-art performance, especially the convolutional neural network (CNN). However, considering that remote sensing images usually contain a variety of ground scenes and objects with different scales, orientations, and spectral characteristics, previous works usually treat important and unnecessary features equally or only apply different weights in the local receptive field, which ignores long-range dependencies; it is still a challenging task to exploit features on different levels and reconstruct images with realistic details. To address these problems, an attention-based generative adversarial network (SRAGAN) is proposed in this article, which applies both local and global attention mechanisms. Specifically, we apply local attention in the SR model to focus on structural components of the earth’s surface that require more attention, and global attention is used to capture long-range interdependencies in the channel and spatial dimensions to further refine details. To optimize the adversarial learning process, we also use local and global attentions in the discriminator model to enhance the discriminative ability and apply the gradient penalty in the form of hinge loss and loss function that combines$L1$pixel loss,$L1$perceptual loss, and relativistic adversarial loss to promote rich details. The experiments show that SRAGAN can achieve performance improvements and reconstruct better details compared with current state-of-the-art SR methods. A series of ablation investigations and model analyses validate the efficiency and effectiveness of our method.
Sébastien Mavromatis, Feng Zhang 0009, Zhenhong Du, Jean Sequeira, Xianwei Zhao, Renyi Liu
IEEE Trans. Geosci. Remote. Sens.4
2022 A Dynamic Pyramid Tilling Method for Traffic Data Stream Based on Flink
abstract
Traffic guidance, traffic management and emergency vehicle traffic all require keeping abreast of traffic status. Intelligent Transportation Systems (ITS) is highly expected to provide real-time traffic condition information service. To achieve this, the capability of handling dynamic data stream collected from multi traffic monitoring sources and serving the public with information timely is essential for ITS. With the wide spread of Internet of Things technology, not only the amount, but also the spatial and temporal resolutions of real-time traffic data have explosive growth, thereby enhancing the difficulty of real-time traffic data processing in ITS. Web pyramid map tiles is wide accepted for massive spatial data service, and the latency of tile generation significantly reduces the timeliness of information transmission and the reliability of services. A Flink-based method for dynamic pyramid tile generation and updating is proposed here. Take advantages of combining grid indexes, employing data partition and window selection mechanisms, and applying iterative computational characteristics for resampling, the distributed dynamic pyramid map tile generation algorithm (DPTG), can quickly visualize real-time spatial traffic data with digital map tiles. Taking the national highway road data from China as an example, the experimental results show that the Flink-based DPTG method has high efficiency and scalability in both batch processing and stream processing mode, which highlights the capability of the proposed method to support real-time traffic monitoring data processing for timely large-scale public service in ITS.
Linshu Hu, Feng Zhang 0009, Mengjiao Qin, Zhiyi Fu, Zhende Chen, Zhenhong Du, Renyi Liu
IEEE Trans. Intell. Transp. Syst.6
2021 Simulating City Expansion Using a CA Urban Growth Model, Through a Case Study of Nairobi, Kenya
abstract
The acceleration of urbanization and industrialization has augmented the development of existing urban centers and hastened urban expansion. Especially in developing countries like Kenya. Recently, Cellular Automata has witnessed significant technological advancements as many CA-based urban models have been developed. These CA-based dynamic spatial urban models provide an improved ability to forecast and assess future urban growth and to create planning scenarios. In this research, we have sought to understand how marginally closer functional towns affect each other growth through time using SLEUTH, a cellular automation urban growth model. The existing data from 1986 to 2015 was used as input and calibration and prediction done to simulate the next 35 years and different indices of expansion employed to expound these changes. Classification results showed that the urban area had expanded from 39.62 km2in 1986 to 367.02 km2in 2015 with most urban growth in the North-east and the west of the Nairobi CBD. The prediction poised the city urban fringes to continue growing at a rate of 5.01% between 2015 and 2020, and slower rate of 2.51% between 2040 and 2050. The urban growth is much determined by breeding from existing urban patches along the road networks with a very small spontaneous growth. This infer that most expansion is determined by existing urban patches
Lingfei Shi, Feng Zhang 0009, Zhenhong Du
IGARSS3
2021 Geographically and temporally neural network weighted regression for modeling spatiotemporal non-stationary relationships
abstract
Geographically weighted regression (GWR) and geographically and temporally weighted regression (GTWR) are classic methods for estimating non-stationary relationships. Although these methods have been widely used in geographical modeling and spatiotemporal analysis, they face challenges in adequately expressing space-time proximity and constructing a kernel with optimal weights. This probably results in an insufficient estimation of spatiotemporal non-stationarity. To address complex non-linear interactions between time and space, a spatiotemporal proximity neural network (STPNN) is proposed in this paper to accurately generate space-time distance. A geographically and temporally neural network weighted regression (GTNNWR) model that extends geographically neural network weighted regression (GNNWR) with the proposed STPNN is then developed to effectively model spatiotemporal non-stationary relationships. To examine its performance, we conducted two case studies of simulated datasets and environmental modeling in coastal areas of Zhejiang, China. The GTNNWR model was fully evaluated by comparing with ordinary linear regression (OLR), GWR, GNNWR, and GTWR models. The results demonstrated that GTNNWR not only achieved the best fitting and prediction performance but also exactly quantified spatiotemporal non-stationary relationships. Further, GTNNWR has the potential to handle complex spatiotemporal non-stationarity in various geographical processes and environmental phenomena.
Sensen Wu, Zhenhong Du, Bo Huang 0001, Feng Zhang 0009, Renyi Liu
Int. J. Geogr. Inf. Sci.3
2020 Geographically neural network weighted regression for the accurate estimation of spatial non-stationarity
abstract
Geographically weighted regression (GWR) is a classic and widely used approach to model spatial non-stationarity. However, the approach makes no precise expressions of its weighting kernels and is insufficient to estimate complex geographical processes. To resolve these problems, we proposed a geographically neural network weighted regression (GNNWR) model that combines ordinary least squares (OLS) and neural networks to estimate spatial non-stationarity based on a concept similar to GWR. Specifically, we designed a spatially weighted neural network (SWNN) to represent the nonstationary weight matrix in GNNWR and developed two case studies to examine the effectiveness of GNNWR. The first case used simulated datasets, and the second case, environmental observations from the coastal areas of Zhejiang. The results showed that GNNWR achieved better fitting accuracy and more adequate prediction than OLS and GWR. In addition, GNNWR is applicable to addressing spatial non-stationarity in various domains with complex geographical processes.
Zhenhong Du, Sensen Wu, Feng Zhang 0009, Renyi Liu
Int. J. Geogr. Inf. Sci.1
2019 A matrix completion-based multiview learning method for imputing missing values in buoy monitoring data
Mengjiao Qin, Zhenhong Du, Feng Zhang 0009, Renyi Liu
Inf. Sci.2
2018 A spatiotemporal regression-kriging model for space-time interpolation: a case study of chlorophyll-a prediction in the coastal areas of Zhejiang, China
abstract
Spatiotemporal kriging (STK) is recognized as a fundamental space-time prediction method in geo-statistics. Spatiotemporal regression kriging (STRK), which combines space-time regression with STK of the regression residuals, is widely used in various fields, due to its ability to take into account both the external covariate information and spatiotemporal autocorrelation in the sample data. To handle the spatiotemporal non-stationary relationship in the trend component of STRK, this paper extends conventional STRK to incorporate it with an improved geographically and temporally weighted regression (I-GTWR) model. A new geo-statistical model, named geographically and temporally weighted regression spatiotemporal kriging (GTWR-STK), is proposed based on the decomposition of deterministic trend and stochastic residual components. To assess the efficacy of our method, a case study of chlorophyll-a (Chl-a) prediction in the coastal areas of Zhejiang, China, for the years 2002 to 2015 was carried out. The results show that the presented method generated reliable results that outperform the GTWR, geographically and temporally weighted regression kriging (GTWR-K) and spatiotemporal ordinary kriging (STOK) models. In addition, employing the optimal spatiotemporal distance obtained by I-GTWR calibration to fit the spatiotemporal variograms of residual mapping is confirmed to be feasible, and it considerably simplifies the residual estimation of STK interpolation.
Zhenhong Du, Sensen Wu, Mei-Po Kwan, Chuanrong Zhang, Feng Zhang 0009, Renyi Liu
Int. J. Geogr. Inf. Sci.1
2018 Multistep-ahead forecasting of chlorophyll a using a wavelet nonlinear autoregressive network
Zhenhong Du, Mengjiao Qin, Feng Zhang 0009, Renyi Liu
Knowl. Based Syst.1
2017 Red tide time series forecasting by combining ARIMA and deep belief network
Mengjiao Qin, Zhihang Li, Zhenhong Du
Knowl. Based Syst.3