VLDB 2026 Research / reviewers in the wild / expert
Bo Xu 0003
dblp:26/1194-3
· DBLP profile ↗
5ranked-venue papers
0as first author
5since 2021 · last 2025
0000-0001-6049-8005ORCID · verified
Domains — the database's venue-derived domains; a paper can count in several
Applied, interdisciplinary, general and emerging computing · 5 · 5 since 2021
| Year | Publication | Venue | Position |
|---|---|---|---|
| 2025 | Higher Order Energy-Optimized Air-Ground Individual Tree Segmentation With Crown Morphology FidelityabstractAccurate acquisition of carbon sequestration information at the individual tree scale is crucial for the refined assessment and management of carbon stocks in urban ecosystems. Traditional optical remote sensing methods, limited by two-dimensional spectral features, are more suitable for large-scale forest carbon stock estimation but struggle to achieve refined quantification at the individual tree level. LiDAR technology, which directly characterizes three-dimensional structural parameters of trees through point clouds, significantly outperforms optical methods in improving both segmentation quantity and morphological accuracy. To achieve a refined assessment and management of the urban carbon storage, on the one hand, we propose to integrate UAV-borne and ground-based point clouds, thereby overcoming the observational limitations of single-source data. On the other hand, to address the problems of under- and over-segmentation and morphological distortion in individual tree segmentation. We have respectively proposed two extreme value segmentation models and a high-order energy morphological optimization model. Specifically, the Canopy Skyline Extremum (CSE) model can solve the under-segmentation problem caused by the absence of canopy extreme points and the over-segmentation problem caused by pseudo-extreme points from the perspective of the canopy dimension. The Vertical Distribution Extremum (VDE) model of tree can solve the under-segmentation problem caused by the sparsity of trunk point clouds from the trunk dimension. For crown morphology fidelity, we construct, a voxel-based spatial correlation graph is constructed to characterize the distribution of branches and leaves. We use the Markov Random Field (MRF) energy optimization framework, integrate the prior knowledge of tree ecology, establish a high-order energy model that satisfies the objective cognitive morphology of trees, dynamically assign the membership relationship of branches and leaves, effectively solve the problem of interlacing and adhesion of branches and leaves between adjacent trees, overcome the problem of tree morphological distortion caused by the "one-size-fits-all" approach in traditional individual tree segmentation, improve the calculation accuracy of the tree crown diameter and canopy volume, and enhance the estimation accuracy of the carbon storage of individual trees to the decimeter level. Experiments were conducted on six datasets from two typical urban scenarios: street trees and landscaped gardens. Results validate the superior performance of our morphology-faithful segmentation. Quantitative comparisons with the state-of-the-arts show that our approach achieves optimal performance in both segmentation accuracy and stability. Further analysis confirms that the proposed morphological optimization preserves reasonable tree shapes, leading to more accurate individual tree attribute calculations. This study verifies that the proposed models significantly enhance the accuracy of urban individual tree segmentation (quantity and morphology), enabling decimeter-level urban carbon stock assessment, and provides a novel technical framework for precise ecosystem carbon sequestration monitoring. Xuming Ge, Min Chen 0015, Han Hu 0005, Bo Xu 0003, Qing Zhu 0012 |
IEEE Trans. Geosci. Remote. Sens. | 4 |
| 2025 | Asymmetric Mamba-CNN Collaborative Architecture for Large-Size Remote Sensing Image Semantic SegmentationabstractLarge-size remote sensing images contain rich geographical information. Efficient and accurate semantic segmentation of these images is of significant importance in various fields. However, the massive memory requirements have hindered the development of semantic segmentation methods for large-size remote sensing images. Most existing methods struggle to balance memory usage, global modeling, and local representation accuracy. To address these issues, we propose a new semantic segmentation method for large-size remote sensing images, Mamba–CNN parallel network (MCPNet), which demonstrates impressive performance. The method is an asymmetric Mamba–convolutional neural network (CNN) hybrid architecture. Given the linear modeling complexity of Mamba, we construct the M-branch based on the visual state space (VSS) model, which processes downsampled images to reduce memory consumption while alleviating Mamba’s local forgetting problem. To further enhance the model’s capability in fine-grained detail extraction, we meticulously design a detail-preserving network (DPN) as the C-branch. This branch employs a split downsampling strategy and multiscale convolutional kernel groups to process large-size images, ensuring the preservation of spatial positional relationships while capturing fine-grained local details. Moreover, to effectively filter redundant information introduced by large-size images and bridge the semantic gap between the features extracted by CNN and Mamba, we propose a multigated feature fusion module (MG-FFM). This module progressively refines heterogeneous feature alignment through a bottom-up hierarchical refinement strategy, achieving a progressive fusion of semantics and details. Our method achieves state-of-the-art (SOTA) performance in terms of mean intersection over union (mIoU) and mF1 score on the self-constructed Yaan UAV dataset and two widely used public datasets (DeepGlobe and Inria Aerial) while consuming less GPU memory. The codes will be available athttps://github.com/fsqy-zhang/MCPNet Min Chen 0015, Lianlei Shan, Caiyi Li, Han Hu 0005, Xuming Ge, Qing Zhu 0012, Bo Xu 0003 |
IEEE Trans. Geosci. Remote. Sens. | 9 |
| 2024 | Semantic Image Translation for Repairing the Texture Defects of Building ModelsabstractThe accurate representation of 3-D building models in urban environments is significantly hindered by challenges such as texture occlusion, blurring, and missing details, which are difficult to mitigate through standard photogrammetric texture mapping pipelines. Current image completion methods often struggle to produce structured results and effectively handle the intricate nature of highly structured façade textures with diverse architectural styles. Furthermore, existing image synthesis methods encounter difficulties in preserving high-frequency details and artificial regular structures, which are essential for achieving realistic façade texture synthesis. To address these challenges, we introduce a novel approach for synthesizing façade texture images that authentically reflect the architectural style from a structured label map, guided by a ground-truth façade image. In order to preserve fine details and regular structures, we propose a regularity-aware multidomain method that capitalizes on frequency information and corner maps. We also incorporate semantic region-adaptive normalization (SEAN) blocks into our generator to enable versatile style transfer. To generate plausible structured images without undesirable regions, we employ image completion techniques to remove occlusions according to semantics prior to image inference. Our proposed method is also capable of synthesizing texture images with specific styles for façades that lack preexisting textures, using manually annotated labels. Experimental results on publicly available façade image and 3-D model datasets demonstrate that our method yields superior results and effectively addresses issues associated with flawed textures. Qisen Shang, Han Hu 0005, Haojia Yu, Bo Xu 0003, Libin Wang 0005, Qing Zhu 0012 |
IEEE Trans. Geosci. Remote. Sens. | 4 |
| 2023 | 3-D Line Segment Reconstruction With Depth Maps for Photogrammetric Mesh Refinement in Man-Made EnvironmentsabstractThree-dimensional (3D) line segments contain richer geometric and structural information than 3D point clouds in man-made environments, which is beneficial for providing constraints to refine point-cloud-based mesh models or build accurate wireframes. However, the efficient reconstruction of 3D line segments with high scene coverage from multi-view images is still challenging. In this study, the depth maps obtained from the point cloud generation procedure are exploited to decrease the search range of two-dimensional (2D) line segment correspondences to improve the efficiency, precision, and recall rate of 2D line segment matching, thereby improving the construction efficiency and scene coverage of 3D line segments. For a line segment on the reference image (called reference line segment) of an image pair, a reliable virtual line segment is produced by projecting several sampled points of the reference line segment onto the search image based on the corresponding depth information. Then, a purely geometrical similarity measurement under the constraints of the virtual line segment is designed to obtain 2D line segment matches. Using the 2D line segment correspondences of all image pairs, a multi-view clustering operation is performed to construct 3D line segments from the redundant 2D matches. Finally, a simple 3D-line-segment-based mesh model refinement method is designed and the reconstructed 3D line segments are employed to improve the quality of the point-cloud-based mesh model. In our experiments, five open-source datasets are adopted to qualitatively and quantitatively evaluate the performance of the proposed 3D line segment reconstruction method and the potential of 3D line segments on mesh model refinement. The experimental results show that the proposed 3D line segment reconstruction method performs better than the state-of-the-art methods. Specifically, on five open-source datasets, our method exhibits an average improvement of 47.28% in the number of reconstructed 3D line segments over the best one among the compared methods. Additionally, the experimental results of the mesh model refinement show that the addition of 3D line segments is beneficial for improving the quality of the point-cloud-based mesh model. Tong Fang, Min Chen 0015, Han Hu 0005, Wen Li 0033, Xuming Ge, Qing Zhu 0012, Bo Xu 0003 |
IEEE Trans. Geosci. Remote. Sens. | 7 |
| 2021 | Multientity Registration of Point Clouds for Dynamic Objects on Complex Floating Platform Using Object SilhouettesabstractThis article is focused on a challenging topic emerging from the registration of point clouds, specifically the registration of dynamic objects with low overlapping ratio. This problem is especially difficult when the static scanner is installed on a floating platform, and the objects it scans are also floating. These issues make most of the automatic registration methods and software solutions invalid. To solve this problem, explicit exploration of the static region is necessary for both the coarse and fine registration steps. Fortunately, determining the corresponding regions can be eased by the intuitive realization that in urban environments, natural objects neither present straight boundaries nor stack vertically. This intuition has guided the authors to develop a robust approach for the detection of static regions using planar structures. Then, silhouettes of the objects are extracted from the planar structures, which assist in the determination of an SE(2) transformation in the horizontal direction by a novel line matching method. The silhouettes also enable identification of the correspondences of planes in the step of fine registration using a variant of the iterative closest point method. Experimental evaluations using point clouds of cargo ships with different sizes and shapes reveal the robustness and efficiency of the proposed method, which gives 100% success and reasonable accuracy in rapid time, suitable for an online system. In addition, the proposed method is evaluated systematically with regard to several practical situations caused by the floating platform, and it demonstrates good robustness to limited scanning time and noise. Feng Wang 0044, Han Hu 0005, Xuming Ge, Bo Xu 0003, Ruofei Zhong, Yulin Ding, Xiao Xie, Qing Zhu 0012 |
IEEE Trans. Geosci. Remote. Sens. | 4 |