Yongcan Yu

dblp:302/9350 · DBLP profile ↗
← Back
10ranked-venue papers
3as first author
10since 2021 · last 2026
—ORCID · conflict

Domains — the database's venue-derived domains; a paper can count in several

Applied, interdisciplinary, general and emerging computing · 6 · 1 first-author · 6 since 2021Artificial intelligence and machine learning · 3 · 2 first-author · 3 since 2021Graphics, computer vision, multimedia, augmented reality and games · 2 · 1 first-author · 2 since 2021Computer networks · 1 · 1 since 2021
YearPublicationVenuePosition
2026 Enhanced Multivariate Fusion TSNN-Based Dynamic Integration Error Detection for Autonomous Platforms' Multibeam-Motion Sensor Systems in Internet of Underwater Things
abstract
A profound understanding of unique underwater terrain characteristics is crucial for the Internet of Underwater Things (IoUT). Currently, using unmanned platforms equipped with Multibeam Echosounder Systems (MBES) for underwater topographic scanning has become a key method and future trend to understand seabed topography. However, non-ideal integration between MBES transducers and motion sensors induces dynamic integration errors (time delay, motion scale, yaw misalignment, lever arm errors), which manifest as high-frequency ship-track orthogonal bathymetric undulations, severely limiting the accurate generation of high-resolution seabed topographic maps. To address this, this study innovatively proposes a multibeam dynamic error detection method based on time-series neural networks (TSNN). Its core innovations are: first, leveraging the characteristic that dynamic errors’ ship-track orthogonal bathymetric undulations are highly correlated with attitude; second, focusing on distinct manifestations of beam-angle-dependent and beam-angle-independent dynamic errors across different beam reception angles, laying a key basis for accurate error inversion. The method involves two core modules: the "MBES Bathymetric Sequence Trend Extraction Module" separates seabed terrain trends from undulations to obtain error-driven fluctuation signals; the "Dynamic Error Regression Module" realizes quantitative prediction of dynamic integration errors. Experimental results show the method can effectively eliminate high-frequency undulations in shallow-to-medium water MBES data, with processing efficiency meeting unmanned platforms’ real-time requirements—providing a new technical approach for high-resolution marine surveying and mapping to support IoUT’s reliable deployment and operation.
Jiawei Long, Jianhu Zhao, Yongcan Yu
IEEE Internet Things J.5
2025 Cooperative Pseudo Labeling for Unsupervised Federated Classification
Kuangpu Guo, Lijun Sheng, Yongcan Yu, Jian Liang 0001, Zilei Wang, Ran He 0001
ICCV3
2024 STAMP: Outlier-Aware Test-Time Adaptation with Stable Memory Replay
Yongcan Yu, Lijun Sheng, Ran He 0001, Jian Liang 0001
ECCV (81)1
2024 Seg2Sonar: A Full-Class Sample Synthesis Method Applied to Underwater Sonar Image Target Detection, Recognition, and Segmentation Tasks
abstract
To overcome the challenges of limited samples, difficult acquisition, under-representation, and labeling in utilizing sonar images and deep learning for target detection, recognition, and segmentation tasks for full-class underwater targets, we propose the Seg2Sonar network based on SPADE. This network generates images through segmentation maps, thus eliminating the need for sample annotation. Additionally, we incorporate the Skip-Layer channel-wise Excitation (SLE) module into the SPADE network to enhance feature extraction ability with minimal training samples. To improve the realism of generated images, we introduce the Focal Frequency Loss (FFL) module, and propose the Elasticity loss (EL) strategy to improve the random combination capability of the network, considering the characteristics of low resolution and severe distortion of sonar images. Furthermore, we propose a weight adjustment (WA) strategy that tackles the challenge of low and unbalanced feature representation with few samples by taking into account the unbalanced distribution of features using prior information. hese four improvements enable efficient sample augmentation of sonar images with limited samples. Building upon the improved Seg2Sonar network, we propose an underwater full-class target augmentation strategy. Based on the imaging characteristics of sonar images, we classify underwater full-class targets into four categories: texture level, group level, shape level, and intensity level. We provide corresponding augmentation strategies by leveraging similar features among sonar target images or adding external radar/optical features to supplement the diversity of features. Our experimental results demonstrate the efficacy of our proposed method in achieving sample augmentation of underwater full-class targets with minimal samples (less than 10) or even zero samples. The approach achieves about 90% accuracy in detection, recognition, and segmentation for all types of targets through deep learning methods. Our findings provide a promising solution for efficient sample augmentation of underwater full-class targets with limited samples.
Chao Huang 0023, Jianhu Zhao, Hongmei Zhang 0002, Yongcan Yu
IEEE Trans. Geosci. Remote. Sens.4
2024 Unsupervised Terrain Reconstruction From Side-Scan Sonar Constrained to the Imaging Mechanism
abstract
To meet the demands of marine scientific research and ocean engineering construction for cost-effective, high-resolution, and large-scale acquisition of seafloor topography, this study proposes an unsupervised terrain reconstruction technique for side-scan sonar (SSS). Unlike conventional methods, this technique does not rely on externally measured depths for initial terrain and ground truth; instead, it constructs these elements based on its own imaging mechanism. Through the development of a rigorous seafloor reflection model, we have designed an unsupervised depth convolutional neural network (DCNN) inversion framework. This framework utilizes the depth of the sea bottom line as the initial terrain and employs the height calculated from the target’s shadow as the ground truth. The proposed network not only extracts absolute topographical information but also captures relative changes in beam pattern and seabed substrate properties. The experimental results demonstrate the method’s excellent accuracy in terrain inversion for areas with water depths ranging from 10 to 14 m, achieving a vertical accuracy exceeding 0.17 m. Moreover, the method exhibits robust resilience to residual radiometric distortions, variations in seabed substrate, noise, and shadows. This innovative approach provides a reliable methodology for acquiring and broadly applying high-resolution seafloor terrain, potentially advancing various fields of marine research and engineering.
Chao Huang 0023, Hongmei Zhang 0002, Jianhu Zhao, Yongcan Yu
IEEE Trans. Geosci. Remote. Sens.4
2024 A Sample Augmentation Method for Side-Scan Sonar Full-Class Images That Can Be Used for Detection and Segmentation
abstract
In order to solve the problems of small samples, acquisition difficulties, under-representation and labeling difficulties in object detection, recognition and segmentation tasks for underwater all-category targets based on sonar images and deep learning methods. we propose a side-scan sonar full-class image sample augmentation method suitable for multi-task scenarios. Based on the superior image generation ability of the diffusion model, we use transfer learning to fine-tune the optical pre-trained model to build a side-scan sonar image generation model. Then, for the object detection task and semantic segmentation task, we use the image content and target shape as guidance information to guide the generation results of the diffusion model respectively. Meanwhile, proposed a mask synthesis method for SSS waterfall image generation based on the working principle of side-scan sonar. The synthesized mask images are used to guide the generation of side-scan sonar waterfall images. Finally, the underwater object detection and segmentation models are trained on the generated data. The experiment results show that training a model with generated data can be effective in improving accuracy.
Jianhu Zhao, Yongcan Yu, Chao Huang 0023
IEEE Trans. Geosci. Remote. Sens.3
2023 Drill-Rep: Repetition counting for automatic shot hole depth recognition based on combined deep learning-based model
Yongcan Yu, Jianhu Zhao, Changhua Yi, Chao Huang 0023, Weiqiang Zhu
Eng. Appl. Artif. Intell.1
2023 Treat Noise as Domain Shift: Noise Feature Disentanglement for Underwater Perception and Maritime Surveys in Side-Scan Sonar Images
abstract
In underwater perception and maritime surveys, due to the scarcity of training data and perturbation of speckle noise, the detection performance of underwater objects in side-scan sonar (SSS) images is limited. To address these problems, we proposed a noise feature disentanglement YOLO (NFD-YOLO) by combining noise-agnostic features learning and attention mechanism. Firstly, we rethink the speckle noise by treating it as the domain shift between the training dataset and real-measured SSS images and build a domain generalization-based (DG-based) underwater object detection framework. Then, we extend YOLOv5 with a feature manipulation module, a noise-agnostic subnetwork, and an auxiliary noise-biased subnetwork for noise features disentanglement, more biases toward noise-agnostic features and less reliance on noise-biased features in underwater object detection, respectively. Finally, the ACmix attention module is introduced for a more powerful learning capacity and attention to the object areas based on a small dataset. According to the experiment results, the proposed NFD-YOLO achieved 75.1% mean average precision (mAP) in the test domain, which increased by 7.5% than YOLOv5, and 75.7% ± 0.4% mAP and 77.5% ± 1.6% mAP for different speckle noise distributions and transfer directions, respectively, which verified its generalization ability and robustness for speckle noise. Therefore, the proposed method can mitigate the effects of speckle noise and provides a new thought to address the speckle noise in underwater object detection with a small dataset, which is of significance and benefits for underwater perception and maritime surveys.
Yongcan Yu, Jianhu Zhao, Chao Huang 0023
IEEE Trans. Geosci. Remote. Sens.1
2022 Comprehensive Sample Augmentation by Fully Considering SSS Imaging Mechanism and Environment for Shipwreck Detection Under Zero Real Samples
abstract
To solve the shortage of training samples when using deep learning to detect shipwrecks, a comprehensive sample augmentation method is proposed. The method fully considers the imaging mechanism and environment of side-scan sonar (SSS), such as acoustic emission and reception, waterbody, target reflection, and seafloor background, and generates diverse and representative shipwreck samples from five aspects of target diversity, target texture, imaging resolution, equipment and environmental noises, and background through a series of novel sample augmentation methods. Under the condition of zero real SSS samples, a detection model of YOLOv5s was established with these amplified samples and achieved a mean average precision (MAP) better than 96% for real SSS data detection.
Chao Huang 0023, Jianhu Zhao, Yongcan Yu, Hongmei Zhang 0002
IEEE Trans. Geosci. Remote. Sens.3
2022 Anisotropic Total Variation Regularized Low-Rank Approximation for SSS Images Radiometric Distortion Correction
abstract
Radiometric distortion caused by the time-varying gain (TVG), beam patterns, angular responses, and sonar altitude variations, highly degrades the quality of side-scan sonar (SSS) images. Thus, radiometric distortion correction becomes a fundamental step for SSS image processing which holds vital importance for geomorphic applications. However, existing methods cannot take the prior information of the acoustic illumination component as well as the feature of seafloor into consideration well, which would easily cause damage to the image and also always be powerless for residual stripe noise. In this paper, a novel radiometric correction method is proposed. First, we give a detailed analysis of the SSS imaging theory based on the Lambert’s law as well as the prior knowledge about the characteristics of SSS images. Then, incorporating the prior of the SSS imaging process, the low-rank constraint is specifically introduced for the illumination component, while the anisotropic total variation (ATV) constraint is used to constraint the albedo component, combining other constraints, a decomposition model is proposed to correct the radiometric distortion based on the SSS imaging theory. And an alternative minimization method has been adopted to solve the proposed model effectively. Experiments proved the validity of the proposed method.
Shaobo Li 0002, Jianhu Zhao, Yongcan Yu, Yunlong Wu 0001, Guojun Zhai
IEEE Trans. Geosci. Remote. Sens.3