VLDB 2026 Research / reviewers in the wild / expert
Bin Ge 0001
dblp:79/3482-1
· DBLP profile ↗
34ranked-venue papers
9as first author
32since 2021 · last 2026
0000-0001-9050-1105ORCID · conflict
Domains — the database's venue-derived domains; a paper can count in several
Graphics, computer vision, multimedia, augmented reality and games · 17 · 7 first-author · 15 since 2021Artificial intelligence and machine learning · 11 · 1 first-author · 11 since 2021Systems, architecture and hardware · 3 · 3 since 2021Computer networks · 2 · 2 since 2021Security and privacy · 1 · 1 first-author · 1 since 2021
| Year | Publication | Venue | Position |
|---|---|---|---|
| 2026 | Arbitrary style transfer with sparse and dense semantic adaptive attention network
Bin Ge 0001, Chenxing Xia, Junshuai Zheng |
Multim. Syst. | 1 |
| 2026 | SharpDepth: self-supervised monocular depth estimation through edge awareness and wavelet frequency domain fusion
Bin Ge 0001, Chenxing Xia, Mengyan Cheng |
Multim. Syst. | 1 |
| 2026 | Leap-mamba: locality-enhanced feature calibration with pixel-region dual-stream vision mamba UNet for medical image segmentation
Bin Ge 0001, Caihan Yue, Chenxing Xia, Junshuai Zheng |
Multim. Syst. | 1 |
| 2026 | Enhancing underwater debris detection: a dimension-aware diffusion and progressive feature enhancement approach
Yongjie Yu, Hui Chen 0030, Shunxiang Zhang, Bin Ge 0001 |
J. Supercomput. | 4 |
| 2025 | DCSFNet: Deeply Coupled Spatial-Frequency Interaction Network for Polyp Segmentation
Chaofan Liu, Chenxing Xia, Bin Ge 0001 |
PRCV (14) | 3 |
| 2025 | MFCL: Multi-feature Contrastive Learning for Deepfake Detection
Junshuai Zheng, Bin Ge 0001, Chenxing Xia, Qing-Ling Yang, Xianjin Fang |
PRCV (2) | 2 |
| 2025 | CSGN:CLIP-driven semantic guidance network for Clothes-Changing Person Re-Identification
Bin Ge 0001, Chenxing Xia, Junming Guan |
Comput. Vis. Image Underst. | 2 |
| 2025 | Multiangle feature fusion network for style transfer
Zhenshan Hu, Bin Ge 0001, Chenxing Xia |
Image Vis. Comput. | 2 |
| 2025 | Camouflaged object detection with integrated feature fusion and boundary optimization
Bin Ge 0001, Xiaolong Peng, Chenxing Xia |
Multim. Syst. | 1 |
| 2025 | Deep hashing with prototype-aware hardness-weighted for supervised cross-modal retrieval
Bin Ge 0001, Mengyan Cheng, Chenxing Xia |
Pattern Anal. Appl. | 1 |
| 2024 | Flow style-aware network for arbitrary style transferabstractResearchers have recently proposed arbitrary style transfer methods based on various model frameworks. Although all of them have achieved good results, they still face the problems of insufficient stylization, artifacts and inadequate retention of content structure. In order to solve these problems, we propose a flow style-aware network (FSANet) for arbitrary style transfer, which combines a VGG network and a flow network. FSANet consists of a flow style transfer module (FSTM), a dynamic regulation attention module (DRAM), and a style feature interaction module (SFIM). The flow style transfer module uses the reversible residue block features of the flow network to create a sample feature containing the target content and style. To adapt the FSTM to VGG networks, we design the dynamic regulation attention module and exploit the sample features both at the channel and pixel levels. The style feature interaction module computes a style tensor that optimizes the fused features. Extensive qualitative and quantitative experiments demonstrate that our proposed FSANet can effectively avoid artifacts and enhance the preservation of content details while migrating style features. Zhenshan Hu, Bin Ge 0001, Chenxing Xia, Wenyan Wu 0008, Guangao Zhou, Baotong Wang |
Comput. Graph. | 2 |
| 2024 | RCNet: Related Context-Driven Network with Hierarchical Attention for Salient Object Detection
Chenxing Xia, Kuanching Li, Bin Ge 0001, Hanling Zhang |
Expert Syst. Appl. | 4 |
| 2024 | IML-SSOD: Interconnected and multi-layer threshold learning for semi-supervised detection
Bin Ge 0001, Chenxing Xia, Shuaishuai Geng |
J. Vis. Commun. Image Represent. | 1 |
| 2024 | Arbitrary style transfer method with attentional feature distribution matching
Bin Ge 0001, Zhenshan Hu, Chenxing Xia, Junming Guan |
Multim. Syst. | 1 |
| 2024 | Triple fusion and feature pyramid decoder for RGB-D semantic segmentation
Bin Ge 0001, Chenxing Xia |
Multim. Syst. | 1 |
| 2024 | Boundary enhancement and refinement network for camouflaged object detection
Chenxing Xia, Huizhen Cao, Xiuju Gao, Bin Ge 0001, Kuanching Li, Xianjin Fang, Yan Zhang 0106, Xingzhu Liang |
Mach. Vis. Appl. | 4 |
| 2024 | PCDR-DFF: multi-modal 3D object detection based on point cloud diversity representation and dual feature fusion
Chenxing Xia, Xubing Li, Xiuju Gao, Bin Ge 0001, Kuanching Li, Xianjin Fang, Yan Zhang 0106 |
Neural Comput. Appl. | 4 |
| 2024 | PCTDepth: Exploiting Parallel CNNs and Transformer via Dual Attention for Monocular Depth EstimationabstractAbstract Monocular depth estimation (MDE) has made great progress with the development of convolutional neural networks (CNNs). However, these approaches suffer from essential shortsightedness due to the utilization of insufficient feature-based reasoning. To this end, we propose an effective parallel CNNs and Transformer model for MDE via dual attention (PCTDepth). Specifically, we use two stream backbones to extract features, where ResNet and Swin Transformer are utilized to obtain local detail features and global long-range dependencies, respectively. Furthermore, a hierarchical fusion module (HFM) is designed to actively exchange beneficial information for the complementation of each representation during the intermediate fusion. Finally, a dual attention module is incorporated for each fused feature in the decoder stage to improve the accuracy of the model by enhancing inter-channel correlations and focusing on relevant spatial locations. Comprehensive experiments on the KITTI dataset demonstrate that the proposed model consistently outperforms the other state-of-the-art methods. Chenxing Xia, Xiuzhen Duan, Xiuju Gao, Bin Ge 0001, Kuanching Li, Xianjin Fang, Yan Zhang 0106 |
Neural Process. Lett. | 4 |
| 2024 | MFCINet: multi-level feature and context information fusion network for RGB-D salient object detection
Chenxing Xia, Difeng Chen, Xiuju Gao, Bin Ge 0001, Kuanching Li, Xianjin Fang, Yan Zhang 0106 |
J. Supercomput. | 4 |
| 2024 | EDFIDepth: enriched multi-path vision transformer feature interaction networks for monocular depth estimation
Chenxing Xia, Mengge Zhang, Xiuju Gao, Bin Ge 0001, Kuanching Li, Xianjin Fang, Yan Zhang 0106, Xingzhu Liang |
J. Supercomput. | 4 |
| 2023 | Coupled locality discriminant analysis with globality preserving for dimensionality reduction
Shuzhi Su, Bin Ge 0001, Xingzhu Liang |
Appl. Intell. | 4 |
| 2023 | IMSFNet: integrated multi-source feature network for salient object detection
Chenxing Xia, Xianjin Fang, Bin Ge 0001, Xiuju Gao, Kuanching Li |
Appl. Intell. | 4 |
| 2023 | Deep reinforcement learning based adaptive threshold multi-tasks offloading approach in MEC
Liting Mu, Bin Ge 0001, Chenxing Xia, Cai Wu |
Comput. Networks | 2 |
| 2022 | Multi-Modality Diversity Fusion Network with Swintransformer for RGB-D Salient Object DetectionabstractMulti-modality complementary information brings new impetus and innovation to saliency object detection (SOD). However, most existing RGB-D SOD methods either indiscriminately handle RGB features and depth features or only take depth features as additional information of RGB subnet-work, ignoring the different roles of two modalities for SOD tasks. To tackle this issue, we propose a novel multi-modality diversity fusion network with SwinTransformer (M2DFNet) for RGB-D SOD from the perspective of the different status of multi-modality, which adequately explores the roles of RGB and depth modalities. To this end, a triple-diversity supervision mechanism (TDSM) and a diversity fusion module (DFM) are designed to parse the function of different modalities. Besides, we designed a dense decoder (DSD) to integrate multi-scale features and transfer gain information from top to bottom, which can improve the performance of SOD. Extensive experiments on five benchmark datasets demonstrate that the proposed M2DFNet outperforms 17 other state-of-the-art (SOTA) RGB-D SOD methods. Songsong Duan, Chenxing Xia, Xiuju Gao, Bin Ge 0001, Hanling Zhang, Kuanching Li |
ICIP | 4 |
| 2022 | Emcenet: Efficient Multi-Scale Context Exploration Network for Salient Object DetectionabstractMulti-scale context is crucial for the accurate salient object detection (SOD) in the real-world scenes. Although current contextual information-based SOD methods have achieved great progress, they may fail to generate precise saliency maps due to their seldom considering the correlation of different scale context during the extraction process. To address these issues, we propose an Efficient Multi-Scale Context Exploration Network (EMCENet) for SOD. Specifically, a progressive multi-scale context extraction (PMCE) module is designed to progressively capture strongly correlated multi-scale context by using multi-receptive-field convolution operations. Afterwards, a hierarchical feature hybrid interaction (HFHI) module is introduced to generate powerful feature representations by adaptively aggregating multi-level features in a hybrid interaction strategy. Extensive experimental results on six public datasets demonstrate that the proposed EMCENet method without any post-processing performs favorably against 13 state-of-the-art SOD methods. Chenxing Xia, Xiuju Gao, Bin Ge 0001, Hanling Zhang, Kuanching Li |
ICIP | 4 |
| 2022 | DAST: Depth-Aware Assessment and Synthesis Transformer for RGB-D Salient Object Detection
Chenxing Xia, Songsong Duan, Xianjin Fang, Bin Ge 0001, Xiuju Gao, Jianhua Cui |
PRICAI (2) | 4 |
| 2022 | CMNet: Cross-Aggregation Multi-branch Network for Salient Object Detection
Chenxing Xia, Xianjin Fang, Bin Ge 0001, Xiuju Gao, Jianhua Cui |
PRICAI (3) | 4 |
| 2022 | Image Encryption Algorithm based on Convolutional Neural Network and Four Square Matrix EncodingabstractTo balance the performance of chaos and the algorithm’s efficiency, a Henon-Modulated-Iterative (HMI) map is constructed in this paper. And the chaotic characteristics of the HMI map are analyzed from three aspects: bifurcation map, Lyapunov exponent and balance degree. On this basis, an image encryption algorithm based on a convolutional neural network (CNN) and four square matrix encoding is proposed. Then, the random sequence generation module is designed using CNN, and the generated random sequence is used to construct the scrambling coordinate matrix and modify the position of the pixel values. Finally, the scrambled image is diffused by four square matrix encoding so that the pixel values are evenly distributed, while non-sequential diffusion using hexadecimal addition and subtraction rules improves the security of the encryption algorithm. Experimental results show that the encryption algorithm proposed in this paper has a good encryption effect and can resist various attacks. Bin Ge 0001, Chenxing Xia, Gaole Dai |
TrustCom | 1 |
| 2022 | GCENet: Global contextual exploration network for RGB-D salient object detection
Chenxing Xia, Songsong Duan, Xiuju Gao, Rongmei Huang, Bin Ge 0001 |
J. Vis. Commun. Image Represent. | 6 |
| 2022 | HDNet: Multi-Modality Hierarchy-Aware Decision Network for RGB-D Salient Object DetectionabstractRGB-D Salient object detection (SOD) is a pixel-level dense prediction task, which can highlight the prominent object in the scene. Recently, Convolution Neural Network (CNN) is widely applied in SOD to generate multi-level features, which are complementary to each other. However, most methods ignore the unique characteristics of multi-level features (high-level and low-level features). Given the effective employment of multi-level features, we propose a novel multi-modality hierarchy-aware decision network (HDNet) by embedding a Swin Transformer as an encoder. The proposed HDNet contains three primary designs: (1) a Swin Transformer encoder is employed instead of a CNN to learn long-range dependencies; (2) a hierarchy-aware feature decision mechanism (HFDM) is proposed to exploit effective local detail cues of low-level features and global semantic information of high-level features, which consists of two sub-modules, namely low-hierarchy edge module (LEM) and high-hierarchy region module (HRM); (3) a decision-based fusion module (DFM) is designed to fuse RGB and depth features under the attribute of multi-level features generated from HFDM. Experiments on five public benchmarks verify that our framework has better performance than the other 18 state-of-the-art algorithms. Chengxing Xia, Songsong Duan, Bin Ge 0001, Hanling Zhang, Kuanching Li |
IEEE Signal Process. Lett. | 3 |
| 2022 | DMINet: dense multi-scale inference network for salient object detection
Chenxing Xia, Xiuju Gao, Bin Ge 0001, Songsong Duan |
Vis. Comput. | 4 |
| 2021 | Image Encryption for Wireless Sensor Networks with Modified Logistic Map and New Hash Algorithm
Bin Ge 0001, Chenxing Xia |
WASA (3) | 2 |
| 2020 | Clustering adaptive canonical correlations for high-dimensional multi-modal data
Shuzhi Su, Xianjin Fang, Gaoming Yang, Bin Ge 0001 |
J. Vis. Commun. Image Represent. | 4 |
| 2019 | Image encryption based on Henon chaotic system with nonlinear term
Bin Ge 0001 |
Multim. Tools Appl. | 2 |