VLDB 2026 Research / reviewers in the wild / expert
Jiang-jiang Liu
dblp:409/9085
· DBLP profile ↗
2ranked-venue papers
0as first author
2since 2021 · last 2025
—ORCID · none
Domains — the database's venue-derived domains; a paper can count in several
Artificial intelligence and machine learning · 2 · 2 since 2021Graphics, computer vision, multimedia, augmented reality and games · 2 · 2 since 2021
Expertise — from the expertise taxonomy: the topics of the expert's papers under the CCF categories. A weight counts papers with recency: 1 for a paper about the topic, 0.3 when the topic is its context, halved every five years.
| Artificial intelligence
2 papers |
3D vision · 57% Vision and language · 19% Image recognition and object detection · 19% |
Topics — the 6 heaviest of 6, each with the papers that count most for it
| Topic | Weight | Papers | Last | Evidence papers |
|---|---|---|---|---|
Computer vision › 3D vision
3d scene understanding |
0.9 | 1 | 2025 | VoxelSplat: Dynamic Gaussian Splatting as an Effective Loss for Occupancy and Flow Prediction · CVPR 2025 |
Computer vision › Image recognition and object detection › object detection
object proposal generation |
0.9 | 1 | 2025 | PropVG: End-To-End Proposal-Driven Visual Grounding with Multi-Granularity Discrimination · ICCV 2025 |
Computer vision › 3D vision
scene flow estimation |
0.9 | 1 | 2025 | VoxelSplat: Dynamic Gaussian Splatting as an Effective Loss for Occupancy and Flow Prediction · CVPR 2025 |
Computer vision › 3D vision › 3d scene understanding
semantic scene completion |
0.9 | 1 | 2025 | VoxelSplat: Dynamic Gaussian Splatting as an Effective Loss for Occupancy and Flow Prediction · CVPR 2025 |
Computer vision › Vision and language
visual grounding |
0.9 | 1 | 2025 | PropVG: End-To-End Proposal-Driven Visual Grounding with Multi-Granularity Discrimination · ICCV 2025 |
Computer vision › Segmentation and scene understanding
referring image segmentation |
0.3 | 1 | 2025 | PropVG: End-To-End Proposal-Driven Visual Grounding with Multi-Granularity Discrimination · ICCV 2025 |
Methods — techniques the papers use, named apart from their topics
self-supervised learning · 0.9proposal-based framework · 0.9multi-granularity discrimination · 0.9contrastive learning · 0.93d gaussian splatting · 0.92d projection supervision · 0.9
| Year | Publication | Venue | Position |
|---|---|---|---|
| 2025 | VoxelSplat: Dynamic Gaussian Splatting as an Effective Loss for Occupancy and Flow PredictionabstractRecent advancements in camera-based occupancy prediction have focused on the simultaneous prediction of 3D semantics and scene flow, a task that presents significant challenges due to specific difficulties, e.g., occlusions and unbalanced dynamic environments. In this paper, we analyze these challenges and their underlying causes. To address them, we propose a novel regularization framework called VoxelSplat. This framework leverages recent developments in 3D Gaussian Splatting to enhance model performance in two key ways: (i) Enhanced Semantics Supervision through 2D Projection: During training, our method decodes sparse semantic 3D Gaussians from 3D representations and projects them onto the 2D camera view. This provides additional supervision signals in the camera-visible space, allowing 2D labels to improve the learning of 3D semantics. (ii) Scene Flow Learning: Our framework uses the predicted scene flow to model the motion of Gaussians, and is thus able to learn the scene flow of moving objects in a self-supervised manner using the labels of adjacent frames. Our method can be seamlessly integrated into various existing occupancy models, enhancing performance without increasing inference time. Extensive experiments on benchmark datasets demonstrate the effectiveness of Voxel-Splat in improving the accuracy of both semantic occupancy and scene flow estimation. The project page and codes are available at https://zzy816.github.io/VoxelSplat-Demo/. Ziyue Zhu, Shenlong Wang, Jin Xie 0001, Jiang-jiang Liu, Jingdong Wang 0001, Jian Yang 0003 |
CVPR | 4 |
| 2025 | PropVG: End-To-End Proposal-Driven Visual Grounding with Multi-Granularity DiscriminationabstractRecent advances in visual grounding have largely shifted away from traditional proposal-based two-stage frameworks due to their inefficiency and high computational complexity, favoring end-to-end direct reference paradigms. However, these methods rely exclusively on the referred target for supervision, overlooking the potential benefits of prominent prospective targets. Moreover, existing approaches often fail to incorporate multi-granularity discrimination, which is crucial for robust object identification in complex scenarios. To address these limitations, we propose PropVG, an end-to-end proposal-based framework that, to the best of our knowledge, is the first to seamlessly integrate foreground object proposal generation with referential object comprehension without requiring additional detectors. Furthermore, we introduce a Contrastive-based Refer Scoring (CRS) module, which employs contrastive learning at both sentence and word levels to enhance the capability in understanding and distinguishing referred objects. Additionally, we design a Multi-granularity Target Discrimination (MTD) module that fuses object- and semantic-level information to improve the recognition of absent targets. Extensive experiments on gRefCOCO (GREC/GRES), Ref-ZOM, R-RefCOCO, and RefCOCO (REC/RES) benchmarks demonstrate the effectiveness of PropVG. The codes and models are available at https://github.com/Dmmm1997/PropVG. Wenxuan Cheng, Jiedong Zhuang, Jiang-jiang Liu, Hongshen Zhao, Zhenhua Feng 0001, Wankou Yang |
ICCV | 4 |