VLDB 2026 Research / reviewers in the wild / expert
Kedi Shen
dblp:351/8370
· DBLP profile ↗
2ranked-venue papers
0as first author
2since 2021 · last 2025
—ORCID · none
Domains — the database's venue-derived domains; a paper can count in several
Graphics, computer vision, multimedia, augmented reality and games · 2 · 2 since 2021Artificial intelligence and machine learning · 1 · 1 since 2021
Expertise — from the expertise taxonomy: the topics of the expert's papers under the CCF categories. A weight counts papers with recency: 1 for a paper about the topic, 0.3 when the topic is its context, halved every five years.
| Artificial intelligence
1 paper |
Generative modeling · 56% 3D vision · 44% |
Topics — the 3 heaviest of 4, each with the papers that count most for it
| Topic | Weight | Papers | Last | Evidence papers |
|---|---|---|---|---|
Computer vision › 3D vision › 3d reconstruction › object reconstruction
3d reassembly |
0.9 | 1 | 2025 | SE(3)-Equivariant Diffusion Models for 3D Object Analysis · IJCAI 2025 |
Machine learning › Generative modeling › diffusion model › geometric diffusion model
equivariant diffusion model |
0.9 | 1 | 2025 | SE(3)-Equivariant Diffusion Models for 3D Object Analysis · IJCAI 2025 |
Machine learning › Generative modeling
diffusion model |
0.3 | 1 | 2025 | SE(3)-Equivariant Diffusion Models for 3D Object Analysis · IJCAI 2025 |
Methods — techniques the papers use, named apart from their topics
lie algebra mapping · 0.9diffusion model · 0.9SE(3)-equivariant neural networks · 0.9
| Year | Publication | Venue | Position |
|---|---|---|---|
| 2025 | SE(3)-Equivariant Diffusion Models for 3D Object AnalysisabstractSE(3)-equivariance is a critical property for capturing pose information in 3D vision tasks, enabling models to handle transformations such as rotations and translations effectively. While equivariant diffusion models have recently demonstrated promise in 3D object reassembly due to their generative and denoising capabilities, they face key challenges when applied to this task. Specifically, traditional diffusion models rely on fixed input sizes, which limits their adaptability to varying part quantities, and their linear noise addition and removal processes struggle to address the inherently nonlinear transformations of 3D parts. To overcome these limitations, this paper proposes an SE(3)-equivariant diffusion model for pose denoising and 3D object reassembly from fragmented parts. The model incorporates an equivariant encoder to extract SE(3)-equivariant features, a Lie algebra mapping to linearize noise addition and removal, and an elastic diffusion framework capable of adapting to varying part quantities and nonlinear transformations. By leveraging these components, the method achieves accurate and robust pose predictions across diverse input configurations. Experiments conducted on the Breaking Bad dataset, a real-world RePAIR and a self-constructed 3D mannequin dataset demonstrate the effectiveness of the proposed model, outperforming state-of-the-art methods across metrics such as root mean square error and part accuracy. Ablation studies further validate the critical contributions of key modules, emphasizing their roles in improving accuracy and robustness for 3D part reassembly tasks. Kedi Shen, Kangxin Chen |
IJCAI | 3 |
| 2025 | A Novel SO(3) Rotational Equivariant Masked Autoencoder for 3D Mesh Object AnalysisabstractEquivariant networks have recently made significant strides in computer vision tasks related to robotic grasping, molecule generation, and 6D pose tracking. In this paper, we explore 3D mesh object analysis based on an equivariant masked autoencoder to reduce the model dependence on large datasets and predict the pose transformation. We employ 3D reconstruction tasks under rotation and masking operations, such as segmentation tasks after rotation, as pretraining to enhance downstream task performance. To mitigate the computational complexity of the algorithm, we first utilize multiple non-overlapping 3D mesh patches with a fixed face size. We then design a rotation-equivariant self-attention mechanism to obtain advanced features. To improve the throughput of the encoder, we design a sparse token merging strategy. Our method achieves comparable performance on equivariant analysis tasks of mesh objects, such as 3D mesh pose transformation estimation, object classification and part segmentation on the ShapeNetCore16, Manifold40, COSEG-aliens, COSEG-vases and Human Body datasets. In the object classification task, we achieve superior performance even when only 10% of the original sample is used. We perform extensive ablation experiments to demonstrate the efficacy of critical design choices in our approach. Kedi Shen |
IEEE Trans. Circuits Syst. Video Technol. | 3 |