VLDB 2026 Research / reviewers in the wild / expert
Ran Chen 0002
dblp:95/6235-2
· DBLP profile ↗
4ranked-venue papers
3as first author
4since 2021 · last 2025
0009-0009-9960-2637ORCID · conflict
Domains — the database's venue-derived domains; a paper can count in several
Graphics, computer vision, multimedia, augmented reality and games · 3 · 2 first-author · 3 since 2021Artificial intelligence and machine learning · 2 · 2 first-author · 2 since 2021Applied, interdisciplinary, general and emerging computing · 1 · 1 first-author · 1 since 2021
Expertise — from the expertise taxonomy: the topics of the expert's papers under the CCF categories. A weight counts papers with recency: 1 for a paper about the topic, 0.3 when the topic is its context, halved every five years.
| Computer graphics and multimedia
1 paper |
Geometric modeling and processing · 100% | |
| Software engineering, system software, and programming languages
1 paper |
Program synthesis and code generation · 100% | |
| Interdisciplinary, comprehensive, and emerging computing
1 paper |
Computing education · 100% |
Topics — the 1 heaviest of 3, each with the papers that count most for it
| Topic | Weight | Papers | Last | Evidence papers |
|---|---|---|---|---|
Computing education
problem generation |
0.3 | 1 | 2025 | GeoUni: A Unified Model for Generating Geometry Diagrams, Problems and Problem Solutions · ACM Multimedia 2025 |
Methods — techniques the papers use, named apart from their topics
unified model · 2.6
| Year | Publication | Venue | Position |
|---|---|---|---|
| 2025 | TopSUMseg: A Topology-Aware Swin Transformer-Mamba Framework for 3D Seismic Fault Image SegmentationabstractSeismic fault image segmentation is crucial for interpreting subsurface geological structures, supporting geologists in resource exploration and structural analysis. However, current deep learning models struggle with single-architecture limitations and the distinctive characteristics of seismic faults, which are distinguished by elongated structures with uneven spatial distributions. To address these challenges, we propose TopSUMseg, a novel Topology-Aware Swin Transformer-Mamba framework for 3D seismic fault image segmentation. Our framework combines Swin Transformer’s local feature extraction with Mamba’s efficient sequence modeling, and boosts 3D spatial modeling in Mamba with a newly designed Global-Local Attention module (GLA). Additionally, we design a Topology-Aware Structural Constraint (TASC) to align predictions with ground-truth structures in the feature space, promoting the modeling of complex fault geometries. Experiments on Thebe, the largest public seismic dataset, demonstrate that TopSUMseg achieves state-of-the-art performance with OIS and ODS scores of 0.879 and 0.875, respectively. Trained entirely from scratch, TopSUMseg nonetheless achieves superior performance compared to extensively pre-trained counterparts. In addition, TopSUMseg maintains a significantly lower parameter count while achieving a favorable trade-off between segmentation performance and time complexity, making it a practical and generalizable solution for real-world seismic fault interpretation. Ran Chen 0002, Jingyang Deng, Zeren Zhang, Ruohua Shi, Jinwen Ma |
ECAI | 1 |
| 2025 | Reframing Multimodal Complex Document Layout Understanding: A Layout-Aware Multi-Source Reasoning Decision FrameworkabstractMultimodal large language models (MLLMs) have achieved significant progress in document understanding. However, complex layout reasoning, characterized by concise answers and cross-page integration, remains a challenge. Unlike conventional semantics-oriented tasks, this task demands accurate visual perception of fine-grained structural elements and logical reasoning across multi-page documents. Existing approaches primarily focus on information extraction and semantic understanding, limiting the capacity of fine-tuned autoregressive models to capture short-answer reasoning signals and generalize to complex layout structures. To address this, we propose the Layout-Aware Multi-Source Reasoning Decision Framework (LAMRD), which reframes complex layout reasoning as a decision-making task over multi-source reasoning paths. In the reasoning path construction stage, LAMRD generates layout-aware reasoning paths by integrating internal visual cues and external knowledge from three complementary perspectives: Visual Structural Awareness (VSA), Logical Reasoning Paths (LRP), and External Knowledge Augmentation (EKA). In the reasoning path decision stage, we employ Group Relative Policy Optimization (GRPO) to train a decision model that produces the final answer based on these paths. We conduct comprehensive evaluations using Qwen2.5-VL-7B-Instruct on the CEP-7K dataset, covering layout structure understanding, information extraction, and logical association. Experimental results demonstrate that LAMRD outperforms advanced MLLMs in accuracy, validating its effectiveness for complex document layout understanding. Ran Chen 0002, Jingyang Deng, Zeren Zhang, Xuefei Tong, Jinwen Ma, Qinghui Shi, Yuanjun Li |
ECAI | 1 |
| 2025 | GeoUni: A Unified Model for Generating Geometry Diagrams, Problems and Problem Solutions
Jo-Ku Cheng, Zeren Zhang, Ran Chen 0002, Jingyang Deng, Ziran Qin, Jinwen Ma |
ACM Multimedia | 3 |
| 2024 | A Fusion Framework of Whitespace Smear Cutting and Swin Transformer for Document Layout Analysis
Ran Chen 0002, Jo-Ku Cheng, Jinwen Ma |
ICIC (6) | 1 |