VLDB 2026 Research / reviewers in the wild / expert
Zhuolin Hao
dblp:289/8268
· DBLP profile ↗
3ranked-venue papers
0as first author
3since 2021 · last 2026
—ORCID · conflict
Domains — the database's venue-derived domains; a paper can count in several
Databases, data management, data science and information retrieval · 3 · 3 since 2021Artificial intelligence and machine learning · 1 · 1 since 2021Applied, interdisciplinary, general and emerging computing · 1 · 1 since 2021
| Year | Publication | Venue | Position |
|---|---|---|---|
| 2026 | When Rules Fall Short: Agent-Driven Discovery of Emerging Content Issues in Short Video Platforms
ChengHui Yu, Hongwei Wang 0004, Junwen Chen 0005, Zixuan Wang 0019, Bingfeng Deng, Zhuolin Hao, Hongyu Xiong, Yang Song 0008 |
WWW | 6 |
| 2025 | Audio-Enhanced Vision-Language Modeling with Latent Space Broadening for High Quality Data ExpansionabstractTransformer-based multimodal models are widely used in industrialscale recommendation, search, and advertising systems for content understanding and relevance ranking.Enhancing labeled training data quality and cross-modal fusion significantly improves model performance, influencing key metrics such as quality view rates and ad revenue.High-quality annotations are crucial for advancing content modeling, yet traditional statistical-based active learning (AL) methods face limitations: they struggle to detect overconfident misclassifications and are less effective in distinguishing semantically similar items in deep neural networks.Additionally, audio information plays an increasing role, especially in short-video platforms, yet most pretrained multimodal architectures primarily focus on text and images.While training from scratch across all three modalities is possible, it sacrifices the benefits of leveraging existing pretrained visual-language (VL) and audio models.To address these challenges, we propose kNN-based Latent Space Broadening (LSB) to enhance AL efficiency, achieving an up to 9% recall improvement at 80% precision on proprietary datasets.Additionally, we introduce Vision-Language Modeling with Audio Enhancement (VLMAE), a mid-fusion approach integrating audio into VL models, yielding up * Author corresponded for this research. Yu Sun 0088, Ruixiao Sun, Chunhui Liu 0002, Fangming Zhou, Ze Jin, Xiang Shen 0001, Zhuolin Hao, Hongyu Xiong |
KDD (2) | 9 |
| 2021 | Generating Personalized Titles Incorporating Advertisement Profile
Jingbing Wang, Zhuolin Hao, Minping Zhou, Jiaze Chen, Zhenqiao Song, Jiandong Yang, Shiguang Ni |
DASFAA (3) | 2 |