VLDB 2026 Research / reviewers in the wild / expert
Zhenhan Guan
dblp:374/7766
· DBLP profile ↗
2ranked-venue papers
0as first author
2since 2021 · last 2024
0009-0005-5364-7961ORCID · corroborated
Domains — the database's venue-derived domains; a paper can count in several
Software engineering, systems software and programming languages · 1 · 1 since 2021Databases, data management, data science and information retrieval · 1 · 1 since 2021
Expertise — from the expertise taxonomy: the topics of the expert's papers under the CCF categories. A weight counts papers with recency: 1 for a paper about the topic, 0.3 when the topic is its context, halved every five years.
| Software engineering, system software, and programming languages
1 paper |
Program synthesis and code generation · 100% |
Topics — the 2 heaviest of 2, each with the papers that count most for it
| Topic | Weight | Papers | Last | Evidence papers |
|---|---|---|---|---|
Program synthesis and code generation
code generation with language models |
0.8 | 1 | 2024 | Preference-Guided Refactored Tuning for Retrieval Augmented Code Generation · ASE 2024 |
Program synthesis and code generation › code generation with language models
retrieval-augmented code generation |
0.8 | 1 | 2024 | Preference-Guided Refactored Tuning for Retrieval Augmented Code Generation · ASE 2024 |
Methods — techniques the papers use, named apart from their topics
retrieval augmentation · 0.8preference optimization · 0.8large language model · 0.8
| Year | Publication | Venue | Position |
|---|---|---|---|
| 2024 | Preference-Guided Refactored Tuning for Retrieval Augmented Code GenerationabstractRetrieval-augmented code generation utilizes Large Language Models as the generator and significantly expands their code generation capabilities by providing relevant code, documentation, and more via the retriever. The current approach suffers from two primary limitations: 1) information redundancy. The indiscriminate inclusion of redundant information can result in resource wastage and may misguide generators, affecting their effectiveness and efficiency. 2) preference gap. Due to different optimization objectives, the retriever strives to procure code with higher ground truth similarity, yet this effort does not substantially benefit the generator. The retriever and the generator may prefer different golden code, and this gap in preference results in a suboptimal design. Additionally, differences in parameterization knowledge acquired during pre-training result in varying preferences among different generators. Yun Xiong, Deze Wang, Zhenhan Guan, Zejian Shi, Haofen Wang, Shanshan Li 0001 |
ASE | 4 |
| 2024 | Exploring on role of location in intelligent news recommendation from data analysis perspectiveabstractLocation factor of recommender systems has been extensively studied in the past decade. However, there is no research thoroughly analyzing location’s role in news recommendation. In this paper, a comprehensive exploration on role of location in news recommendation is presented. First of all, based on analysis of real news datasets, we find that news recommendation differs from spatial item recommendation. Location affects news consumption behaviors of users with two-fold aspects including geographic feature and semantic feature. Regarding geographic feature, location influences news recommendation according to region rather than latitude-longitude level. Furthermore, interesting news topics are also impacted by semantic feature of location. Semantic feature may play a more positive role than geographic feature. The novel findings consistently manifest that, as non-spatial items, news differ from spatial items in that location influences users' selection in terms of different pattern and degree. In summary, geographic and semantic features influence reading preference through mapping locations into special topics. Changing of location topics leads to varying of reading preference. The news datasets in this paper belong to check in data. NewsREEL dataset is from a company, and it is provided by German researcher. The location data in Twitter dataset is also check in data. NetEase news dataset are collected from NetEase news websites, and the type of location data is city or region. Pengtao Lv, Lei Shi 0030, Zhenhan Guan, Yanfeng Fan, Kaiyang Zhong, Muhammet Deveci |
Inf. Sci. | 4 |