VLDB 2026 Research / reviewers in the wild / expert
Chaobo Sun
dblp:150/4566
· DBLP profile ↗
5ranked-venue papers
1as first author
4since 2021 · last 2026
0009-0007-3150-0656ORCID · corroborated
Domains — the database's venue-derived domains; a paper can count in several
Artificial intelligence and machine learning · 3 · 1 first-author · 2 since 2021Databases, data management, data science and information retrieval · 3 · 3 since 2021Graphics, computer vision, multimedia, augmented reality and games · 2 · 1 first-author · 1 since 2021
| Year | Publication | Venue | Position |
|---|---|---|---|
| 2026 | LexInstructEval: Lexical Instruction Following Evaluation for Large Language ModelsabstractThe ability of Large Language Models (LLMs) to precisely follow complex and fine-grained lexical instructions is a cornerstone of their utility and controllability. However, evaluating this capability remains a significant challenge. Current methods either rely on subjective and costly human evaluation or on automated ``LLM-as-a-judge'' systems, which suffer from inherent biases and unreliability. Existing programmatic benchmarks, while objective, often lack the expressiveness to test intricate, compositional constraints at a granular level. To address these limitations, we introduce LexInstructEval, a new benchmark and evaluation framework for fine-grained lexical instruction following. Our framework is built upon a formal, rule-based grammar that deconstructs complex instructions into a canonical (Procedure, Relation, Value) triplet. This grammar enables the systematic generation of a diverse dataset through a multi-stage, human-in-the-loop pipeline and facilitates objective verification via a transparent, programmatic engine. We release our dataset and open-source evaluation tools to facilitate further research into the controllability and reliability of LLMs. Baiqiao Su, Chaobo Sun, Hengtong Lu, Kaike Zhang |
AAAI | 4 |
| 2023 | Dialog-to-Actions: Building Task-Oriented Dialogue System via Action-Level GenerationabstractEnd-to-end generation-based approaches have been investigated and applied in task-oriented dialogue systems. However, in industrial scenarios, existing methods face the bottlenecks of reliability (e.g., domain-inconsistent responses, repetition problem, etc) and efficiency (e.g., long computation time, etc). In this paper, we propose a task-oriented dialogue system via action-level generation. Specifically, we first construct dialogue actions from large-scale dialogues and represent each natural language (NL) response as a sequence of dialogue actions. Further, we train a Sequence-to-Sequence model which takes the dialogue history as the input and outputs a sequence of dialogue actions. The generated dialogue actions are transformed into verbal responses. Experimental results show that our light-weighted method achieves competitive performance, and has the advantage of reliability and efficiency. Yuncheng Hua, Xiangyu Xi, Guanwei Zhang, Chaobo Sun, Guanglu Wan, Wei Ye 0004 |
SIGIR | 5 |
| 2022 | Unified Knowledge Prompt Pre-training for Customer Service DialoguesabstractDialogue bots have been widely applied in customer service scenarios to provide timely and user-friendly experience. These bots must classify the appropriate domain of a dialogue, understand the intent of users, and generate proper responses. Existing dialogue pre-training models are designed only for several dialogue tasks and ignore weakly-supervised expert knowledge in customer service dialogues. In this paper, we propose a novel unified knowledge prompt pre-training framework, UFA (Unified Model F or All Tasks), for customer service dialogues. We formulate all the tasks of customer service dialogues as a unified text-to-text generation task and introduce a knowledge-driven prompt strategy to jointly learn from a mixture of distinct dialogue tasks. We pre-train UFA on a large-scale Chinese customer service corpus collected from practical scenarios and get significant improvements on both natural language understanding (NLU) and natural language generation (NLG) benchmarks. Keqing He 0001, Jingang Wang, Chaobo Sun, Wei Wu 0014 |
CIKM | 3 |
| 2022 | A Low-Cost, Controllable and Interpretable Task-Oriented Chatbot: With Real-World After-Sale Services as ExampleabstractThough widely used in industry, traditional task-oriented dialogue systems suffer from three bottlenecks: (i) difficult ontology construction (e.g., intents and slots); (ii) poor controllability and interpretability; (iii) annotation-hungry. In this paper, we propose to represent utterance with a simpler concept named Dialogue Action, upon which we construct a tree-structured TaskFlow and further build task-oriented chatbot with TaskFlow as core component. A framework is presented to automatically construct TaskFlow from large-scale dialogues and deploy online. Our experiments on real-world after-sale customer services show TaskFlow can satisfy the major needs, as well as reduce the developer burden effectively. Xiangyu Xi, Chenxu Lv, Yuncheng Hua, Wei Ye 0004, Chaobo Sun, Shuaipeng Liu, Fan Yang 0087, Guanglu Wan |
SIGIR | 5 |
| 2014 | Object Ranking on Deformable Part Models with Bagged LambdaMART
Chaobo Sun, Xiaojie Wang 0006, Peng Lu 0007 |
ACCV (2) | 1 |