Wenbin Li 0012

dblp:27/1736-12 · DBLP profile ↗
← Back
11ranked-venue papers
3as first author
11since 2021 · last 2026
0000-0002-3258-1116ORCID · conflict

Domains — the database's venue-derived domains; a paper can count in several

Artificial intelligence and machine learning · 8 · 1 first-author · 8 since 2021Databases, data management, data science and information retrieval · 7 · 2 first-author · 7 since 2021Graphics, computer vision, multimedia, augmented reality and games · 3 · 3 since 2021Applied, interdisciplinary, general and emerging computing · 1 · 1 first-author · 1 since 2021
YearPublicationVenuePosition
2026 PORCA: Root Cause Analysis with Partially Observed Data
Chang Gong 0001, Di Yao 0001, Jin Wang 0007, Wenbin Li 0012, Lanting Fang, Yongtao Xie, Kaiyu Feng, Peng Han 0005, Jingping Bi
ICDE4
2026 DPBL: Denoised Player Behavior Representation Learning
abstract
The video game industry has emerged as a significant economic force, driving extensive research on optimizing the gaming environment and improving gaming experiences. Among these endeavors, player behavior representation learning has become a critical way to model valuable player properties and is beneficial for a wide range of downstream tasks. However, some common factors, such as login rewards and daily tasks, can trigger similar behaviors among different players, which are informative and noisy for learning high-quality player behavior representations. Existing methods ignore the low signal-to-noise ratio in player behavior data and waste too much modeling capacity on less informative behaviors, resulting in their learned representations being noisy. In this paper, we propose a novel model for Denoised Player Behavior representation Learning, namely DPBL, which consists of two key modules. The first module extracts various player behavior patterns and isolates them from less informative noise. The second module utilizes the extracted patterns to refine the embedding of each behavior and eliminates noise. To optimize DPBL, two contrastive learning strategies are proposed to identify the noise that should be eliminated and to learn distinguishable representations, respectively. With the above design, DPBL is capable of mitigating the impact of noise in the data and learning high-quality representations that effectively capture player characteristics. We conducted extensive experiments on two real-world datasets, and DPBL outperforms all baselines on various downstream tasks with an improvement of 1.4% ∼ 18.1%. The results also show that DPBL achieves an improvement of 5.6% ∼ 23.0% in the denoising experiments, which proves that DPBL is more robust to noisy behaviors. Code is available at https://github.com/LwbXc/DPBL.
Wenbin Li 0012, Di Yao 0001, Zijie Xu 0006, Chang Gong 0001, Quanliang Jing, Runze Wu 0001, Haining Tan, Zhipeng Hu, Tangjie Lv, Changjie Fan, Jingping Bi
IEEE Trans. Games1
2025 Effective and Efficient Representation Learning for Flight Trajectories
abstract
Flight trajectory data plays a vital role in the traffic management community, especially for downstream tasks such as trajectory prediction, flight recognition, and anomaly detection. Existing works often utilize handcrafted features and design models for different tasks individually, which heavily rely on domain expertise and are hard to extend. We argue that different flight analysis tasks share the same useful features of the trajectory. Jointly learning a unified representation for flight trajectories could be beneficial for improving the performance of various tasks. However, flight trajectory representation learning (TRL) faces two primary challenges, \ie unbalanced behavior density and 3D spatial continuity, which disable recent general TRL methods. In this paper, we propose Flight2Vec, a flight-specific representation learning method to address these challenges. Specifically, a behavior-adaptive patching mechanism is used to inspire the learned representation to pay more attention to behavior-dense segments. Moreover, we introduce a motion trend learning technique that guides the model to memorize not only the precise locations, but also the motion trend to generate better representations. Extensive experimental results demonstrate that Flight2Vec significantly improves performance in downstream tasks such as flight trajectory prediction, flight recognition, and anomaly detection.
Wenbin Li 0012, Di Yao 0001, Jingping Bi
AAAI2
2024 Optimistic Value Instructors for Cooperative Multi-Agent Reinforcement Learning
abstract
In cooperative multi-agent reinforcement learning, decentralized agents hold the promise of overcoming the combinatorial explosion of joint action space and enabling greater scalability. However, they are susceptible to a game-theoretic pathology called relative overgeneralization that shadows the optimal joint action. Although recent value-decomposition algorithms guide decentralized agents by learning a factored global action value function, the representational limitation and the inaccurate sampling of optimal joint actions during the learning process make this problem still. To address this limitation, this paper proposes a novel algorithm called Optimistic Value Instructors (OVI). The main idea behind OVI is to introduce multiple optimistic instructors into the value-decomposition paradigm, which are capable of suggesting potentially optimal joint actions and rectifying the factored global action value function to recover these optimal actions. Specifically, the instructors maintain optimistic value estimations of per-agent local actions and thus eliminate the negative effects caused by other agents' exploratory or sub-optimal non-cooperation, enabling accurate identification and suggestion of optimal joint actions. Based on the instructors' suggestions, the paper further presents two instructive constraints to rectify the factored global action value function to recover these optimal joint actions, thus overcoming the RO problem. Experimental evaluation of OVI on various cooperative multi-agent tasks demonstrates its superior performance against multiple baselines, highlighting its effectiveness.
Jianqi Wang, Yujing Hu, Shaokang Dong, Wenbin Li 0012, Tangjie Lv, Changjie Fan, Yang Gao 0001
AAAI6
2024 CausalTAD: Causal Implicit Generative Model for Debiased Online Trajectory Anomaly Detection
abstract
Trajectory anomaly detection, aiming to estimate the anomaly risk of trajectories given the Source-Destination (SD) pairs, has become a critical problem for many real-world applications. Existing solutions directly train a generative model for observed trajectories and calculate the conditional generative probability$P(T \vert C)$as the anomaly risk, where$T$and$C$represent the trajectory and SD pair respectively. However, we argue that the observed trajectories are confounded by road network preference which is a common cause of both SD distribution and trajectories. Existing methods ignore this issue limiting their generalization ability on out-of-distribution trajectories. In this paper, we define the debiased trajectory anomaly detection problem and propose a causal implicit generative model, namely CausalTAD, to solve it. CausalTAD adopts do-calculus to eliminate the confounding bias of road network preference and estimates$P(T\vert do(C))$as the anomaly criterion. Extensive experiments show that CausalTadcan not only achieve superior performance on trained trajectories but also generally improve the performance of out-of-distribution data, with improvements of 2.1% ~ 5.7% and 10.6% ~ 32.7% respectively.
Wenbin Li 0012, Di Yao 0001, Chang Gong 0001, Xiaokai Chu, Quanliang Jing, Yunxia Fan, Jingping Bi
ICDE1
2024 AnomalyLLM: Few-Shot Anomaly Edge Detection for Dynamic Graphs Using Large Language Models
abstract
Detecting anomaly edges for dynamic graphs aims to identify edges significantly deviating from the normal pattern and can be applied in various domains, such as cybersecurity, financial transactions and AIOps. With the evolving of time, the types of anomaly edges are emerging and the labeled anomaly samples are few for each type. Current methods are either designed to detect randomly inserted edges or require sufficient labeled data for model training, which harms their applicability for real-world applications. In this paper, we study this problem by cooperating with the rich knowledge encoded in large language models(LLMs) and propose a method, namely AnomalyLLM. To align the dynamic graph with LLMs, AnomalyLLM pretrains a dynamic-aware encoder to generate the representations of edges and reprograms the edges using the prototypes of word embeddings. Along with the encoder, we design an in-context learning framework that integrates the information of a few labeled samples to achieve few-shot anomaly detection. Experiments on four datasets reveal that AnomalyLlmcan not only significantly improve the performance of few-shot anomaly detection, but also achieve superior results on new anomalies without any update of model parameters.
Di Yao 0001, Lanting Fang, Zhetao Li, Wenbin Li 0012, Kaiyu Feng, Xiaowen Ji, Jingping Bi
ICDM5
2024 STAR: Spatio-Temporal State Compression for Multi-Agent Tasks with Rich Observations
Yujing Hu, Shangdong Yang, Tangjie Lv, Changjie Fan, Wenbin Li 0012, Chongjie Zhang, Yang Gao 0001
IJCAI6
2024 CausalMMM: Learning Causal Structure for Marketing Mix Modeling
abstract
In online advertising, marketing mix modeling (MMM) is employed to predict the gross merchandise volume (GMV) of brand shops and help decision-makers to adjust the budget allocation of various advertising channels. Traditional MMM methods leveraging regression techniques can fail in handling the complexity of marketing. Although some efforts try to encode the causal structures for better prediction, they have the strict restriction that causal structures are prior-known and unchangeable. In this paper, we define a new causal MMM problem that automatically discovers the interpretable causal structures from data and yields better GMV predictions. To achieve causal MMM, two essential challenges should be addressed: (1) Causal Heterogeneity. The causal structures of different kinds of shops vary a lot. (2) Marketing Response Patterns. Various marketing response patterns i.e., carryover effect and shape effect, have been validated in practice. We argue that causal MMM needs dynamically discover specific causal structures for different shops and the predictions should comply with the prior known marketing response patterns. Thus, we propose CausalMMM that integrates Granger causality in a variational inference framework to measure the causal relationships between different channels and predict the GMV with the regularization of both temporal and saturation marketing response patterns. Extensive experiments show that CausalMMM can not only achieve superior performance of causal structure learning on synthetic datasets with improvements of 5.7%\sim 7.1%, but also enhance the GMV prediction results on a representative E-commerce platform.
Chang Gong 0001, Di Yao 0001, Lei Zhang 0206, Wenbin Li 0012, Yueyang Su, Jingping Bi
WSDM5
2023 Causal Discovery from Temporal Data
abstract
Temporal data representing chronological observations of complex systems can be ubiquitously collected in smart industry, medicine, finance and etc. In the last decade, many tasks have been studied for mining temporal data and offered significant value for various applications. Among these tasks, causal discovery aims to understand the underlying generation mechanism of temporal data and has attracted much research attention. According to whether the data is calibrated, existing causal discovery approaches can be divided into two subtasks, i.e., multivariate time-series causal discovery, and event sequence causal discovery. Previous tutorials or surveys have primarily focused on causal discovery from time-series data and disregarded the second ones. In this tutorial, we elucidate the correlation between the two subtasks and provide a comprehensive review of the existing solutions. Moreover, we offer some potential applications and summarize new perspectives for discovering causal relations from temporal data. We hope the audiences can obtain a systematic overview of this topic and inspire some new ideas for their own research.
Chang Gong 0001, Di Yao 0001, Chuzhe Zhang, Wenbin Li 0012, Jingping Bi, Lun Du, Jin Wang 0007
KDD4
2022 Few-shot Learning for Trajectory-based Mobile Game Cheating Detection
abstract
With the emerging of smartphones, mobile games have attracted billions of players and occupied most of the share for game companies. On the other hand, mobile game cheating, aiming to gain improper advantages by using programs that simulate the players' inputs, severely damages the game's fairness and harms the user experience. Therefore, detecting mobile game cheating is of great importance for mobile game companies. Many PC game-oriented cheating detection methods have been proposed in the past decades, however, they can not be directly adopted in mobile games due to the concern of privacy, power, and memory limitations of mobile devices. Even worse, in practice, the cheating programs are quickly updated, leading to the label scarcity for novel cheating patterns. To handle such issues, we in this paper introduce a mobile game cheating detection framework, namely FCDGame, to detect the cheats under the few-shot learning framework. FCDGame only consumes the screen sensor data, recording users' touch trajectories, which is less sensitive and more general for almost all mobile games. Moreover, a Hierarchical Trajectory Encoder and a Cross-pattern Meta Learner are designed in FCDGame to capture the intrinsic characters of mobile games and solve the label scarcity problem, respectively. Extensive experiments on two real online games show that FCDGame achieves almost 10% improvements in detection accuracy with only few fine-tuned samples.
Yueyang Su, Di Yao 0001, Xiaokai Chu, Wenbin Li 0012, Jingping Bi, Runze Wu 0001, Shize Zhang, Jianrong Tao
KDD4
2022 FingFormer: Contrastive Graph-based Finger Operation Transformer for Unsupervised Mobile Game Bot Detection
abstract
This paper studies the task of detecting bots for online mobile games. Considering the fact of lacking labeled cheating samples and restricted available data in the real detection systems, we aim to study the finger operations captured by screen sensors to infer the potential bots in an unsupervised way. In detail, we introduce a Transformer-style detection model, namely FingFormer. It studies the finger operations in the format of graph structure in order to capture the spatial and temporal relatedness between the two hands’ operations. To optimize the model in an unsupervised way, we introduce two contrastive learning strategies to refine both finger moving patterns and players’ operation habits. We conduct extensive experiments under different experimental environments, including the synthetic dataset, the offline dataset, as well as the large-scale online data flow from three mobile games. The multi-facet experiments illustrate the proposed model is both effective and general to detect the bots for different mobile games.
Wenbin Li 0012, Xiaokai Chu, Yueyang Su, Di Yao 0001, Runze Wu 0001, Shize Zhang, Jianrong Tao, Jingping Bi
WWW1