VLDB 2026 Research / reviewers in the wild / expert
Yingfei Wang
dblp:165/2570
· DBLP profile ↗
9ranked-venue papers in the field
1as first author
8since 2021 · last 2025
—ORCID · conflict
Domains — venue-derived; a paper can count in several
Data Mining & Knowledge Discovery · 8Database Systems & Data Management · 1 (1 first)
| Year | Publication | Venue | Position |
|---|---|---|---|
| 2025 | 3rd Workshop on Causal Inference and Machine Learning in PracticeabstractThe 3rd Workshop on Causal Inference and Machine Learning in Practice at KDD 2025 aims to bring together researchers, industry professionals, and practitioners to explore the application of causal inference within machine learning models. As causal machine learning techniques gain traction across industries, practical challenges related to trustworthiness, robustness, and fairness remain at the forefront. This workshop will provide a forum to discuss methodologies for evaluating causal models in real-world scenarios and explore innovative applications that integrate causal inference with generative AI (GenAI) and large language models (LLMs). Topics of interest include using GenAI and LLMs to facilitate causal inference tasks and leveraging causal inference techniques for evaluating and improving GenAI/LLM models. Building on the success of the previous workshop editions at KDD 2023 and KDD 2024, which attracted over 200 and 250 participants, respectively, this workshop will continue fostering collaboration between academia and industry. Through invited talks, contributed papers, and interactive discussions, we will address key challenges and opportunities at the intersection of causal inference and machine learning. As the field continues to evolve, this workshop serves as a crucial platform for knowledge exchange and innovation, driving forward the application of causal techniques in machine learning and AI. Jeong-Yoon Lee, Totte Harinen, Paul Lo, Huigang Chen, Sichao Yin, Roland Stevenson, Jingshen Wang, Yingfei Wang, Zeyu Zheng 0002 |
KDD (2) | 11 |
| 2025 | A/B Test and Online Experiment Under Diminishing Marginal Effects: Regret Minimization and Statistical InferenceabstractWhen large online platforms test a new strategy to implement with their user traffic, the phenomenon of diminishing marginal effects may arise. For example, when a strategy is implemented on 100% of the user traffic, the expected per-user effect can be lower compared to the expected per-user effect when a strategy is implemented on 10% of the user traffic, potentially due to limits of overall budget, resource, attention or content involved with that strategy. This diminishing marginal effect phenomenon brings an additional delicacy to online sequential experiments. In particular, for the classical goal of achieving the largest expected reward, the optimal decision may no longer be assigning 100% traffic to one strategy, but instead a mixture of strategies. We deliver two tasks for online sequential experiments in presence of diminishing marginal effect: (1) Adaptively identify the optimal traffic allocation to maximize the expected cumulative reward and (2) Construct valid central limit theorem (which is critically needed for A/B tests in online platforms) to perform reliable statistical inference for the expected reward under the optimal traffic allocation that is a priori unknown. We show that classical algorithms can fail to deliver the second task, especially because the statistical inference task presents its own difficulty. We develop a new online algorithm that leverages an additional smoothness condition on how the marginal effects change to achieve both tasks. We prove that this algorithm obtains the best achievable expected cumulative reward. Further, crucially for online platforms' need to do trustworthy statistical inference, the algorithm is proved to enjoy a valid central limit theorem. The theoretical findings are illustrated through numerical experiments. Jingxu Xu, Yuhang Wu 0011, Yingfei Wang, Zeyu Zheng 0002 |
KDD (2) | 3 |
| 2025 | Long short-term search session-based document re-ranking model
Yingfei Wang, Xintao Chu |
Knowl. Inf. Syst. | 4 |
| 2024 | 2nd Workshop on Causal Inference and Machine Learning in PracticeabstractThe workshop's rationale stems from the escalating interest in causal inference and machine learning methodologies within various industrial contexts. This surge in demand underscores the importance for both scholars and practitioners to exchange knowledge and best practices regarding the application of these techniques to tackle real-world challenges. Yet, applying causal machine learning techniques in real-world scenarios presents a range of challenges not addressed in the academic literature. This workshop aims to address the challenges for practical causal machine learning and explore new industry use cases. The workshop will provide a forum for practitioners and researchers to exchange ideas and explore new collaborations. Moreover, this workshop aims to capitalize on the success and achievements of the KDD 2023 Workshop titled "Causal Inference and Machine Learning in Practice". Jeong-Yoon Lee, Totte Harinen, Paul Lo, Huigang Chen, Zeyu Zheng 0002, Hasta Vanchinathan, Yingfei Wang, Roland Stevenson |
KDD | 10 |
| 2024 | Probabilistic graph model and neural network perspective of click models for web search
Yingfei Wang, Xintao Chu |
Knowl. Inf. Syst. | 2 |
| 2023 | Causal Inference and Machine Learning in Practice: Use Cases for Product, Brand, Policy and BeyondabstractThe increasing demand for data-driven decision-making has led to the rapid growth of machine learning applications in various industries. However, the ability to draw causal inferences from observational data remains a crucial challenge. In recent years, causal inference has emerged as a powerful tool for understanding the effects of interventions in complex systems. Combining causal inference with machine learning has the potential to provide a deeper understanding of the underlying mechanisms and to develop more effective solutions to real-world problems. Jeong-Yoon Lee, Keith Battocchi, Fabio Vera, Totte Harinen, Huigang Chen, Zeyu Zheng 0002, Yingfei Wang, Xinwei Ma |
KDD | 11 |
| 2023 | 2nd Workshop on Multi-Armed Bandits and Reinforcement Learning: Advancing Decision Making in E-Commerce and BeyondabstractThe areas of reinforcement learning and multi-armed bandits have recently seen significant innovation, while many application domains, such as e-commerce, are full of problems and challenges to which vanilla RL or MAB methods cannot directly apply. This workshop aims at filling this communication gap by creating a platform for researchers and practitioners from both the method/theory side and application side of the community. Having this platform now instead of at a later time is beneficial to all sides of the community: practitioners and frontline scientists are able to avoid re-inventing existing techniques; theory-oriented researchers can find motivation in industry problems, working within more realistic settings, and making real-world impact. The 2nd Multi-armed Bandits and Reinforcement Learning Workshop was a half day workshop co-located with the 29th ACM SIGKDD Conference on Knowledge Discovery & Data Mining (KDD 2023) in Long Beach, California. Yingfei Wang, Daniel R. Jiang, Jinghai He, Zeyu Zheng 0002 |
KDD | 2 |
| 2021 | Multi-Armed Bandits and Reinforcement Learning: Advancing Decision Making in E-Commerce and BeyondabstractThe areas of reinforcement learning and multi-armed bandits have recently seen significant innovation, while many application domains, such as e-commerce, are full of problems and challenges to which vanilla RL or MAB methods cannot directly apply. This workshop aims at filling this communication gap by creating a platform for researchers and practitioners from both the method/theory side and application side of the community. Having this platform now instead of at a later time is beneficial to all sides of the community: practitioners and frontline scientists are able to avoid re-inventing existing techniques; theory-oriented researchers can find motivation in industry problems, working within more realistic settings, and making real-world impact. The 1st Multi-armed Bandits and Reinforcement Learning Workshop was a full day workshop co-located with the 27th ACM SIGKDD Conference on Knowledge Discovery & Data Mining (KDD 2021) in Singapore. Daniel R. Jiang, Yingfei Wang |
KDD | 4 |
| 2017 | Learning Online Trends for Interactive Query Auto-CompletionabstractQuery auto-completion (QAC) is widely used by modern search engines to assist users by predicting their intended queries. Most QAC approaches rely on deterministic batch learning algorithms trained from past query log data. However, query popularities keep changing all the time and QAC operates in a real-time scenario where users interact with the search engine continually. So, ideally, QAC must be timely and adaptive enough to reflect time-sensitive changes in an online fashion. Second, due to the vertical position bias, a query suggestion with a higher rank tends to attract more clicks regardless of user's original intention. Hence, in the long run, it is important to place some lower ranked yet potentially more relevant queries to higher positions to collect more valuable user feedbacks. In order to tackle these issues, we propose to formulate QAC as a ranked Multi-Armed Bandits (MAB) problem which enjoys theoretical soundness. To utilize prior knowledge from query logs, we propose to use Bayesian inference and Thompson Sampling to solve this MAB problem. Extensive experiments on large scale datasets show that our QAC algorithm has the capacity to adaptively learn temporal trends, and outperforms existing QAC algorithms in ranking qualities. Yingfei Wang, Hua Ouyang, Hongbo Deng, Yi Chang 0001 |
IEEE Trans. Knowl. Data Eng. | 1 |