VLDB 2026 Research / reviewers in the wild / expert
Tingqiao Xu
dblp:419/7810
· DBLP profile ↗
2ranked-venue papers
2as first author
2since 2021 · last 2026
0009-0002-2688-7212ORCID · reported
Domains — the database's venue-derived domains; a paper can count in several
Artificial intelligence and machine learning · 1 · 1 first-author · 1 since 2021Databases, data management, data science and information retrieval · 1 · 1 first-author · 1 since 2021
Expertise — from the expertise taxonomy: the topics of the expert's papers under the CCF categories. A weight counts papers with recency: 1 for a paper about the topic, 0.3 when the topic is its context, halved every five years.
| Artificial intelligence
2 papers |
Efficient and distributed learning · 38% Vision and language · 38% Language models and text generation · 24% | |
| Databases, data mining, and information retrieval
1 paper |
Information retrieval · 100% |
Topics — the 3 heaviest of 6, each with the papers that count most for it
| Topic | Weight | Papers | Last | Evidence papers |
|---|---|---|---|---|
Information retrieval
e-commerce search |
1.0 | 1 | 2026 | Learning to Trust: Dynamic Utilization of Retrieval-Augmented Generation for E-commerce Search Relevance · SIGIR 2026 |
Information retrieval
retrieval-augmented generation |
1.0 | 1 | 2026 | Learning to Trust: Dynamic Utilization of Retrieval-Augmented Generation for E-commerce Search Relevance · SIGIR 2026 |
Natural language and speech › Language models and text generation
large language model reasoning |
0.3 | 1 | 2026 | Learning to Trust: Dynamic Utilization of Retrieval-Augmented Generation for E-commerce Search Relevance · SIGIR 2026 |
Methods — techniques the papers use, named apart from their topics
group relative policy optimization · 2.9retrieval-augmented generation · 2.0reinforcement learning · 2.0chain-of-thought · 2.0vision priors · 0.9expert fusion · 0.9OCR · 0.9
| Year | Publication | Venue | Position |
|---|---|---|---|
| 2026 | Learning to Trust: Dynamic Utilization of Retrieval-Augmented Generation for E-commerce Search RelevanceabstractAccurately estimating query-item relevance is vital for e-commerce ranking and conversion. While Large Language Models (LLMs) excel at reasoning, they often lack specialized knowledge required for long-tail or fast-evolving queries, necessitating Retrieval-Augmented Generation (RAG). However, production environments face three critical challenges: (1) external context is inherently noisy and inconsistent; (2) extreme latency budgets prohibit multi-stage processing or refinement; and (3) the model must simultaneously assess relevance and context-trust within a unified inference pass. We propose DyKnow-RAG, a reinforcement learning framework that teaches LLMs to learn to trust through dynamic utilization of external knowledge. Built on Group Relative Policy Optimization (GRPO), DyKnow-RAG utilizes a dual-group rollout strategy (parametric-only vs. with-context) and a posterior-driven inter-group advantage scaling mechanism. This enables the model to optimize context utilization without human process labels or extra inference overhead. Our pipeline further integrates structured Chain-of-Thought (CoT) and an uncertainty-prioritized RL pool to stabilize training. Offline evaluations show significant Macro-F1 and Accuracy gains, particularly on noise-sensitive query slices. Importantly, DyKnow-RAG has been deployed in Taobao's production system, serving hundreds of millions of active users and billions of daily search requests. Controlled A/B tests demonstrate consistent lifts in key business metrics, including GSB and Item Goodrate, while maintaining a p99 latency under 400ms. This work provides a scalable and deployable paradigm for operationalizing noisy RAG under extreme efficiency constraints of large-scale industrial search. Tingqiao Xu, Shaowei Yao, Chenhe Dong, Zerui Huang, Dan Ou, Haihong Tang, Bo Zheng 0007 |
SIGIR | 1 |
| 2025 | VERITAS: Leveraging Vision Priors and Expert Fusion to Improve Multimodal DataabstractThe quality of supervised fine-tuning (SFT) data is crucial for the performance of large multimodal models (LMMs), yet current data enhancement methods often suffer from factual errors and hallucinations due to inadequate visual perception.To address this challenge, we propose VERITAS, a pipeline that systematically integrates vision priors and multiple state-of-the-art LMMs with statistical methods to enhance SFT data quality.VERITAS leverages visual recognition models (RAM++) and OCR systems (PP-OCRv4) to extract structured vision priors, which are combined with images, questions, and answers.Three LMMs (GPT-4o, Gemini-2.5-Pro,Doubao-1.5-pro)evaluate the original answers, providing critique rationales and scores that are statistically fused into a high-confidence consensus score serving as ground truth.Using this consensus, we train a lightweight critic model via Group Relative Policy Optimization (GRPO), enhancing reasoning capabilities efficiently.Each LMM then refines the original answers based on the critiques, generating new candidate answers; we select the highest-scoring one as the final refined answer.Experiments across six multimodal benchmarks demonstrate that models fine-tuned with data processed by VERITAS consistently outperform those using raw data, particularly in text-rich and fine-grained reasoning tasks.Our critic model exhibits enhanced capability comparable to state-of-theart LMMs while being significantly more efficient.We release our pipeline, datasets, and model checkpoints to advance research in multimodal data optimization. Tingqiao Xu, Ziru Zeng |
EMNLP | 1 |