Jingzhi Fang

dblp:239/4453 · DBLP profile ↗
← Back
9ranked-venue papers
5as first author
6since 2021 · last 2026
—ORCID · conflict

Domains — the database's venue-derived domains; a paper can count in several

Databases, data management, data science and information retrieval · 6 · 4 first-author · 4 since 2021Artificial intelligence and machine learning · 2 · 2 since 2021Human-computer interaction and ubiquitous computing · 1 · 1 first-authorApplied, interdisciplinary, general and emerging computing · 1 · 1 first-author
YearPublicationVenuePosition
2026 Mil: Cost-guided Minimum Makespan Scheduling for Applications of Multiple LLMs
Jingzhi Fang, Yanyan Shen
Proc. VLDB Endow.1
2025 DIDS: Domain Impact-aware Data Sampling for Large Language Model Training
abstract
Weijie Shi, Jipeng Zhang, Yaguang Wu, Jingzhi Fang, Shibo Zhang, Yao Zhao, Hao Chen, Ruiyuan Zhang, Yue Cui, Jia Zhu, Sirui Han, Jiajie Xu, Xiaofang Zhou. Proceedings of the 2025 Conference on Empirical Methods in Natural Language Processing. 2025.
Yaguang Wu, Jingzhi Fang, Ruiyuan Zhang, Yue Cui 0001, Jia Zhu 0003, Sirui Han, Jiajie Xu 0001, Xiaofang Zhou 0001
EMNLP4
2025 Semantic-guided Diverse Decoding for Large Language Model
abstract
Diverse decoding of large language models is crucial for applications requiring multiple semantically distinct responses, yet existing methods primarily achieve lexical rather than semantic diversity. This limitation significantly constrains Best-of-N strategies, group-based reinforcement learning, and data synthesis. While temperature sampling and diverse beam search modify token distributions or apply n-gram penalties, they fail to ensure meaningful semantic differentiation. We introduce Semantic-guided Diverse Decoding (SemDiD), operating directly in embedding space that balances quality with diversity through three complementary mechanisms: orthogonal directional guidance, dynamic inter-group repulsion, and position-debiased probability assessment. SemDiD harmonizes these competing objectives using adaptive gain functions and constraint optimization, ensuring both quality thresholds and maximal semantic differentiation. Experiments show SemDiD consistently outperforms existing methods, improving Best-of-N coverage by 1.4-5.2% across diverse tasks and accelerating RLHF training convergence by 15% while increasing accuracy by up to 2.1%.
Yue Cui 0001, Yaguang Wu, Jingzhi Fang, Mengze Li 0001, Sirui Han, Jia Zhu 0003, Jiajie Xu 0001, Xiaofang Zhou 0001
NeurIPS4
2024 STile: Searching Hybrid Sparse Formats for Sparse Deep Learning Operators Automatically
abstract
Sparse operators, i.e., operators that take sparse tensors as input, are of great importance in deep learning models. Due to the diverse sparsity patterns in different sparse tensors, it is challenging to optimize sparse operators by seeking an optimal sparse format, i.e., leading to the lowest operator latency. Existing works propose to decompose a sparse tensor into several parts and search for a hybrid of sparse formats to handle diverse sparse patterns. However, they often make a trade-off between search space and search time: their search spaces are limited in some cases, resulting in limited operator running efficiency they can achieve. In this paper, we try to extend the search space in its breadth (by doing flexible sparse tensor transformations) and depth (by enabling multi-level decomposition). We formally define the multi-level sparse format decomposition problem, which is NP-hard, and we propose a framework STile for it. To search efficiently, a greedy algorithm is used, which is guided by a cost model about the latency of computing a sub-task of the original operator after decomposing the sparse tensor. Experiments of two common kinds of sparse operators, SpMM and SDDMM, are conducted on various sparsity patterns, and we achieve 2.1-18.0× speedup against cuSPARSE on SpMMs and 1.5 - 6.9× speedup against DGL on SDDMM. The search time is less than one hour for any tested sparse operator, which can be amortized.
Jingzhi Fang, Yanyan Shen, Yue Wang 0012, Lei Chen 0002
Proc. ACM Manag. Data1
2024 Efficient Training of Graph Neural Networks on Large Graphs
abstract
Graph Neural Networks (GNNs) have gained significant popularity for learning representations of graph-structured data. Mainstream GNNs employ the message passing scheme that iteratively propagates information between connected nodes through edges. However, this scheme incurs high training costs, hindering the applicability of GNNs on large graphs. Recently, the database community has extensively researched effective solutions to facilitate efficient GNN training on massive graphs. In this tutorial, we provide a comprehensive overview of the GNN training process based on the graph data lifecycle, covering graph preprocessing, batch generation, data transfer, and model training stages. We discuss recent data management efforts aiming at accelerating individual stages or improving the overall training efficiency. Recognizing the distinct training issues associated with static and dynamic graphs, we first focus on efficient GNN training on static graphs, followed by an exploration of training GNNs on dynamic graphs. Finally, we suggest some potential research directions in this area. We believe this tutorial is valuable for researchers and practitioners to understand the bottleneck of GNN training and the advanced data management techniques to accelerate the training of different GNNs on massive graphs in diverse hardware settings.
Yanyan Shen, Lei Chen 0002, Jingzhi Fang, Xin Zhang 0101, Shihong Gao
Proc. VLDB Endow.3
2021 ETO: Accelerating Optimization of DNN Operators by High-Performance Tensor Program Reuse
abstract
Recently, deep neural networks (DNNs) have achieved great success in various applications, where low inference latency is important. Existing solutions either manually tune the kernel library or utilize search-based compilation to reduce the operator latency. However, manual tuning requires significant engineering effort, and the huge search space makes the search cost of the search-based compilation unaffordable in some situations. In this work, we propose ETO, a framework for speeding up DNN operator optimization based on reusing the information of performant tensor programs. Specifically, ETO defines conditions for the information reuse between two operators. For operators satisfying the conditions, based on the performant tensor program information of one operator, ETO uses a reuse-based tuner to significantly prune the search space of the other one, and keeps optimization effectiveness at the same time. In this way, for a set of operators, ETO first determines the information reuse relationships among them to reduce the total search time needed, and then tunes the operators either by the backend compiler or by the reuse-based tuner accordingly. ETO further increases the reuse opportunities among the operators by injecting extra operators as bridges between two operators which do not satisfy the reuse conditions. Compared with various existing methods, the experiments show that ETO is effective and efficient in optimizing DNN operators.
Jingzhi Fang, Yanyan Shen, Yue Wang 0012, Lei Chen 0002
Proc. VLDB Endow.1
2020 A Detection Algorithm of Giant Panda in Wild Video Image Based on Wavelet-SSD Network
abstract
The image and video of the giant panda captured by the infrared camera in the field are of great significance to the research and protection of the giant panda. However, there are few pandas in the field and many useless data are also captured by the infrared camera, which brings many difficulties to the detection of the panda image. In order to quickly find the panda image from a large number of images, this paper proposes a special method of fast detection of giant panda based on the WL-SSD (wavelet feature single shot multibox detection) network using Spatial domain and frequency domain feature. The algorithm extracts the frequency domain features of the giant panda image through wavelet feature network. The experimental results show that the giant panda detection algorithm proposed in this paper better use of the high frequency characteristics of the giant panda texture feature, and it can find the panda image in high accuracy and detection rate from a large number of images.
Jingzhi Fang, Hengyi Yang, Shaoxiang Hu
SMC1
2020 Two-sided online bipartite matching in spatial data: experiments and analysis
Jingzhi Fang, Yuxiang Zeng, Balz Maag, Yongxin Tong, Lingyu Zhang 0001
GeoInformatica2
2020 Optimizing DNN Computation Graph using Graph Substitutions
Jingzhi Fang, Yanyan Shen, Yue Wang 0012, Lei Chen 0002
Proc. VLDB Endow.1