Shangyou Wang

dblp:371/4892 · DBLP profile ↗
← Back
2ranked-venue papers
0as first author
2since 2021 · last 2025
0009-0001-5738-985XORCID · corroborated

Domains — the database's venue-derived domains; a paper can count in several

Databases, data management, data science and information retrieval · 2 · 2 since 2021

Expertise — from the expertise taxonomy: the topics of the expert's papers under the CCF categories. A weight counts papers with recency: 1 for a paper about the topic, 0.3 when the topic is its context, halved every five years.

Databases, data mining, and information retrieval
1 paper
Graph data management · 75% Distributed and cloud data management · 25%
Artificial intelligence
1 paper
Graph learning · 87% Efficient and distributed learning · 13%
Computer architecture, parallel and distributed computing, and storage systems
1 paper
GPUs and heterogeneous computing · 100%

Topics — the 8 heaviest of 8, each with the papers that count most for it

TopicWeightPapersLastEvidence papers
Machine learning › Graph learning
graph neural network training
0.912025
Capsule: An Out-of-Core Training Mechanism for Colossal GNNs · Proc. ACM Manag. Data 2025
Machine learning › Graph learning › graph neural network training
out-of-core GNN training
0.912025
Capsule: An Out-of-Core Training Mechanism for Colossal GNNs · Proc. ACM Manag. Data 2025
GPUs and heterogeneous computing
GPU memory management
0.912025
Capsule: An Out-of-Core Training Mechanism for Colossal GNNs · Proc. ACM Manag. Data 2025
Distributed and cloud data management › distributed query processing
communication cost optimization
0.812024
Play like a Vertex: A Stackelberg Game Approach for Streaming Graph Partitioning · Proc. ACM Manag. Data 2024
Graph data management
distributed graph processing
0.812024
Play like a Vertex: A Stackelberg Game Approach for Streaming Graph Partitioning · Proc. ACM Manag. Data 2024
Graph data management
graph partitioning
0.812024
Play like a Vertex: A Stackelberg Game Approach for Streaming Graph Partitioning · Proc. ACM Manag. Data 2024
Graph data management › graph partitioning
streaming graph partitioning
0.812024
Play like a Vertex: A Stackelberg Game Approach for Streaming Graph Partitioning · Proc. ACM Manag. Data 2024
Machine learning › Efficient and distributed learning › distributed training
large-scale training
0.312025
Capsule: An Out-of-Core Training Mechanism for Colossal GNNs · Proc. ACM Manag. Data 2025

Methods — techniques the papers use, named apart from their topics

out-of-core training · 1.7GPU kernels · 1.7stackelberg game · 0.8skewness-aware clustering · 0.8
YearPublicationVenuePosition
2025 Capsule: An Out-of-Core Training Mechanism for Colossal GNNs
abstract
Cutting-edge platforms of graph neural networks (GNNs), such as DGL and PyG, harness the parallel processing power of GPUs to extract structural information from graph data, achieving state-of-the-art (SOTA) performance in fields such as recommendation systems, knowledge graphs, and bioinformatics. Despite the computational advantages provided by GPUs, these GNN platforms struggle with scalability challenges due to the colossal graphical structures processed and the limited memory capacities of GPUs. In response, this work introduces Capsule, a new out-of-core mechanism for large-scale GNN training. Unlike existing out-of-core GNN systems, which use main or secondary memory as operative memory and use CPU kernels during non-backpropagation computation, Capsule uses GPU memory and GPU kernels. By substantially leveraging the parallelization capabilities of GPUs, Capsule significantly enhances GNN training efficiency. In addition, Capsule can be smoothly integrated to mainstream open-source GNN frameworks, DGL and PyG, in a play-and-plug manner. Through a prototype implementation and comprehensive experiments on real datasets, we demonstrate that Capsule can achieve up to a 12.02× improvement in runtime efficiency, while using only 22.24% of the main memory, compared to SOTA out-of-core GNN systems.
Yongan Xiang, Zezhong Ding 0001, Shangyou Wang, Xike Xie, Shaohua Kevin Zhou
Proc. ACM Manag. Data4
2024 Play like a Vertex: A Stackelberg Game Approach for Streaming Graph Partitioning
abstract
In the realm of distributed systems tasked with managing and processing large-scale graph-structured data, optimizing graph partitioning stands as a pivotal challenge. The primary goal is to minimize communication overhead and runtime cost. However, alongside the computational complexity associated with optimal graph partitioning, a critical factor to consider is memory overhead. Real-world graphs often reach colossal sizes, making it impractical and economically unviable to load the entire graph into memory for partitioning. This is also a fundamental premise in distributed graph processing, where accommodating a graph with non-distributed systems is unattainable. Currently, existing streaming partitioning algorithms exhibit a skew-oblivious nature, yielding satisfactory partitioning results exclusively for specific graph types. In this paper, we propose a novel streaming partitioning algorithm, the Skewness-aware Vertex-cut Partitioner (S5P ), designed to leverage the skewness characteristics of real graphs for achieving high-quality partitioning. S5P offers high partitioning quality by segregating the graph's edge set into two subsets, head and tail sets. Following processing by a skewness-aware clustering algorithm, these two subsets subsequently undergo a Stackelberg graph game. Our extensive evaluations conducted on substantial real-world and synthetic graphs demonstrate that, in all instances, the partitioning quality of S5P surpasses that of existing streaming partitioning algorithms, operating within the same load balance constraints. For example, S5P can bring up to a 51% improvement in partitioning quality compared to the top partitioner among the baselines. Lastly, we showcase that the implementation of S5P results in up to an 81% reduction in communication cost and a 130% increase in runtime efficiency for distributed graph processing tasks on PowerGraph.
Zezhong Ding 0001, Yongan Xiang, Shangyou Wang, Xike Xie, Shaohua Kevin Zhou
Proc. ACM Manag. Data3