EDBT 2026 Demo / reviewers in the wild / expert
Chuanghui Yin
dblp:343/8993
· DBLP profile ↗
1ranked-venue papers
0as first author
1since 2021 · last 2024
0009-0000-2186-528XORCID · reported
Domains — the database's venue-derived domains; a paper can count in several
Systems, architecture and hardware · 1 · 1 since 2021
Expertise — from the expertise taxonomy: the topics of the expert's papers under the CCF categories. A weight counts papers with recency: 1 for a paper about the topic, 0.3 when the topic is its context, halved every five years.
| Computer architecture, parallel and distributed computing, and storage systems
1 paper |
Hardware accelerators and domain-specific architectures · 30% High-performance computing · 30% GPUs and heterogeneous computing · 30% |
Topics — the 4 heaviest of 4, each with the papers that count most for it
| Topic | Weight | Papers | Last | Evidence papers |
|---|---|---|---|---|
GPUs and heterogeneous computing › CPU-GPU heterogeneous computing
CPU-GPU heterogeneous systems |
0.8 | 1 | 2024 | Efficient Utilization of Multi-Threading Parallelism on Heterogeneous Systems for Sparse Tensor Contraction · IEEE Trans. Parallel Distributed Syst. 2024 |
Hardware accelerators and domain-specific architectures › sparsity exploitation
sparse tensor computation |
0.8 | 1 | 2024 | Efficient Utilization of Multi-Threading Parallelism on Heterogeneous Systems for Sparse Tensor Contraction · IEEE Trans. Parallel Distributed Syst. 2024 |
High-performance computing › tensor computation
sparse tensor contraction |
0.8 | 1 | 2024 | Efficient Utilization of Multi-Threading Parallelism on Heterogeneous Systems for Sparse Tensor Contraction · IEEE Trans. Parallel Distributed Syst. 2024 |
Parallel and multicore computing
pipeline parallelism |
0.2 | 1 | 2024 | Efficient Utilization of Multi-Threading Parallelism on Heterogeneous Systems for Sparse Tensor Contraction · IEEE Trans. Parallel Distributed Syst. 2024 |
Methods — techniques the papers use, named apart from their topics
parallel pipeline · 0.8fine-grained partitioning · 0.8
| Year | Publication | Venue | Position |
|---|---|---|---|
| 2024 | Efficient Utilization of Multi-Threading Parallelism on Heterogeneous Systems for Sparse Tensor ContractionabstractMany fields of scientific simulation, such as chemistry and condensed matter physics, are increasingly eschewing dense tensor contraction in favor of sparse tensor contraction. In this work, we center around binary sparse tensor contraction (SpTC) which has the challenges of index matching and accumulation. To address these difficulties, we present GSpTC, an efficient element-wise SpTC framework on CPU-GPU heterogeneous systems. GSpTC first introduces a fine-grained partitioning strategy based on element-wise tensor contraction. By analyzing and selecting appropriate dimension partitioning strategies, we can efficiently utilize the multi-threading parallelism on GPUs and optimize the overall performance of GSpTC. In particular, GSpTC leverages multi-threading parallelism on GPUs for the contraction phase and merging phase, which greatly accelerates the computation phase in sparse tensor contraction computations. Furthermore, GSpTC employs parallel pipeline technology to hide the data transmission time between the host and the device, further enhancing its performance. As a result, GSpTC achieves an average performance improvement of 267% compared to the previous state-of-the-art framework Sparta. Guoqing Xiao 0001, Chuanghui Yin, Yuedan Chen, Mingxing Duan, Kenli Li 0001 |
IEEE Trans. Parallel Distributed Syst. | 2 |