EDBT 2026 Demo / reviewers in the wild / expert
Sookwang Lee
dblp:391/8793
· DBLP profile ↗
2ranked-venue papers in the field
0as first author
2since 2021 · last 2025
0009-0002-3326-278XORCID · corroborated
Domains — venue-derived; a paper can count in several
Big Data, Cloud & Distributed Data Systems · 2
| Year | Publication | Venue | Position |
|---|---|---|---|
| 2025 | DEPUTY: A DPU-Based Network Offloading Architecture with Minimal CPU Involvement for Stable Network Performance
Yuchan Lee, Sooho Jang, Sookwang Lee, Shin-Young Ahn, Jaehwan Lee 0001 |
IEEE Big Data | 3 |
| 2024 | Efficient Data-parallel Distributed DNN Training for Big Dataset under Heterogeneous GPU ClusterabstractTraining large-scale deep neural networks (DNNs) using a large number of parameters requires significant computational resources. Despite the rapid advancements in GPU technology, limited budgets have forced many institutions to gradually build GPU servers, leading to growing challenges with resource heterogeneity. However, most open-source distributed deep-learning libraries use synchronous training algorithms that perform better on homogeneous GPUs than on heterogeneous GPUs. Therefore, many researchers have struggled to efficiently conduct distributed training on heterogeneous GPU clusters owing to the straggler problem. In this study, we introduce Efficient Distributed Deep learning lIbrary based on SoftMemoryBox(EDDIS), a novel data-parallel distributed deep learning library. EDDIS overcomes the scalability limitations caused by heterogeneity, enabling the efficient utilization of heterogeneous GPU resources. EDDIS trains DNNs synchronously, asynchronously, and in a hybrid manner and supports TensorFlow and PyTorch. In a heterogeneous GPU environment, EDDIS’s three training modes—synchronous, asynchronous, and hybrid synchronous—accelerate distributed DNN training by approximately 8.2x, 19x, and 18.7x on 16 nodes, respectively, compared to a single node. In particular, the EDDIS hybrid synchronous training mode achieves training speeds that are 2.8 times faster than PyTorch DDP and 2.3 times faster than Horovod when training the Yolov5m model. Shin-Young Ahn, Sookwang Lee, Hyeonseong Choi |
IEEE Big Data | 2 |