EDBT 2026 Demo / reviewers in the wild / expert
Weihang Jiang
dblp:64/5433
· DBLP profile ↗
10ranked-venue papers
4as first author
0since 2021 · last 2009
—ORCID · none
Domains — the database's venue-derived domains; a paper can count in several
Systems, architecture and hardware · 8 · 4 first-authorSoftware engineering, systems software and programming languages · 3Databases, data management, data science and information retrieval · 2 · 2 first-author
Expertise — from the expertise taxonomy: the topics of the expert's papers under the CCF categories. A weight counts papers with recency: 1 for a paper about the topic, 0.3 when the topic is its context, halved every five years.
| Computer architecture, parallel and distributed computing, and storage systems
6 papers |
Storage systems · 52% Energy-efficient computing · 18% Interconnection networks and networks-on-chip · 11% | |
| Software engineering, system software, and programming languages
2 papers |
Program analysis · 48% Concurrent programming · 32% Software testing · 21% |
Topics — the 20 heaviest of 21, each with the papers that count most for it
| Topic | Weight | Papers | Last | Evidence papers |
|---|---|---|---|---|
Storage systems
storage reliability |
0.2 | 2 | 2008 | Are disks the dominant contributor for storage failures - A comprehensive study of storage subsystem failure characteristics · ACM Trans. Storage 2008 Are Disks the Dominant Contributor for Storage Failures? A Comprehensive Study of Storage Subsystem Failure Characteristics · FAST 2008 |
Concurrent programming
concurrency bug detection |
0.1 | 2 | 2007 | MUVI: automatically inferring multi-variable access correlations and detecting related semantic and concurrency bugs · SOSP 2007 A study of interleaving coverage criteria · ESEC/SIGSOFT FSE 2007 |
Storage systems › storage reliability
failure characterization |
0.1 | 1 | 2008 | Are Disks the Dominant Contributor for Storage Failures? A Comprehensive Study of Storage Subsystem Failure Characteristics · FAST 2008 |
Program analysis › static analysis
bug detection |
0.1 | 1 | 2007 | MUVI: automatically inferring multi-variable access correlations and detecting related semantic and concurrency bugs · SOSP 2007 |
Program analysis › error detection
semantic bug detection |
0.1 | 1 | 2007 | MUVI: automatically inferring multi-variable access correlations and detecting related semantic and concurrency bugs · SOSP 2007 |
Program analysis
source code analysis |
0.1 | 1 | 2007 | MUVI: automatically inferring multi-variable access correlations and detecting related semantic and concurrency bugs · SOSP 2007 |
Software testing › concurrency testing
thread interleaving coverage |
0.1 | 1 | 2007 | A study of interleaving coverage criteria · ESEC/SIGSOFT FSE 2007 |
Energy-efficient computing
power-performance tradeoff |
0.1 | 1 | 2007 | Managing energy-performance tradeoffs for multithreaded applications on multiprocessor architectures · SIGMETRICS 2007 |
Energy-efficient computing › power management
memory power management |
0.1 | 1 | 2006 | DMA-aware memory energy management · HPCA 2006 |
High-performance computing
cluster computing |
0.0 | 1 | 2003 | Performance Comparison of MPI Implementations over InfiniBand, Myrinet and Quadrics · SC 2003 |
Interconnection networks and networks-on-chip
cluster interconnect |
0.0 | 1 | 2003 | Performance Comparison of MPI Implementations over InfiniBand, Myrinet and Quadrics · SC 2003 |
Interconnection networks and networks-on-chip › cluster interconnect
infiniband |
0.0 | 1 | 2003 | Performance Comparison of MPI Implementations over InfiniBand, Myrinet and Quadrics · SC 2003 |
Cloud and datacenter computing
log analysis |
0.0 | 1 | 2009 | Understanding Customer Problem Troubleshooting from Storage System Logs · FAST 2009 |
Storage systems › storage reliability
disk failure |
0.0 | 1 | 2008 | Are Disks the Dominant Contributor for Storage Failures? A Comprehensive Study of Storage Subsystem Failure Characteristics · FAST 2008 |
Storage systems › storage reliability
RAID |
0.0 | 1 | 2008 | Are disks the dominant contributor for storage failures - A comprehensive study of storage subsystem failure characteristics · ACM Trans. Storage 2008 |
Software testing › test adequacy
coverage criteria |
0.0 | 1 | 2007 | A study of interleaving coverage criteria · ESEC/SIGSOFT FSE 2007 |
Processor architecture and microarchitecture
multicore design |
0.0 | 1 | 2007 | Managing energy-performance tradeoffs for multithreaded applications on multiprocessor architectures · SIGMETRICS 2007 |
Parallel and multicore computing › thread-level parallelism
multithreaded applications |
0.0 | 1 | 2007 | Managing energy-performance tradeoffs for multithreaded applications on multiprocessor architectures · SIGMETRICS 2007 |
Performance modeling and evaluation › simulation › discrete-event simulation
trace-driven simulation |
0.0 | 1 | 2006 | DMA-aware memory energy management · HPCA 2006 |
Performance modeling and evaluation
benchmarking |
0.0 | 1 | 2003 | Performance Comparison of MPI Implementations over InfiniBand, Myrinet and Quadrics · SC 2003 |
Methods — techniques the papers use, named apart from their topics
field data analysis · 0.1failure correlation analysis · 0.1static code analysis · 0.1performance-guaranteed energy management · 0.1lock-set · 0.1interleaving exploration · 0.1happens-before · 0.1trace-driven simulation · 0.1microbenchmarking · 0.0application benchmarking · 0.0
| Year | Publication | Venue | Position |
|---|---|---|---|
| 2009 | Understanding Customer Problem Troubleshooting from Storage System Logs
Weihang Jiang, Chongfeng Hu, Shankar Pasupathy, Arkady Kanevsky, Zhenmin Li, Yuanyuan Zhou 0001 |
FAST | 1 |
| 2008 | Are Disks the Dominant Contributor for Storage Failures? A Comprehensive Study of Storage Subsystem Failure Characteristics
Weihang Jiang, Chongfeng Hu, Yuanyuan Zhou 0001, Arkady Kanevsky |
FAST | 1 |
| 2008 | Are disks the dominant contributor for storage failures - A comprehensive study of storage subsystem failure characteristicsabstractBuilding reliable storage systems becomes increasingly challenging as the complexity of modern storage systems continues to grow. Understanding storage failure characteristics is crucially important for designing and building a reliable storage system. While several recent studies have been conducted on understanding storage failures, almost all of them focus on the failure characteristics of one component—disks—and do not study other storage component failures. This article analyzes the failure characteristics of storage subsystems. More specifically, we analyzed the storage logs collected from about 39,000 storage systems commercially deployed at various customer sites. The dataset covers a period of 44 months and includes about 1,800,000 disks hosted in about 155,000 storage-shelf enclosures. Our study reveals many interesting findings, providing useful guidelines for designing reliable storage systems. Some of our major findings include: (1) In addition to disk failures that contribute to 20--55% of storage subsystem failures, other components such as physical interconnects and protocol stacks also account for a significant percentage of storage subsystem failures. (2) Each individual storage subsystem failure type, and storage subsystem failure as a whole, exhibits strong self-correlations. In addition, these failures exhibit “bursty” patterns. (3) Storage subsystems configured with redundant interconnects experience 30--40% lower failure rates than those with a single interconnect. (4) Spanning disks of a RAID group across multiple shelves provides a more resilient solution for storage subsystems than within a single shelf. Weihang Jiang, Chongfeng Hu, Yuanyuan Zhou 0001, Arkady Kanevsky |
ACM Trans. Storage | 1 |
| 2007 | Managing energy-performance tradeoffs for multithreaded applications on multiprocessor architecturesabstractIn modern computers, non-performance metrics such as energy consumption have become increasingly important, requiring tradeoff with performance. A recent work has proposed performance-guaranteed energy management, but it is designed specifically for sequential applications and cannot be used to a large class of multithreaded applications running on high end computers and data servers. Weihang Jiang, Yuanyuan Zhou 0001, Sarita V. Adve |
SIGMETRICS | 2 |
| 2007 | A study of interleaving coverage criteriaabstractConcurrency bugs are becoming increasingly important due to the prevalence of concurrent programs. A fundamental problem of concurrent program bug detection and testing is that the interleaving space is too large to be thoroughly explored. Practical yet effective interleaving coverage criteria are desired to systematically explore the interleaving space and effectively expose concurrency bugs. Shan Lu 0001, Weihang Jiang, Yuanyuan Zhou 0001 |
ESEC/SIGSOFT FSE | 2 |
| 2007 | MUVI: automatically inferring multi-variable access correlations and detecting related semantic and concurrency bugsabstractSoftware defects significantly reduce system dependability. Among various types of software bugs, semantic and concurrency bugs are two of the most difficult to detect. This paper proposes a novel method, called MUVI, that detects an important class of semantic and concurrency bugs. MUVI automatically infers commonly existing multi-variable access correlations through code analysis and then detects two types of related bugs: (1) inconsistent updates--correlated variables are not updated in a consistent way, and (2) multi-variable concurrency bugs--correlated accesses are not protected in the same atomic sections in concurrent programs.We evaluate MUVI on four large applications: Linux, Mozilla,MySQL, and PostgreSQL. MUVI automatically infers more than 6000 variable access correlations with high accuracy (83%).Based on the inferred correlations, MUVI detects 39 new inconsistent update semantic bugs from the latest versions of these applications, with 17 of them recently confirmed by the developers based on our reports.We also implemented MUVI multi-variable extensions to tworepresentative data race bug detection methods (lock-set and happens-before). Our evaluation on five real-world multi-variable concurrency bugs from Mozilla and MySQL shows that the MUVI-extension correctly identifies the root causes of four out of the five multi-variable concurrency bugs with 14% additional overhead on average. Interestingly, MUVI also helps detect four new multi-variable concurrency bugs in Mozilla that have never been reported before. None of the nine bugs can be identified correctly by the original race detectors without our MUVI extensions. Shan Lu 0001, Chongfeng Hu, Xiao Ma 0014, Weihang Jiang, Zhenmin Li, Raluca A. Popa, Yuanyuan Zhou 0001 |
SOSP | 5 |
| 2006 | DMA-aware memory energy managementabstractAs increasingly larger memories are used to bridge the widening gap between processor and disk speeds, main memory energy consumption is becoming increasingly dominant. Even though much prior research has been conducted on memory energy management, no study has focused on data servers, where main memory is predominantly accessed by DMAs instead of processors. In this paper, we study DMA-aware techniques for memory energy management in data servers. We first characterize the effect of DMA accesses on memory energy and show that, due to the mismatch between memory and I/O bus band-widths, significant energy is wasted when memory is idle but still active during DMA transfers. To reduce this waste, we propose two novel performance-directed energy management techniques that maximize the utilization of memory devices by increasing the level of concurrency between multiple DMA transfers from different I/O buses to the same memory device. We evaluate our techniques using a detailed trace-driven simulator, and storage and database server traces. The results show that our techniques can effectively minimize the amount of idle energy waste during DMA transfers and, consequently, conserve up to 38.6% more memory energy than previous approaches while providing similar performance. Vivek Pandey, Weihang Jiang, Yuanyuan Zhou 0001, Ricardo Bianchini |
HPCA | 2 |
| 2004 | High performance MPI-2 one-sided communication over InfiniBandabstractMany existing MPI-2 one-sided communication implementations are built on top of MPI send/receive operations. Although this approach can achieve good portability, it suffers front high communication overhead and dependency on remote process for communication progress. To address these problems, we propose a high performance MPI-2 one-sided communication design over the InfiniBand Architecture. In our design, MPI-2 one-sided communication operations such as MPI-Put, MPI-Get and MPI-Accumulate are directly mapped to InfiniBand Remote Direct Memory Access (RDMA) operations. Our design has been implemented based on MPICH2 over InfiniBand. We present detailed design issues for this approach and perform a set of microbenchmarks to characterize different aspects of its performance. Our performance evaluation shows that compared with the design based on MPI send/receive, our design can improve throughput up to 77%, and reduce latency and synchronization overhead up to 19% and 13%, respectively. Under certain process skew, the bad impact can be significantly reduced by new design, from 41% to nearly 0%. It also can achieve better overlap of communication and computation. Weihang Jiang, Jiuxing Liu, Hyun-Wook Jin, Dhabaleswar K. Panda 0001, William Gropp, Rajeev Thakur |
CCGRID | 1 |
| 2004 | Design and Implementation of MPICH2 over InfiniBand with RDMA SupportabstractSummary form only given. For several years, MPI has been the de facto standard for writing parallel applications. One of the most popular MPI implementations is MPICH. Its successor, MPICH2, features a completely new design that provides more performance and flexibility. To ensure portability, it has a hierarchical structure based on which porting can be done at different levels. In this paper, we present our experiences in designing and implementing MPICH2 over InfiniBand. Because of its high performance and open standard, InfiniBand is gaining popularity in the area of high-performance computing. Our study focuses on optimizing the performance of MPl-1 functions in MPICH2. One of our objectives is to exploit remote direct memory access (RDMA) in InfiniBand to achieve high performance. We have based our design on the RDMA channel interface provided by MP1CH2, which encapsulates architecture-dependent communication functionalities into a very small set of functions. Starting with a basic design, we apply different optimizations and also propose a zero-copy-based design. We characterize the impact of our optimizations and designs using microbenchmarks. We have also performed an application-level evaluation using the NAS parallel benchmarks. Our optimized MPICH2 implementation achieves 7.6/spl mu/s latency and 857 MB/s bandwidth, which are close to the raw performance of the underlying InfiniBand layer. Our study shows that the RDMA channel interface in MPICH2 provides a simple, yet powerful, abstraction that enables implementations with high performance by exploiting RDMA operations in InfiniBand. To the best of our knowledge, this is the first high-performance design and implementation ofMPICH2 on InfiniBand using RDMA support. Jiuxing Liu, Weihang Jiang, Pete Wyckoff, Dhabaleswar K. Panda 0001, David Ashton, Darius Buntinas, William Gropp, Brian R. Toonen |
IPDPS | 2 |
| 2003 | Performance Comparison of MPI Implementations over InfiniBand, Myrinet and QuadricsabstractIn this paper, we present a comprehensive performance comparison of MPI implementations over Infini-Band, Myrinet and Quadrics. Our performance evaluation consists of two major parts. The first part consists of a set of MPI level micro-benchmarks that characterize different aspects of MPI implementations. The second part of the performance evaluation consists of application level benchmarks. We have used the NAS Parallel Benchmarks and the sweep3D benchmark. We not only present the overall performance results, but also relate application communication characteristics to the information we acquired from the micro-benchmarks. Our results show that the three MPI implementations all have their advantages and disadvantages. For our 8-node cluster, InfiniBand can offer significant performance improvements for a number of applications compared with Myrinet and Quadrics when using the PCI-X bus. Even with just the PCI bus, InfiniBand can still perform better if the applications are bandwidth-bound. Jiuxing Liu, B. Chandrasekaran 0001, Jiesheng Wu, Weihang Jiang, Sushmitha P. Kini, Weikuan Yu, Darius Buntinas, Pete Wyckoff, Dhabaleswar K. Panda 0001 |
SC | 4 |