Performance Bug Analysis and Detection for Distributed Storage and Computing Systems

Jiaxin Li,Haryadi S Gunawi,Xiaohui Gu,Shan Lu,Dongsheng Li,Yiming Zhang,Feng Huang

doi:10.1145/3580281

Abstract

This article systematically studies 99 distributed performance bugs from five widely deployed distributed storage and computing systems (Cassandra, HBase, HDFS, Hadoop MapReduce and ZooKeeper). We present the TaxPerf database, which collectively organizes the analysis results as over 400 classification labels and over 2,500 lines of bug re-description. TaxPerf is classified into six bug categories (and 18 bug subcategories) by their root causes; resource, blocking, synchronization, optimization, configuration, and logic. TaxPerf can be used as a benchmark for performance bug studies and debug tool designs. Although it is impractical to automatically detect all categories of performance bugs in TaxPerf, we find that an important category of blocking bugs can be effectively solved by analysis tools. We analyze the cascading nature of blocking bugs and design an automatic detection tool called PCatch , which (i) performs program analysis to identify code regions whose execution time can potentially increase dramatically with the workload size; (ii) adapts the traditional happens-before model to reason about software resource contention and performance dependency relationship; and (iii) uses dynamic tracking to identify whether the slowdown propagation is contained in one job. Evaluation shows that PCatch can accurately detect blocking bugs of representative distributed storage and computing systems by observing system executions under small-scale workloads.

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

R Discovery Prime

R Discovery Prime

Performance Bug Analysis and Detection for Distributed Storage and Computing Systems

Abstract

Talk to us

Similar Papers

More From: ACM Transactions on Storage

Lead the way for us

Journal: ACM Transactions on Storage	Publication Date: Jun 19, 2023
Citations: 1

Similar Papers

Security Versus Performance Bugs: How Bugs are Handled in the Chromium Project
Amrit Rajbhandari ... Farjana Z Eishita
-
Amrit Rajbhandari, et. al.Amrit Rajbhandari ... Farjana Z Eishita
25 May 2022
25 May 2022

Bug Characteristics in Blockchain Systems: A Large-Scale Empirical Study
Zhiyuan Wan ... Xin Xia
-
Zhiyuan Wan, et. al.Zhiyuan Wan ... Xin Xia
01 May 2017
01 May 2017

Discovering, reporting, and fixing performance bugs
Adrian Nistor ... Tian Jiang
-
Adrian Nistor, et. al.Adrian Nistor ... Tian Jiang
01 May 2013
01 May 2013

On Automatic Detection of Performance Bugs
Sokratis Tsakiltsidis ... Andriy Miranskyy
-
Sokratis Tsakiltsidis, et. al.Sokratis Tsakiltsidis ... Andriy Miranskyy
01 Oct 2016
01 Oct 2016

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

Performance Bug Analysis and Detection for Distributed Storage and Computing Systems

Abstract

Talk to us

Similar Papers

More From: ACM Transactions on Storage