4.4 Article

Diagnosing Root Causes of Intermittent Slow Queries in Cloud Databases

期刊

PROCEEDINGS OF THE VLDB ENDOWMENT
卷 13, 期 8, 页码 1176-1189

出版社

ASSOC COMPUTING MACHINERY
DOI: 10.14778/3389133.3389136

关键词

-

资金

  1. Alibaba Group through Alibaba Innovative Research Program
  2. National Key R&D Program of China [2019YFB1802504]
  3. National Research Center for Information Science and Technology (BNRist) key projects
  4. National Natural Science Foundation of China [61902200]
  5. China Postdoctoral Science Foundation [2019M651015]

向作者/读者索取更多资源

With the growing market of cloud databases, careful detection and elimination of slow queries are of great importance to service stability. Previous studies focus on optimizing the slow queries that result from internal reasons (e.g., poorly-written SQLs). In this work, we discover a different set of slow queries which might be more hazardous to database users than other slow queries. We name such queries Intermittent Slow Queries (iSQs), because they usually result from intermittent performance issues that are external (e.g., at database or machine levels). Diagnosing root causes of iSQs is a tough but very valuable task. This paper presents iSQUAD, Intermittent Slow QUery Anomaly Diagnoser, a framework that can diagnose the root causes of iSQs with a loose requirement for human intervention. Due to the complexity of this issue, a machine learning approach comes to light naturally to draw the interconnection between iSQs and root causes, but it faces challenges in terms of versatility, labeling overhead and interpretability. To tackle these challenges, we design four components, i.e., Anomaly Extraction, Dependency Cleansing, TypeOriented Pattern Integration Clustering (TOPIC) and Bayesian Case Model. iSQUAD consists of an offline clustering & explanation stage and an online root cause diagnosis & update stage. DBAs need to label each iSQ cluster only once at the offline stage unless a new type of iSQs emerges at the online stage. Our evaluations on real-world datasets from Alibaba OLTP Database show that iSQUAD achieves an iSQ root cause diagnosis average F1score of 80.4%, and outperforms existing diagnostic tools in terms of accuracy and efficiency.

作者

我是这篇论文的作者
点击您的名字以认领此论文并将其添加到您的个人资料中。

评论

主要评分

4.4
评分不足

次要评分

新颖性
-
重要性
-
科学严谨性
-
评价这篇论文

推荐

暂无数据
暂无数据