Logo image
Anomaly Detection in Cloud-Native Systems
Conference proceeding   Peer reviewed

Anomaly Detection in Cloud-Native Systems

F Lomio, S Moreschini, Xiaozhou Li and V Lenarduzzi
2022 48th Euromicro Conference on Software Engineering and Advanced Applications (SEAA), pp.100-103
Euromicro Conference on Software Engineering and Advanced Applications (Gran Canaria, 31/08/2022–02/09/2022)
2022
Handle:
https://hdl.handle.net/10863/52202

Abstract

Anomaly detection Empirical study Kafka metrics Machine Learning
Companies develop cloud-native systems deployed on public and private clouds. Since private clouds have limited resources, the systems should run efficiently by keeping performance related anomalies under control. The goal of this work is to understand whether a set of five performance-related KPIs depends on the metrics collected at runtime by Kafka, Zookeeper, and other tools (168 different metrics). We considered four weeks worth of runtime data collected from a system running in production. We trained eight Machine Learning algorithms on three weeks worth of data and tested them on one week’s worth of data to compare their prediction accuracy and their training and testing time. It is possible to detect performance-related anomalies with a very high level of accuracy (higher than 95% AUC) and with very limited training time (between 8 and 17 minutes). Machine Learning algorithms can help to identify runtime anomalies and to detect them efficiently. Future work will include the identification of a proactive approach to recognize the root cause of the anomalies and to prevent them as early as possible.
url
https://doi.org/10.1109/SEAA56994.2022.00023View

Details

Metrics

1 Record Views
Logo image