Comparison and Adaptation of Two Strategies for Anomaly Detection in Load Profiles Based on Methods from the Fields of Machine Learning and Statistics
- 1 Limón GmbH, Kassel, Germany
- 2 Limón GmbH, Kassel, Germany
- 3 Department for Sustainable Products and Processes (Upp), University Kassel, Kassel, Germany
Abstract
The Federal Office for Economic Affairs and Export Control (BAFA) of Germany promotes digital concepts for increasing energy efficiency as part of the “Pilotprogramm Einsparz ä hler”. Within this program, Limón GmbH is developing software solutions in cooperation with the University of Kassel to identify efficiency potentials in load profiles by means of automated anomaly detection. Therefore, in this study two strategies for anomaly detection in load profiles are evaluated. To estimate the monthly load profile, strategy 1 uses the artificial neural network LSTM (Long Short-Term Memory), with a data period of one month (1 M) or three months (3 M), and strategy 2 uses the smoothing method PEWMA (Probalistic Exponential Weighted Moving Average). By comparing with original load profile data, residuals or summed residuals of the sequence lengths of two, four, six and eight hours are identified as an anomaly by exceeding a predefined threshold. The thresholds are defined by the Z-Score test, i . e ., residuals greater than 2, 2.5 or 3 standard deviations are considered anomalous. Furthermore, the ESD (Extreme Studentized Deviate) test is used to set thresholds by means of three significance level values of 0.05, 0.10 and 0.15, with a maximum of k = 40 iterations. Five load profiles are examined, which were obtained by the cluster method k -Means as a representative sample from all available data sets of the Limón GmbH. The evaluation shows that for strategy 1 a maximum F 1 -value of 0.4 (1 M) and for all examined companies an average F 1 -value of maximum 0.24 and standard deviation of 0.09 (1 M) could be achieved for the investigation on single residuals. In variant 3 M the highest F 1 -value could be achieved with an average F 1 -value of 0.21 and standard deviation of 0.06 (3 M) for summed residuals of the partial sequence length of four hours. The PEWMA-based strategy 2 did not show a higher anomaly detection efficacy compared to strategy 1 in any of the investigated companies.
- European Commission (2014) Communication from the Commission to the European Parliament, the Council, the European Economic and Social Committee and the Committee of the Regions a Policy Framework for Climate and Energy in the Period from 2020 to 2030. Document 52014DC0015. https://eur-lex.europa.eu/legal-content/EN/TXT/?uri=CELEX%3A52014DC0015&qid=1611915593867
- European Commission (2019) Communication from the Commission to the European Parliament, the Council, the European Economic and Social Committee and the Committee of the Regions the European Green Deal. Document 52014DC0015. https://eur-lex.europa.eu/legal-content/EN/TXT/?uri=COM:2019:640:FIN
- Krewitt, W., Nienhaus, K., Kleßmann, C., Capone, C., Stricker, E., Graus, W., Hoogwijk, M., Supersberger, N., Winterfeld, U. and Samadi, S. (2009) Role and Potential of Renewable Energy and Energy Efficiency for Global Energy Supply. https://www.umweltbundesamt.de/publikationen/role-potential-of-renewable-energy-energy
- Federal Office of Economics and Export Control (2021) Federal Funding for Pilot Program on Energy-Saving Meters. https://www.bafa.de/DE/Energie/Energieeffizienz/Einsparzaehler/einsparzaehler_node.html
- Chandola, V., Banerjee, A. and Kumar, V. (2009) Anomaly Detection: A Survey. ACM Computing Surveys, 41, 1-58. https://doi.org/10.1145/1541880.1541882
- Aggarwal, C.C. (2017) Outlier Analysis. 5th Edition, Springer International Publishing, Cham.
- Blázquez-García, A., Conde, A., Mori, U. and Lozano J.A. (2020) A Review on outlier/Anomaly Detection in Time Series Data. arXiv preprint, arXiv: 2002.04236v1, 1-32. https://arxiv.org/abs/2002.04236
- Gupta, M., Gao, J., Aggarwal, C.C. and Han, J. (2014) Outlier Detection for Temporal Data: A Survey. IEEE Transactions on Knowledge and Data Engineering, 26, 2250-2267. https://doi.org/10.1109/TKDE.2013.184
- Wang, X., Lin, J., Patel, N. and Braun, M. (2018) Exact Variable-Length Anomaly Detection Algorithm for Univariate and Multivariate Time Series. Data Mining and Knowledge Discovery, 32, 1806-1844. https://doi.org/10.1007/s10618-018-0569-7
- Makridakis, S., Spiliotis, E. and Assimakopoulos, V. (2018) Statistical and Machine Learning Forecasting Methods: Concerns and Ways Forward. PLoS ONE, 13, e0194889. https://doi.org/10.1371/journal.pone.0194889
- Hochreiter, S. and Schmidhuber, J. (1997) Long Short-Term Memory. Neural computation, 9, 1735-1780. https://doi.org/10.1162/neco.1997.9.8.1735