MEGA Hub

A Standardized Framework for Machine Learning in Power System Protection

Authors

Do you know Julian Oelhaf?You can claim authorship or link another user.Do you know Georg Kordowich?You can claim authorship or link another user.Do you know Paula Andrea Pérez-Toro?You can claim authorship or link another user.Do you know Christian Bergler?You can claim authorship or link another user.Do you know Johann Jäger?You can claim authorship or link another user.Do you know Andreas Maier?You can claim authorship or link another user.Do you know Siming Bayer?You can claim authorship or link another user.

Abstract

Studies of machine-learning-based power-system protection increasingly report near-perfect scores, yet the meaning of those scores depends strongly on the evaluation setting. Protection task, physical scope, measurements, timing, targets, preprocessing, and validation often vary jointly and remain incompletely specified. This paper proposes a standardization-oriented framework that treats evaluation design as part of the scientific contribution. It defines seven required study dimensions: protection objective, physical scope, observability, timing and decision windows, targets and sample validity, validation protocol, and evaluation outputs. The framework is instantiated in a bounded case study on the public PROTECT-90 electromagnetic-transient benchmark, comprising 9022 simulated episodes from a 90 kV double-line topology, for onset-conditioned fault classification and localization. Under centralized sensing, simulation-metadata-aligned 20 ms windows, and episode-grouped validation, a multi-layer perceptron (MLP) achieved a five-fold mean macro-averaged F1 score of 0.991 +/- 0.001 for classification and a localization mean absolute error of 10.20 +/- 0.25% of line length (mean +/- std across episode-grouped folds). Extending the decision horizon to 50 ms preserved this task-dependent performance asymmetry, while reduced observability approximately doubled the MLP localization error but had little effect on classification. A synchronized two-ended conventional locator outperformed the learning locators under its richer clean information set, and measurement degradation showed that clean predictive performance did not determine robustness. The framework turns evaluation assumptions into explicit, reproducible evidence and provides a basis for more comparable, auditable evaluation and future certification-oriented assessment of machine-learning protection functions.

Community

00

Publication notes

Author note
32 pages, 4 figures, 26 tables. Code: https://github.com/julianoelhaf/protection-eval-framework. Dataset: PROTECT-90, doi:10.5281/zenodo.21109169