Validating Machine-learned Diagnostic Classifiers in Safety Critical Applications with Imbalanced Populations

Daniel Wade; Andrew Wilson; Abraham Reddy; Raj Bharadwaj

doi:10.36001/phmconf.2018.v10i1.192

Validating Machine-learned Diagnostic Classifiers in Safety Critical Applications with Imbalanced Populations

PDF

Published Sep 24, 2018

DOI https://doi.org/10.36001/phmconf.2018.v10i1.192

Daniel Wade

United States Army AMRDEC

Andrew Wilson

United States Army AMRDEC

Abraham Reddy

Honeywell Aerospace

Raj Bharadwaj

Honeywell Aerospace

Abstract

Data science techniques such as machine learning are rapidly becoming available to engineers building models from system data, such as aircraft operations data. These techniques require validation for use in fielded systems providing recommendations to operators or maintainers. The methods for validating and testing machine learned algorithms generally focus on model performance metrics such as accuracy or F1-score. Many aviation datasets are highly imbalanced, which can invalidate some underlying assumptions of machine learning models. Two simulations are performed to show how some common performance metrics respond to imbalanced populations. The results show that each performance metric responds differently to a sample depending on the imbalance ratio between two classes. The results indicate that traditional methods for repairing underlying imbalance in the sample may not provide the rigorous validation necessary in safety critical applications. The two simulations indicate that authorities must be cautious when mandating metrics for model acceptance criteria because they can significantly influence the model parameters.

How to Cite

Wade, D., Wilson, A., Reddy, A., & Bharadwaj, R. (2018). Validating Machine-learned Diagnostic Classifiers in Safety Critical Applications with Imbalanced Populations. Annual Conference of the PHM Society, 10(1). https://doi.org/10.36001/phmconf.2018.v10i1.192

Abstract 740 | PDF Downloads 763

Keywords

machine-learned model validation

Issue

Vol. 10 No. 1 (2018): Proceedings of the Annual Conference of the PHM Society 2018

Section

Technical Research Papers

This work is licensed under a Creative Commons Attribution 3.0 Unported License.

The Prognostic and Health Management Society advocates open-access to scientific data and uses a Creative Commons license for publishing and distributing any papers. A Creative Commons license does not relinquish the author’s copyright; rather it allows them to share some of their rights with any member of the public under certain conditions whilst enjoying full legal protection. By submitting an article to the International Conference of the Prognostics and Health Management Society, the authors agree to be bound by the associated terms and conditions including the following:

As the author, you retain the copyright to your Work. By submitting your Work, you are granting anybody the right to copy, distribute and transmit your Work and to adapt your Work with proper attribution under the terms of the Creative Commons Attribution 3.0 United States license. You assign rights to the Prognostics and Health Management Society to publish and disseminate your Work through electronic and print media if it is accepted for publication. A license note citing the Creative Commons Attribution 3.0 United States License as shown below needs to be placed in the footnote on the first page of the article.

First Author et al. This is an open-access article distributed under the terms of the Creative Commons Attribution 3.0 United States License, which permits unrestricted use, distribution, and reproduction in any medium, provided the original author and source are credited.

##plugins.themes.bootstrap3.article.main##

##plugins.themes.bootstrap3.article.sidebar##

Abstract

How to Cite

##plugins.themes.bootstrap3.article.details##

Most read articles by the same author(s)