Root Cause Investigation Best Practices Guide
Roland Duphily, Harold Harder, Rodney Morehead, Joe Haman, Helen Gjerde, Susanne Dubois, Thomas T. Stout, David A. Ward, Thomas Reinsel, Jim Loman
Abstract
Roland Duphily, Harold Harder, Rodney Morehead, Joe Haman, Helen Gjerde, Susanne Dubois, Thomas T. Stout, David A. Ward, Thomas Reinsel, Jim Loman
Abstract
Abstract : A multi-discipline team composed of representatives from different organizations in national security space has developed the following industry best practices guidance document for conducting consistent and successful root cause investigations. The desired outcome of a successful root cause investigation process is the conclusive determination of root causes, contributing factors and undesirable conditions. This provides the necessary information to define corrective actions that can be implemented to prevent recurrence of the associated failure or anomaly. This guide also addresses the realities of complex system failures and technical and programmatic constraints in the event root causes are not determined. The analysis to determine root causes begins with a single engineer for most problems. For more complex problems, identify an anomaly investigation lead / team and develop a plan to collect and analyze data available before the failure, properly define the problem, establish a timeline of events, select the root cause analysis methods to use a long with any software tools to help the process. This guide focuses on specific early actions associated with the broader Root Cause Corrective Action (RCCA) process. The focus is the early root cause investigation steps of the RCCA process associated with space system anomalies during ground testing and on -orbit operations that significantly impact the RCA step of the RCCA process. Starting with a confirmed significant anomaly we discuss the collection and classification of data, what determines a good problem definition and what helps the anomaly investigation team select methods and software tools and also know w hen they have identified and confirmed the root cause or causes.
OpenAlex reports 12 citations for this work. Citation counts describe recorded attention and do not establish research quality.
A contribution statement is not available in the OpenAlex record.
Method details are not available in the OpenAlex metadata.
Findings are not separately available in the OpenAlex metadata.
Limitations are not available in the OpenAlex metadata.
Application details are not available in the OpenAlex metadata.
Abstract : A multi-discipline team composed of representatives from different organizations in national security space has developed the following industry best practices guidance document for conducting consistent and successful root cause investigations. The desired outcome of a successful root cause investigation process is the conclusive determination of root causes, contributing factors and undesirable conditions. This provides the necessary information to define corrective actions that can be implemented to prevent recurrence of the associated failure or anomaly. This guide also addresses the realities of complex system failures and technical and programmatic constraints in the event root causes are not determined. The analysis to determine root causes begins with a single engineer for most problems. For more complex problems, identify an anomaly investigation lead / team and develop a plan to collect and analyze data available before the failure, properly define the problem, establish a timeline of events, select the root cause analysis methods to use a long with any software tools to help the process. This guide focuses on specific early actions associated with the broader Root Cause Corrective Action (RCCA) process. The focus is the early root cause investigation steps of the RCCA process associated with space system anomalies during ground testing and on -orbit operations that significantly impact the RCA step of the RCCA process. Starting with a confirmed significant anomaly we discuss the collection and classification of data, what determines a good problem definition and what helps the anomaly investigation team select methods and software tools and also know w hen they have identified and confirmed the root cause or causes.
Key concepts: Root cause analysis, Root cause, Root (linguistics), Timeline, Process (computing), Engineering, Plan (archaeology), Risk analysis (engineering)