by Brent Voelker
June 23, 2026
In our last post on High Reliability Organizations (HROs), we discussed a tension between two major tenants, Sensitivity to Operations and Cultivation of Expertise. This week, we extend our prior discussion of Vigilance for Failure by relating it to another hallmark of HROs, Resilience. In product development contexts, both principles are operationalized through Novellum’s unique change management process implementation.
A strong culture of vigilance for failure depends on leadership responsive to the rapid flow of emerging information. Employees must be confident that raising concerns will promptly bring needed attention and benefit the organization. The more quickly leaders show up and help solve problems, the more team members embrace the process and make reports. In this way, vigilance for failure becomes a daily operational discipline rather than a periodic management exercise. With proper planning and resources, reports of anomalies are not causes of delay but valuable interruptions to homogenized thinking and sources of critical operational data.
Novellum operationalizes these principles through requirements-based compliance assessment, formal problem reporting, and change management processes. Requirements that lack compliance evidence are actively identified and managed. Any known or anticipated failure to comply with a requirement is reported as a non-compliance, and every problem report is reviewed without prejudgment or premature dismissal. The organization adopts a posture of zero tolerance for unresolved non-compliance, not because perfection is expected, but because unresolved discrepancies present a risk of a serious product defect downstream. Such practices ensure that weak signals are elevated rather than normalized and that organizational learning occurs before failures propagate through the product or program.
Vigilance for failure also requires resisting the tendency to focus on symptoms and papering over deeper causes. Accordingly, Novellum’s change management process emphasizes rigorous root cause analysisbefore solution development begins. Root causes must be identified and understood before corrective actions are authorized, and both product and process root causes are examined. By resisting superficial fixes and insisting on evidence-based diagnosis, the change process converts individual problem reports into deeper institutional knowledge. Failures become opportunities to improve both the product and the work processes used to create it. Reliability grows out of the continuous refinement of organizational understanding and improved workflow.
By embedding the problem reporting and change management process in daily activity, the organization enacts Resilience, which concerns the collective ability to adapt, recover, and continue functioning when confronted with unexpected conditions. Weick and Sutcliffe describe resilience as the capability to absorb disturbances, recombine expertise, improvise effectively, and restore performance without catastrophic degradation. Importantly, resilience does not assume that failures can be eliminated. Instead, resilient organizations acknowledge that surprises are inevitable in complex systems and therefore build the capacity to respond effectively when those surprises occur.
The Novellum change management process operationalizes resilience through a structured but adaptable framework. The gated change board process requires disciplined progression from problem identification through root cause analysis, solution development, validation, and verification. This rigor ensures that corrective actions are not ad hoc responses but measured interventions informed by evidence. Yet the process also allows requirements, compliance status, and technical understanding to evolve continuously as new information emerges. Problems are not viewed as disruptions to a static baseline; they are expected features of a learning system. Resilience emerges not from avoiding change but from managing change deliberately and effectively.
Several criteria specifically reinforce organizational adaptability. Interdisciplinary change boards bring together diverse expertise to evaluate problems and develop comprehensive corrective actions. Assigned tasks, due dates, objective evidence, and accountability mechanisms ensure that effective solutions move efficiently from concept to implementation. Rapid notification and response requirements reduce the latency between problem discovery and corrective action. Together, these practices create an organization capable of mobilizing knowledge quickly when confronted with uncertainty. In HRO terms, resilience is supported by the ability to bring the right expertise to the problem at the right time and to translate learning into action without delay.
Metrics and leadership engagement further strengthen resilience. The change management process establishes standards for cycle time, tracks aging problem reports, identifies overdue actions, and provides visibility into organizational performance. Leadership and frontline managers routinely review these metrics and intervene to remove obstacles, allocate resources, and sustain momentum. This operating philosophy aligns with the HRO emphasis on maintaining awareness of operational realities rather than relying solely on plans, assumptions, or the uncoordinated actions of individual experts. Resilience therefore becomes measurable and manageable rather than remaining an abstract cultural aspiration.
Taken together, vigilance for failure and resilience form a mutually reinforcing system. Vigilance creates the flow of information necessary to detect emerging threats, while resilience provides the organizational capability to respond effectively once those threats are identified. The Novellum change management process institutionalizes both principles through accessible problem reporting, objective review, rigorous root cause analysis, interdisciplinary collaboration, evidence-based corrective action, and continuous leadership oversight. In doing so, it transforms two foundational concepts from HRO theory into concrete behaviors and management practices. The result is an organization that neither ignores failures nor becomes consumed by them, but instead uses every anomaly, non-compliance, and unexpected outcome as an opportunity to increase reliability, strengthen learning, and improve performance.
Let us show you how to build vigilance and resilience in your organization. Call us today.
