Skip to content

When systems fail: why orchestration will define the UK’s resilience edge

20 May 20264 min read
Guest Insights
When systems fail: why orchestration will define the UK’s resilience edge

Emily Wells-Cole

Principal, Netcompany

For years, digital resilience has been treated primarily as a technology challenge: protecting systems, strengthening cyber defences, and maintaining uptime, but this view is increasingly outdated. As digital infrastructure becomes more embedded in essential services, resilience must be understood not just as system protection, but as the ability to sustain operations under any conditions.

For the UK, this shift is particularly significant. Critical national infrastructure (CNI) – from energy and transport to water and communications – depends on a web of interconnected systems and external dependencies. When disruption occurs, the consequences ripple across the economy and society.

From protection to continuity

Traditional approaches to resilience are built around prevention and containment. When an incident occurs, systems are isolated, access is restricted, and operations are paused to limit damage. While necessary, this approach can come at a cost: the loss of service continuity.

A more mature view of resilience starts from a different premise, that disruption is inevitable. Moving from “how do we stop this entirely?”, to “how do we continue to operate when it happens?”

This requires organisations to prioritise outcomes over assets, and means ensuring essential services can continue – even if parts of the underlying infrastructure are unavailable, compromised, or need to be rebuilt.

The growing importance of external dependencies

One of the most important – and overlooked – aspects of resilience is the role of external factors. Digital systems don’t operate in isolation, they depend on power, connectivity, and complex supply chains. They’re also exposed to environmental risks and physical threats, from extreme weather to unauthorised drones.

Yet many resilience strategies remain focused on the internal, with limited visibility of the broader conditions that can impact operations. For UK organisations, particularly those within CNI, this is becoming more acute. Climate-related risks, energy volatility, and geopolitical uncertainty are all increasing the likelihood of disruption originating outside the traditional IT perimeter. Understanding these dependencies is now a critical component of resilience planning.

Orchestration as the foundation

Addressing this challenge doesn’t require organisations to rebuild their technology estates from scratch. Instead, it requires a different approach: orchestration.

Orchestration is about connecting systems, data, and decision-making processes so that organisations can respond to disruption in a coordinated and timely way. It enables different teams – from IT and operations to risk and executive leadership – to act on a shared understanding of what’s happening and what needs to happen next.

Crucially, orchestration extends beyond internal systems to incorporate signals from the external environment – bringing together insights on infrastructure dependencies, environmental conditions, and physical risks. By combining these perspectives, organisations can move from reactive responses to proactive, scenario-based planning. They can anticipate how different types of disruption might unfold and define coordinated responses.

Designing for rebuild, not just recovery

Another critical shift is the move from recovery to rebuild. Modern, cloud-based infrastructure has made it technically feasible to recreate environments quickly and at scale. However, many organisations are still structured around the assumption that systems should be preserved and restored, rather than replaced when necessary.

Designing for rebuild means accepting that in some scenarios the fastest and safest path to continuity is to stand up new environments and switch operations over. That depends on having clear processes, predefined scenarios, and the ability to coordinate action across multiple systems and teams.

Again, orchestration plays a central role, enabling organisations to execute this quickly and with confidence.

A UK-wide opportunity

For the UK, there’s an opportunity to lead in how resilience is defined and delivered. With growing focus on cyber resilience, infrastructure investment and wider digital transformation, there’s increasing recognition that a more joined-up approach is needed across industry and government – one that treats digital infrastructure as a connected system, rather than standalone assets.

The next step is to move from discussion to execution and embed orchestration-led resilience into strategy and delivery. That means broadening resilience thinking beyond IT to include external dependencies, from physical infrastructure through to environmental and operational factors.

It also requires closer alignment between technology, operations and risk teams, so decisions are based on a shared view. Alongside this, organisations need better real-time visibility across internal and external environments to enable faster, more coordinated responses. Resilience needs to be designed in from the outset.

Resilience as a competitive advantage

Resilience is often framed as a defensive priority, something organisations must invest in to avoid failure. But it is increasingly becoming a competitive advantage.

Organisations that can adapt quickly, maintain continuity, and respond decisively are better positioned to serve customers, protect reputations, and navigate uncertainty.

In a world where disruption is no longer the exception but the norm, resilience by design, underpinned by orchestration, will define the organisations that succeed.