Resilience engineered, not assumed.
Many organisations have a resilience problem they haven't understood yet. The same incidents keep recurring. We help you understand why, and fix it.
You'll recognise this.
Start with diagnosis,
then build from there.
First we find out why the same incidents keep happening. Then we fix what's causing them.
Your organisation keeps having the same incidents despite doing incident reviews, chaos engineering experiments, architecture and service reviews. Something organisational is broken but you can't articulate why.
We embed with your teams to see how the work really happens. We watch how teams interact, what pressures they're under, and where practice and policy don't match.
- Embedded observation of your practices (incident reviews, GameDays, ORRs, chaos experiments, etc.)
- Interviews across engineering, operations, and leadership
- Analysis of your feedback loops and organisational patterns limiting resilience
- Written report with prioritised, actionable recommendations
- Readout with your leadership team
You want to change how your organisation learns in order to improve its resilience capability.
We work with your teams for six to twelve months, often on multiple projects at once depending on how fast we can change your organisation (it takes time). Each work stream is done so your team can take over once we leave.
We first focus on operational excellence and resilience to reduce the firefighting and free capacity for proactive work. When a topic needs a specialist, we bring one in from the Collective.
- Incident analysis
- Operational readiness reviews
- Change management
- On-call and incident command
- Observability and tooling
- Availability measured around critical customer workflows
- Chaos engineering and GameDays
- Ownership and roles
- A shared resilience vocabulary
- Leadership support, up to the CEO
No prescriptive checklists. We study your organisation's culture first
Founder-led. The person you talk to is the person doing the work
Diagnosis before prescription. We don't sell solutions to problems we haven't seen
What you need to hear, not what you want to hear
The work speaks
for itself.
Resilium Labs is led by Adrian Hornsby, a former Principal Engineer at AWS and author of Why We Still Suck at Resilience, now working alongside a curated collective of specialists. Here is what people who have worked with him say.
Writing on resilience,
organisations, and AI.
Free tools. No sign-up.
Each one surfaces a pattern most teams never measure.
Ready to find out what's actually broken?
The best first step is a conversation to understand your current challenges and resilience goals. We'll help you figure out which step in the journey makes sense for your organisation.
Book a call