Built from operational experience.
Recoverable exists because most reliability problems aren't tooling problems — they're lifecycle problems. The causes are upstream, the consequences are downstream, and nobody's looking at both.

Built by someone who's been on every side of an incident.#
Recoverable is built by Hugo van der Horst. Seven years at Adyen — spanning Technical Support, Platform Reliability, Site Reliability Engineering, and Platform Observability — including leading the EMEA 24/7 on-call support team for one of Europe's largest payment platforms.
Operational scale
Led incident response across a platform processing hundreds of billions in annual payment volume. Understands what reliability means when minutes of downtime have real financial impact.
Every side of an incident
Has led the team responding to incidents, owned the incident management programme, and built the observability tooling that detects them.
Evidence-first
The assessment draws directly from operational experience — not frameworks pulled from a textbook. Every recommendation is grounded in what actually works at scale.
Prepare. Respond. Learn.#
The assessment is structured around three phases of the incident lifecycle. Most teams over-invest in one phase and under-invest in the others.
Before things break
Change validation, blast radius awareness, alert signal quality. Are you catching problems before they reach production — or just hoping your tests are enough?
When things are breaking
Ownership clarity, runbook quality, escalation paths, on-call coverage. When something breaks at 2am, does the right person know — and know what to do?
After things are fixed
Post-mortem quality, action item follow-through, learning propagation. Are you actually getting better — or just documenting that you didn't?
"Trust is hard to build and easy to lose. Reliability is how you keep it."
Find out where your reliability breaks down.
No commitment, no pitch. A 30-minute conversation to understand whether the problem you're facing is one we can help with.