The resilience block
Every important service, run as one loop - the blast radius mapped and the clock managed before the scramble.
The pain
“When something breaks we scramble - no one can see the blast radius, and the regulatory clock starts before we have understood what happened.”
An incident is the moment the firm’s dependencies all matter at once, and it is exactly the moment no one can see them. The error rate ticks up in monitoring, a complaint spike forms in service, a supplier’s failure lands in ops - each in its own place, none joined into “which important service is now at risk, and what does it touch.” So the response is a manual scramble to reconstruct the map while the impact tolerance erodes and the notification clock, which started the moment the service breached, runs unmanaged. In regulated firms this is the important business service and its impact tolerances; in distribution, platform and logistics uptime; in legal, matter-critical systems.
What the block is
The object is the important business service and its incidents - one continuous picture per service, from normal operation through disruption, response, recovery and control update. Everything it produces feeds one live record: service health and incidents, impact tolerances, dependencies, notifications and control changes. A model reads it permanently. Your monitoring, service-management and incident systems stay exactly where they are - the block reads from them, it does not replace them.
The loop in action
Each output states what it saw, why it matters, what it recommends, and how confident it is - every fact linked back to its source, in an order an operations lead can read and a regulator can examine.
You set the dial
Every capability has four positions: observe - it reads silently and measures its own accuracy; recommend - it surfaces the case and a proposed action to a named owner; draft - it prepares the action for a human signature; act - it executes within agreed bounds and logs everything.
Everything starts at observe. Promotion is earned on the measured record and reversed instantly if it doesn’t hold. Customer and regulatory notifications, and any containment that touches production, stay behind a named human - permanently, if that is your policy. The mechanics are the immune gate described in our framework.
How it lands
It starts narrow: one important business service, the block at observe against your own past outcomes. Nothing is replaced, nothing is migrated - the first signals flow from the monitoring and incident systems you already run. From there it is a bounded engagement - blueprint, build, handover - and your team owns the result outright: the signal map, the judgement rules, the dial settings, the working loop.
The backtest
Take three past incidents: one contained within tolerance, one that surprised you, and one where the regulatory clock started before you understood what had happened. What did your own systems know, and when?
That is the first question the diagnostic answers. Take the ten-minute diagnostic, or write to hello@somai.studio.
Runs on our framework · related reading: Most AI spend never reaches the P&L