Ops Pillar · Managed Service Ops

Run the managed
application estate.

Managed Service Ops (MSOps) turns application management into one governed control surface. Every managed service, its dependencies, the business processes it carries and the SLAs it owes are understood together, so an incident is measured by real business impact and handed to Sentinel AI to resolve.

Incident command Blast-radius impact Service topology SLA & error budgets Governed process mapping
What is Managed Service Ops

One control surface for the whole managed estate.

Application management usually runs on tickets, tribal knowledge and spreadsheets, disconnected from what each service actually carries for the business. Managed Service Ops models the estate as it really is: every service, its dependencies, the business processes riding on it and the SLAs it owes. So when something breaks, MSOps already knows the blast radius, ranks the incident by real impact, and hands Sentinel AI a clean picture to investigate and resolve.

Incident command

Every incident, ranked by what it costs the business.

MSOps runs a single incident board across the managed estate, with a blast overlay on every active incident. Impact is projected from the same model that powers the What-If Simulator, so triage is ordered by business consequence, not alert timestamp. Only Sherlock-verified closures leave the board, so "resolved" means proven, not assumed.

  • Active, verifying and closed states in one view
  • Blast overlay ranks incidents by business impact
  • Sherlock-verified closure before an incident clears
Live incidentsImpact-ranked
HIGHpayments-gateway degradedactive
MEDnightly-reconcile behind windowactive
CHKauth-service fix verifyingSherlock
OKorders-api verified closedclosed
Service topology

The whole estate, without the hairball.

MSOps never draws the full graph at once. Semantic zoom moves from estate, to service group, to a service's two-hop neighborhood, and search jumps straight to any service. For a full-estate read, a 3D mesh clusters services by group, sizes them by traffic and colours them by risk, with incident services pulsing so the pressure points are obvious.

  • Semantic zoom: estate, group, service neighborhood
  • 3D estate mesh, clustered by group and sized by traffic
  • Search any service to open its 2-hop ego view
Service topologySemantic zoom
Healthy Warning Critical External
What-If simulator

See the ripple before you touch anything.

Fail anything, a service, an external dependency, a node, a namespace or a whole cluster, and MSOps shows the ripple across services, business processes and revenue. The model is replica-aware: losing one of three replicas degrades a service, losing the only replica takes it down. So a node drain or a maintenance window is planned with its true business impact in view, not discovered after the change.

  • Simulate service, dependency, node, namespace or cluster loss
  • Ripple across services, processes and revenue
  • Replica-aware degradation, not just up or down
What-If simulatorBlast radius
Simulated failure
Drop payments-gateway Only replica
Services hit
Several Down
Processes
Checkout At risk
Revenue
Bearing Elevated
Application & process impact

What each application looks like right now.

MSOps derives current application state from live incident blast, KPI pressure and external degradation together, so you see impact without waiting for an incident to be declared. Business processes come from the ProcessOps pillar plus discovered flows, mapped to the services they ride on, with impacted processes ranked first. Click a module, drill into its processes, then the services underneath.

  • Live application state from blast, KPIs and externals
  • Business processes mapped to the services they ride on
  • External dependencies treated as first-class blast sources
Application impactCurrent state
Checkout
Pressured Degraded
Onboarding
Nominal Healthy
Reconciliation
Batch late Watch
Statements
External dep Blast
SLA & error-budget command

The commercial layer of managed operations.

MSOps tracks the contract side of every managed service: SLO attainment month-to-date, error-budget burn, breach forecast and service credits at risk, per contract and per service. It counts verified MTTR only, so Sherlock-closed incidents count and unverified fixes do not. The number you report to a customer is the number you can stand behind.

  • SLO attainment and error-budget burn per contract
  • Breach forecast and service credits at risk
  • Verified MTTR only, Sherlock-closed incidents count
SLA commandMonth to date
SLO attainment
On track Healthy
Error budget
Burning Watch
Breach forecast
Low Stable
Credits at risk
Contained None due
Governed mapping & scoring

Nothing enters the model until a human confirms it.

MSOps mines the call graph and ProcessOps definitions to suggest service-to-process mappings with evidence, but nothing enters the blast model until an operator confirms it. The blast-radius score is an explainable weighted composite, not a black box: tune the factor weights and the ranking preview updates live, so you see exactly what a change does. Onboarding a process declares who owns it, which services it rides on, its fallbacks and the KPIs that define healthy.

  • Evidence-backed mapping suggestions, operator-confirmed
  • Explainable weighted blast score, no black box
  • Process onboarding: owners, services, fallbacks, KPIs
Blast score modelExplainable
Revenue exposureHigh
Process criticalityHigh
Dependents hitMed
Replica headroomMed
External fragilityLow
Operator confirmation required before any edge joins the model
Powered by Sentinel AI

MSOps measures impact. Sentinel resolves it.

Managed Service Ops does more than show you what is happening. The blast model, topology, SLA state and process map all feed Sentinel AI, the intelligence component at the core of Ops Singularity. Sentinel runs the OIAO loop over the managed estate and resolves issues through governed, reversible Action Tickets, ordered by the business impact MSOps already computed.

Sherlock closes the root-cause loop and ProcBot executes the fix, every step explained with citations and fully audited.

1
Observe
MSOps ranks incidents by blast radius across services, processes and SLAs.
2
Investigate
Sentinel AI finds root cause across the estate graph and picks the right procedure.
3
Act
ProcBot executes the approved MOP through a reversible, audited Action Ticket.
4
Optimize
Sherlock verifies the fix, and only then does the incident and its MTTR count.

See Managed Service Ops on your estate.

Book a walkthrough and see incident command, blast-radius impact, SLA governance and autonomous resolution on a managed estate that looks like yours.