02 — Capability · BC-1230.20
Observability
See what production is doing from the outside in — metrics, logs, traces and synthetic checks — well enough to answer new questions during an incident without shipping new code.
- Monitoring
- Telemetry
- Production Visibility
In scope
- Metrics, logging and distributed tracing pipelines
- Alert design and routing
- Synthetic and real-user monitoring of customer journeys
Out of scope
- Business analytics on product usage (see BC-1260)
- Security event monitoring (see BC-760)
Realized by · 0
- No product in the catalog yet.
Used in · 1
Build it · 5
- pattern Canary Release via BC-1230.20.20
- pattern Distributed Tracing via BC-1230.20.10
- pattern Structured Logging via BC-1230.20.10
- stack Open observability stack
- book Observability Engineering
Decomposes into · 4
- BC-1230.20.10Telemetry CollectionInstrument services and collect metrics, logs and traces with consistent naming and retention across the product.
- BC-1230.20.20Alerting & Signal DesignAlert on symptoms customers feel, route to the right responder, and prune alerts that nobody acts on.
- BC-1230.20.30Synthetic & Real-User MonitoringExercise key customer journeys continuously from outside and measure what real users experience, so outages are caught from their side.
- BC-1230.20.40Dashboards & Service Health ViewsPresent service health per team and per customer-facing surface so a responder can orient in seconds.