The incident, investigated
What broke, what changed in the minutes before it, which customers felt it, and whether it is still getting worse. Metrics, logs and traces answer in one query rather than across three tools with three query languages.
- Correlate a latency change against every deploy in the window
- Follow one slow request from the edge to the database call
- Compare this hour to the same hour last week, per service