SLOPager

SLOPager is a web app (with optional Slack/Teams and mobile push) that turns noisy monitoring into incident pages tied to real user impact. Instead of paging on raw CPU, error-rate spikes, or single-service alerts, it pages when an SLO burn-rate or customer-journey signal crosses a threshold (checkout failures, login latency, API availability). It ingests metrics from Datadog/Prometheus and traces from OpenTelemetry, then correlates them to a small set of business-critical “golden flows.” The product focuses on fast setup: pick a flow template, map endpoints, and define SLOs; the system generates routing rules, dedupes alerts, and opens an incident with context (recent deploys, top failing endpoints, suspected blast radius). It also produces post-incident summaries and tracks alert quality (false pages, time-to-ack, time-to-mitigate) so teams can iteratively reduce noise.

← Back to idea list