We stopped waking up the whole team.
Before Relay every alert went to a channel and whoever was awake took it. Now one person is paged, with the runbook. Our on-call survey went from dread to fine.
Be on call less.
We invite teams every two weeks, smallest first. Tell us where your alerts come from and what the worst part of the night is.
Datadog, Grafana, Sentry, CloudWatch or a webhook. Ten minutes, no agent.
Who owns what, in a file in your repo. Relay reads it and keeps it current.
Start from ours. Relay opens it the next time the alert fires.
The right person is paged, the status page speaks, and the timeline is kept. You read it in the morning.
Beta teams pay nothing until general availability, and Starter stays free after it.
What they say after a quarter.
We stopped waking up the whole team.
Before Relay every alert went to a channel and whoever was awake took it. Now one person is paged, with the runbook. Our on-call survey went from dread to fine.
The status page is finally honest.
We used to update it after the fact. Now it is written from inside the incident and customers see it before they notice anything. Support tickets during incidents fell by half.
Postmortems take an hour.
The timeline is already there, with the clock. We argue about what to change instead of what happened.
Be on call less.
Teams are invited every two weeks, smallest first. Beta teams pay nothing until general availability.