Every deploy, just after rollout
Watch the canary
Read the golden metrics and flag any anomaly on any server. Watch the canary while it is at 1% and tell me before we widen it.
Lands in #deploys in Slack
Compare against the last good window
Compare error rate and p99 against the same window before this deploy. If anything regressed, name the commit that did it.
Lands in #deploys in Slack
Check the neighbours
Check that every service depending on this one is still healthy, not just the one that shipped.
Lands in PagerDuty incident note