4 ms·
IMO - log every 5xx (4xxs too), ensure your logs/error messages are groupable by message (e.g. keep variables like ids in a metadata object separate from the me
by yashap 3y ago
IMO - log every 5xx (4xxs too), ensure your logs/error messages are groupable by message (e.g. keep variables like ids in a metadata object separate from the message), and keep a close eye on errors grouped by message and HTTP status code. Then fix any 5xx error that’s too common - IMO, for most cases, this means more than a handful of the same 5xx error most days. And of course, don’t just rely on ppl happening to look at dashboards, alert/page on significant 5xx volume.
But there’s no need to send every single error to Slack individually, just keep an eye on counts by error message.