4 ms·
AWS also say they do something interesting: > When adding jitter to scheduled work, we do not select the jitter on each host randomly. Instead, we use a consis
by ramchip 2y ago
AWS also say they do something interesting:
> When adding jitter to scheduled work, we do not select the jitter on each host randomly. Instead, we use a consistent method that produces the same number every time on the same host. This way, if there is a service being overloaded, or a race condition, it happens the same way in a pattern. We humans are good at identifying patterns, and we're more likely to determine the root cause. Using a random method ensures that if a resource is being overwhelmed, it only happens - well, at random. This makes troubleshooting much more difficult.
https://aws.amazon.com/builders-library/timeouts-retries-and-backoff-with-jitter/ https://aws.amazon.com/builders-library/timeouts-retries-and...
- cpeterso 2y agoI’ve read a suggestion to use prime numbers for retry timers to reduce the chance of multiple timers synchronizing if they have common factors. I don’t know if that’s a real concern, but it wouldn’t hurt to pick a random prime number instead of some other random number.
- password4321 2y agoIIS picked 29 hours as the smallest prime over 24. https://serverfault.com/questions/348493/why-does-the-iis-worker-process-recycle-every-29-hours-and-not-every-24-hours/348502#348502 https://serverfault.com/questions/348493/why-does-the-iis-wo...