3 ms·
What surprised me here is how close Aurora got. That's some magic right there. I've always held that if you want resilience, you just cannot rely on local sto
by regularfry 1y ago
What surprised me here is how close Aurora got. That's some magic right there.
I've always held that if you want resilience, you just cannot rely on local storage. No matter how many times you've got data replicated locally, you're still at risk of the whole machine failing - best case falling off the network, worst case trashing all its disks in some weird failure state as the RAID firmware decides today is the day to Just Not. And while you might technically still be able to recover the data, you're still offline.
You just need your data to be off the machine already when that happens. Not to say that all access needs to go over the network - local caching ought to go a long way here - but the default should be to switch to another machine and recycle the failed one.
Relevant to the article, this is independent of the speed and reliability of the actual hardware. It was true in 2010, it's true now.