3 ms·
failure of 2 servers is a risk for the proposed design. for a v2, the system could use Reed-Solomon on the backup-server to add entropy and support the simultan
by nanch 13y ago
failure of 2 servers is a risk for the proposed design. for a v2, the system could use Reed-Solomon on the backup-server to add entropy and support the simultaneous failure of N servers.
Yes, a distributed design with the backup blocks distributed across servers would be adequate as well, even without the RS blowup factor.
- chris_va 13y agoThis is your best bet: Distributing blocks to all servers evenly. Using RS to encode your blocks to limit storage space requirements, and make recovery obvious. Also, make sure you don't buy all of your hard drives from the same end manufacturer. Most of these types of systems assume random hard drive failure... turns out drives can have highly correlated failures.