4 ms·
It handles millions, but it can be a lot faster to just pipe output from tar through the ssh connection.
by perbu 5y ago
It handles millions, but it can be a lot faster to just pipe output from tar through the ssh connection.
- hotpotamus 5y agoGot any benchmarks/write ups on the subject? I did a bit of testing myself a long time ago and basically the answer just ended up being to use rsync because any differences were marginal. That said, I didn't test with millions of files.
- perbu 5y agoI think perhaps this was a bigger issue back in the day when we were using rotating harddisks. In those day doing a seek would be a lot slower than doing a write. Today seeks are mostly instants, so maybe my experience isn't valid anymore.
- porker 5y agoYour experience is valid today but for a different reason: if you're comparing millions of tiny files there's a lot of back and forth. If you're streaming a single archive, it only checks if that single file has been modified. Like everyone here I've no benchmarks but have got burned trying to rsync around too many small files.
- nani_o 5y agoI got into the « lot of small files » case, and I ended up generating a file list, split it and feed multiple rsync instances with xargs.