3 ms·Lessons learned scaling LLM training and inference with RDMA (2024)1 points by ArcVRArthur 2y ago