4 ms·
Have you tested this on big models involving multi-gpu communication, or any plans?
by papersnake 4y ago
Have you tested this on big models involving multi-gpu communication, or any plans?
- ipiszy 4y agoFor now it's for single GPU inference only.