4 ms·
Are you aiming for Nvidia hardware with rust-cuda, or looking to integrate with non-Nvidia hardware?
by RRRozie 2y ago
Are you aiming for Nvidia hardware with rust-cuda, or looking to integrate with non-Nvidia hardware?
- zackangelo 2y agoWe used candle[0], which uses cudarc and the metal crate under the hood. That means we run on nvidia hardware in production and can test locally on macbooks with smaller models. I would certainly like to use non nvidia hardware but at this point it's not a priority. The subset of tensor operations needed to run the forward pass of LLMs isn't as large as you'd think though. [0] https://github.com/huggingface/candle https://github.com/huggingface/candle