4 ms·
Self plug: run llama.cpp as an inference server on a spot instance anywhere: https://cedana.readthedocs.io/en/latest/examples.html#running-llama-cpp-inference h
by nravic 3y ago
Self plug: run llama.cpp as an inference server on a spot instance anywhere: https://cedana.readthedocs.io/en/latest/examples.html#running-llama-cpp-inference https://cedana.readthedocs.io/en/latest/examples.html#runnin...
- dabei 3y agoLooks cool, joined the waitlist.