4 ms·
Could this work with llama.cpp, since it’s the engine behind Ollama? I usually build llama.cpp from source and download quantized (GGUF) models from Huggingfac
by car 2y ago
Could this work with llama.cpp, since it’s the engine behind Ollama?
I usually build llama.cpp from source and download quantized (GGUF) models from Huggingface, haven’t used Ollama this far.