3 ms·
Clone llama.cpp project from GitHub. Download LLM models from HuggingFace. TheBloke user posts lot of models in GGUF format. GPU is not needed to run these. 13B
by svjatoslav 3y ago
Clone llama.cpp project from GitHub. Download LLM models from HuggingFace. TheBloke user posts lot of models in GGUF format. GPU is not needed to run these. 13B models should be fast modern computer CPU. llama.cpp offers batch mode, interactive chat mode and also web server mode.