3 ms·
I'd recommended using llama.cpp and The Bloke's GGUF version of this model! https://github.com/ggerganov/llama.cpp/ https://github.com/ggerganov/llama.cpp/ htt
by schmeichel 3y ago
I'd recommended using llama.cpp and The Bloke's GGUF version of this model!
https://github.com/ggerganov/llama.cpp/ https://github.com/ggerganov/llama.cpp/
https://huggingface.co/TheBloke/MonadGPT-GGUF https://huggingface.co/TheBloke/MonadGPT-GGUF