3 ms·
This model will run on Ollama, llama.cpp, and other tools: ollama run mistral or for llama.cpp, thebloke has uploaded the GGUF models here: https://huggingfa
by mchiang 3y ago
This model will run on Ollama, llama.cpp, and other tools:
ollama run mistral
or for llama.cpp, thebloke has uploaded the GGUF models here:
https://huggingface.co/TheBloke/Mistral-7B-v0.1-GGUF/tree/main https://huggingface.co/TheBloke/Mistral-7B-v0.1-GGUF/tree/ma...
and you can run it
really looking forward to the chat fine-tuned models that doesn't seem to be available yet.
- brucethemoose2 3y agoOh, that means its a llama architecture model! Is the tokenizer the same? It may "work" without actually working optimally until llama.cpp patches it in. And the instruct model was just uploaded.