3 ms·
Something like ollama run hf.co/ngxson/GLM-4.7-Flash-GGUF:Q4_K_M It's really fast! But, for now it outputs garbage because there is no (good) template. So
by ljouhet 9mo ago
Something like
ollama run hf.co/ngxson/GLM-4.7-Flash-GGUF:Q4_K_M
It's really fast! But, for now it outputs garbage because there is no (good) template. So I'll wait for a model/template on ollama.com
- jmorgan 9mo agoIt's available (with tool parsing, etc.): https://ollama.com/library/glm-4.7-flash https://ollama.com/library/glm-4.7-flash but requires 0.14.3 which is in pre-release (and available on Ollama's GitHub repo)