5 ms·
Did you try the MLX model instead? In general MLX tends provide much better performance than GGUF/Llama.cpp on macOS.
by smcleod 6mo ago
Did you try the MLX model instead? In general MLX tends provide much better performance than GGUF/Llama.cpp on macOS.