3 ms·
Should be a bit faster if you run an MLX version of the model with LM Studio instead. Ollama doesn't support MLX. Qwen3-Coder is in the same ballpark and maybe
by eli 11mo ago
Should be a bit faster if you run an MLX version of the model with LM Studio instead. Ollama doesn't support MLX.
Qwen3-Coder is in the same ballpark and maybe a bit better at coding
- ZeroCool2u 11mo agoLM Studio will run dynamic quants from Unsloth too. Much nicer than Ollama.