3 ms·
Not only the quantization, but what’s available via ollama is magistral-small (for local inference), not the -medium variant.
by samtheprogram 1y ago
Not only the quantization, but what’s available via ollama is magistral-small (for local inference), not the -medium variant.