4 ms·
Full weights aren't available, only Q2, Q4, Q5 quants via miqudev.
by MallocVoidstar 3y ago
Full weights aren't available, only Q2, Q4, Q5 quants via miqudev.
- throwaway9274 3y agoUnquantized model is here: https://huggingface.co/152334H/miqu-1-70b-sf https://huggingface.co/152334H/miqu-1-70b-sf This strikes me as less a leak and more clever marketing from Mistral.
- MallocVoidstar 3y agoThat isn't unquantized, it's de-quantized. They went from Q5 to fp16 for use in Pytorch instead of the GGUF ecosystem.
- Taek 3y agoI never thought people would be upscaling models by increasing quantization precision. The rationale makes sense bit its also a goofy outcome.
- nullc 3y agoYou should be able to upscale and fine tune to recover performance, I suppose! Clearly we should train a diffusion model to denoise the weights of LLM transformer models. Yo dawg.
- throwaway9274 3y agoYes, that’s correct. Good correction.