Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
rajatgupta314
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
4 ms
·
1.
▲
by
rajatgupta314
1y ago
Is this the full weight model or quantized version? The GGUFs distributed on Hugging Face labeled as MXFP4 quantization have layers that are quantized to int8 (q8_0) instead of bf16 as suggested by OpenAI. Example looking at blk.0.attn_k.we