4 ms·
this is insanely misleading, you can't run anything close to current chatgpt or gemini on local hardware
by hhh 1mo ago
this is insanely misleading, you can't run anything close to current chatgpt or gemini on local hardware
- colingauvin 1mo agoGLM 5.3 Flash? Qwen 3.8 Flash Next? I believe those both are as good as the best Gemini, competitive with Terra.
- notnullorvoid 1mo agoYes they are quite good, but are not able to run on a 16GB RX 9070. Quantized Qwen 3.8 Flash Next could maybe run eventually on that card with a highly optimized inference engine that dynamically caches the hottest layer experts. Even then you run into some hard limits.
- sharms 1mo agoI am running GLM 5.3 across 2x DGX Sparks and was doing comparisons and it absolutely can beat Gemini. Yesterday it corrected a poor Fable 5 response even
- hhh 1mo agoI should have clarified, you can’t run it on their local hardware, which is a 16gb gpu. Glm-5.3 and k3 are of course near the frontier.
- NoYouAre 1mo ago[flagged]