Comment by hhh
13 hours ago
this is insanely misleading, you can't run anything close to current chatgpt or gemini on local hardware
13 hours ago
this is insanely misleading, you can't run anything close to current chatgpt or gemini on local hardware
I am running GLM 5.3 across 2x DGX Sparks and was doing comparisons and it absolutely can beat Gemini. Yesterday it corrected a poor Fable 5 response even
I should have clarified, you can’t run it on their local hardware, which is a 16gb gpu. Glm-5.3 and k3 are of course near the frontier.
GLM 5.3 Flash? Qwen 3.8 Flash Next? I believe those both are as good as the best Gemini, competitive with Terra.
Yes they are quite good, but are not able to run on a 16GB RX 9070.
Quantized Qwen 3.8 Flash Next could maybe run eventually on that card with a highly optimized inference engine that dynamically caches the hottest layer experts. Even then you run into some hard limits.
[flagged]