Comment by hhh

13 hours ago

this is insanely misleading, you can't run anything close to current chatgpt or gemini on local hardware

I am running GLM 5.3 across 2x DGX Sparks and was doing comparisons and it absolutely can beat Gemini. Yesterday it corrected a poor Fable 5 response even

  • I should have clarified, you can’t run it on their local hardware, which is a 16gb gpu. Glm-5.3 and k3 are of course near the frontier.

GLM 5.3 Flash? Qwen 3.8 Flash Next? I believe those both are as good as the best Gemini, competitive with Terra.

  • Yes they are quite good, but are not able to run on a 16GB RX 9070.

    Quantized Qwen 3.8 Flash Next could maybe run eventually on that card with a highly optimized inference engine that dynamically caches the hottest layer experts. Even then you run into some hard limits.

You are insanely non-technical then. Yes you can. Skill issue

You probably pay $200/mo for text-to-text!