← Back to context

Comment by anslopic

12 hours ago

[flagged]

this is insanely misleading, you can't run anything close to current chatgpt or gemini on local hardware

  • I am running GLM 5.3 across 2x DGX Sparks and was doing comparisons and it absolutely can beat Gemini. Yesterday it corrected a poor Fable 5 response even

    • I should have clarified, you can’t run it on their local hardware, which is a 16gb gpu. Glm-5.3 and k3 are of course near the frontier.

  • GLM 5.3 Flash? Qwen 3.8 Flash Next? I believe those both are as good as the best Gemini, competitive with Terra.

    • Yes they are quite good, but are not able to run on a 16GB RX 9070.

      Quantized Qwen 3.8 Flash Next could maybe run eventually on that card with a highly optimized inference engine that dynamically caches the hottest layer experts. Even then you run into some hard limits.

  • You are insanely non-technical then. Yes you can. Skill issue

    You probably pay $200/mo for text-to-text!

> You have good enough hardware to run good models comparable with Gemini and ChatGPT.

That is at best misleading and at worst outright misinformation.

  • If you have 2TB of VRAM you can’t run one of the big models which are comparable?

    • The post was replying to someone with 16GB. (And also: no, even the best open weight models are not as good as what you can use on your ChatGPT subscription. They’ve gotten a lot better, but not that much better.)

      5 replies →