← Back to context

Comment by mrngld

3 days ago

Depends on what level of intelligence you're wanting to use. A vanishingly small number of people can or would want to go to the hardware expense of running something like GLM 5.3 Flash, much less something like K3.

And if you want Astra/Fable/Opus frontier level, then there's no option at all.

But if you don't need that, or you don't need speed... That opens up the discussion. I've been impressed even with how Siri's been doing with the Apple Foundation Models in MacOS/iOS 27 given how small they are.

Edit: I can't even fully spec the M5 Ultra Mac Studio you'd need for GLM5.3 Flash since 512GB isn't available yet, but it's already at $9500 for 256GB RAM.

> A vanishingly small number of people can or would want to go to the hardware expense of running something like GLM 5.3 Flash, much less something like K3.

It's probably worth letting the user specify their actual costs in such a tool. I run a Framework Desktop 128GB that I bought before memory prices got crazy; the current retail price is almost double what I actually paid a year ago.

  • Dear person,

      I am sorry if this looks out of the blue, but I was trying to answer to an old thread, and HackerNews seems to lock old threads and doesn't allow for private messages (unless I am too dumb to figure it out, which is always a possibility).
    

    Anyway, I wanted to reply to the very thoughtful comment you left my after I inquired about your responsibilities at a co/op(1) and I wanted to tell you that I am very proud of you and that you are the type of "hacker" that I aspire to be :)

    I hope this doesn't come as too forward and I wish you anything but the best.

    (1) : https://news.ycombinator.com/item?id=48383220