Comment by epolanski

6 days ago

I don't see a difference in capabilities between k3 and fable, but k3 is slow and expensive. Burned through my monthly plan in 3 days.

I'd guess the slowness is mostly due to there currently being only one provider, Moonshot AI. And they are overwhelmed with demand.

Let's judge the speed of the model when its weights are released and every inference provider on the planet offers it, so demand can spread out a bit.

It's the same topic with token budget comparisons and subscription pricing - don't people understand that this doesn't really matter for open weights models? The pricing is going to be determined by the inference providers, and until they had a chance to evaluate the model on their infra and set token prices accordingly, one doesn't really have anything tangible to compare with other open models nor with closed ones.

  • Even if hardware capacity increases, it seems clear it uses way more tokens, so I don't expect parity with other competitors on that front.

    On the other hand I expect K3 future refinements to be massive and more efficient.

    • The speed at which tokens are crunched, even on the same hardware, differs between models as well. Using more tokens is only a problem if they are processed at the same speed as with a comparison model.

      3 replies →