Comment by epolanski
6 days ago
Even if hardware capacity increases, it seems clear it uses way more tokens, so I don't expect parity with other competitors on that front.
On the other hand I expect K3 future refinements to be massive and more efficient.
The speed at which tokens are crunched, even on the same hardware, differs between models as well. Using more tokens is only a problem if they are processed at the same speed as with a comparison model.
Using more tokens is a significant problem if you pay per token?
Not if the price per token is significantly lower.
Also this arm of the discussion was about speed, not price.
> Using more tokens is a significant problem if you pay per token?
Yoh have posts in this thread suggesting that Fable is 5x more expensive than Kimi K3.