Comment by Slartie
7 days ago
The speed at which tokens are crunched, even on the same hardware, differs between models as well. Using more tokens is only a problem if they are processed at the same speed as with a comparison model.
7 days ago
The speed at which tokens are crunched, even on the same hardware, differs between models as well. Using more tokens is only a problem if they are processed at the same speed as with a comparison model.
Using more tokens is a significant problem if you pay per token?
Not if the price per token is significantly lower.
Also this arm of the discussion was about speed, not price.
> Using more tokens is a significant problem if you pay per token?
Yoh have posts in this thread suggesting that Fable is 5x more expensive than Kimi K3.