← Back to context

Comment by Slartie

7 days ago

The speed at which tokens are crunched, even on the same hardware, differs between models as well. Using more tokens is only a problem if they are processed at the same speed as with a comparison model.

Using more tokens is a significant problem if you pay per token?

  • Not if the price per token is significantly lower.

    Also this arm of the discussion was about speed, not price.

  • > Using more tokens is a significant problem if you pay per token?

    Yoh have posts in this thread suggesting that Fable is 5x more expensive than Kimi K3.