← Back to context

Comment by _aavaa_

3 days ago

Artificial analysis, the place where they get this data from, has a cost vs time chart.

Go to https://artificialanalysis.ai/ and scroll down to the second graph under “Speed & Latency”.

I think this is the most import graph on their page. I wish they would let us filter by intelligence, or pass rate, and then see this graph. This is the tradeoff that actually matters, cost/token or tok/s can be very misleading (take glm-5.3-flash as an example).