← Back to context

Comment by datadrivenangel

4 days ago

The complaint about not being able to switch between linear and log is valid, which is what I did for making a 3D speed/cost/quality frontier application for a recent meetup talk: https://www.williamangel.net/apps/model_performance.html

Because speed is important, as the reasoning and hardware determine both cost and speed. it's a three dimensional tradeoff.

I really like the 3D version, but I strongly believe you need to consider the number of tokens required to complete a task, it heavily impacts the results for certain models that rely heavily on test time compute (Glm-5.3-flash is the newest example).