Comment by walrus01
19 hours ago
Pricing at $0.25 and $0.75 already puts its cost well above reasonably reputable inference providers for deepseek v4 flash or qwen 3.8-flash-next or similar class of open weight LLMs that fit in under 170GB of RAM, so I don't see the point. I think this is probably also stupider than laguna s 2.1 which can also be very cheap to serve.
The point is the speed.
If the model provides me with bad results because it's dumb, I don't care how quickly it does it.
But there are lots of use cases where a relatively "dumb" model is good enough.
Is it possible to construct a control system where bad, fast and cheap can become good, fast, and cheap through repeated sampling and a strong spec/eval harness?
I am trying to keep an open mind with AI, but I also have little understanding of control theory, trying to learn.
2 replies →
fast results that you need to verify are better than slow (allegedly better) results that you still need to verify. REPL vs batch.
1 reply →