Comment by tyre
3 hours ago
> There is a reason why we hear less about this idea of smaller expert models, because large strong models to the tasks just as good.
Smaller models are cheaper, sometimes faster. I agree that the “we’re an LLM fine-tuned for X” hasn’t worked out because you can just train Claude to do X (and Anthropic will), but not burning Opus/Fable tokens on dumb-but-token-heavy tasks is good sense.
As we move from “integrate AI into Y” to “optimize the ROI on Y”, we’ll see more of this.
No comments yet
Contribute on Hacker News ↗