← Back to context

Comment by tyre

3 hours ago

> There is a reason why we hear less about this idea of smaller expert models, because large strong models to the tasks just as good.

Smaller models are cheaper, sometimes faster. I agree that the “we’re an LLM fine-tuned for X” hasn’t worked out because you can just train Claude to do X (and Anthropic will), but not burning Opus/Fable tokens on dumb-but-token-heavy tasks is good sense.

As we move from “integrate AI into Y” to “optimize the ROI on Y”, we’ll see more of this.