← Back to context

Comment by Waterluvian

2 hours ago

Every time weird stuff happened this week it was because the model choice in VSCode got set back to “auto” and some other model was trying its best. 5.5 is what I just set it to. Even Fable feels worse for my use cases.

Even Anthropic rates Fable lower than 5.5 on pretty much all benchmarks.

"Why does Fable even exist" is a very very reasonable question right now.

  • You are making the mistake of classifying models on a single linear axis, or even a multi axis basis set of all benchmarks. That just isn’t true. Each model is unique in its skills and capabilities and the way it approaches problems, in a way that is not represented in benchmarks. Fable is better at reviewing things. I don’t know how to explain it well but it is true. I trust Fable to do thorough reviews (sometimes too thorough) and to present its information in a dense but ordered way. Its output is equivalent to what you used to get from security firms doing code reviews. Having Opus do the work, and have fable do reviews (of the plan and implementation) is a good combo.

  • It's a class of model not a static one. There'll be Fable 5.5 that's even better than Opus 5.5.

    • Although we never had a Claude Opus 5.1. Fable went 5 → 5.1 and Opus went 5 → 5.5.

      I assume Fable 5.1 is the first re-tuned (is there a better term for this?) version of Fable 5 and Opus 5.5 is the first re-tuned version of Opus 5, and maybe the Opus one just came out better for some reason?

  • Because Anthropic releases their different model level's at very different points in time, they seem to always have one model that is by far an away the best to use for everything. Haiku 4.5 is almost a year old. Sonnet is fine, but idk if it has any real benefits over Opus. It's only been since Fable has been released that you get to choose between Fable and Opus, but not with 5.5 there is no reason to use Fable.

    I feel like instead of releasing fable, they should have released it as Opus 5, then their next Opus release they would call Sonnet, and their next Sonnet release they would have called Haiku. I don't know if their pricing structure would have been able to support that, but Anthropic has always been the least competitive regarding token pricing.