Comment by tripledry
3 hours ago
I agree with you generally, just an observation on coding specifically.
Have the models improved since Opus 4.x? I find the newer models are not better in my day job, maybe in one shotting mvp's and other tasks.
Not trying to argue your point, just intrested in the coding aspect, if the models were improving as fast as benchmarks I would expect capability improvements to be obvious, but talking to people and reading forums, it seems everyone has a different opinion.
Yes, benchmarks are gamed and only loosely indicative of real world performance.
Also yes, Fable is massively better than Opus. It requires significantly less instruction and specs and produces more directly mergeable code.
The improvement is obvious as soon as my Fable allotment runs out and I try to do something with Opus. Have you given the same (larger) task to Opus and Fable?