3 ms·
I agree with you generally, just an observation on coding specifically. Have the models improved since Opus 4.x? I find the newer models are not better in my d
by tripledry 13d ago
I agree with you generally, just an observation on coding specifically.
Have the models improved since Opus 4.x? I find the newer models are not better in my day job, maybe in one shotting mvp's and other tasks.
Not trying to argue your point, just intrested in the coding aspect, if the models were improving as fast as benchmarks I would expect capability improvements to be obvious, but talking to people and reading forums, it seems everyone has a different opinion.
- blfr 13d agoYes, benchmarks are gamed and only loosely indicative of real world performance. Also yes, Fable is massively better than Opus. It requires significantly less instruction and specs and produces more directly mergeable code. The improvement is obvious as soon as my Fable allotment runs out and I try to do something with Opus. Have you given the same (larger) task to Opus and Fable?
- tripledry 13d agoI have not experimented that much, my observation is mostly that people seem to have different experiences with the capabilities. Personally I still use mainly Opus and find it handles most tasks quite well (without burning all my corporate quota).