Opus 5 is terrible. I'd even say it's a step backwards from 4.8. I'm getting high error rates from it, and then it catches the error, and then it sometimes errors the error fix (!).
Just today I had to switch another agent to Fable with the instruction, "Please clean up the mess that Opus 5 made, thanks"
The other day, Sol called Opus 5's handoff (a skill I have that is basically a compaction, but just written to a file not tied to one LLM) "incoherent", that was a new one.
Opus 4.8 or Fable (at great expense) are the only ones that aren't frustrating for me.
Thanks, that’s interesting to know. I don’t know much about LLMs so I use 5 because it’s a bigger number than 4.8.
Every time when Opus 5 needs a design decision and presents me with suggestions/recommendations, I switch to Fable and ask it to think again, and it almost always replies something like "Actually my previous suggestions were wrong" and describes in detail a bunch of ways in which Opus 5's suggestions were indeed complete garbage.
Same here. Regularly reverting back to Opus 4.8 after 5.0 being terrible.
Anthropic does this all the time (ruins their models for users) while they screw around with system prompts. Oh but it's for your own good of course! They know what's best for us all, if we would just give them a monopoly.
I can't wait until OpenAI/Grok/Chinese models surpass them enough that their main character syndrome and smug doomerism no longer draws much media attention.
Interesting. My experience has been similar. Opus 4.8 was awesome. Opus 5 feels a little off, although I can't put my finger on exactly what it is.