logoalt Hacker News

mikert89 • today at 5:15 PM • 3 replies • view on HN

have you tried opus 5.5? anthropic is way ahead, atleast in terms of publicly available models


Replies

monideas • today at 5:16 PM

Have you tried Astra? Way ahead?

➕ show 4 replies
Razengan • today at 5:28 PM

Someone else's experience with Opus 5.5: https://news.ycombinator.com/item?id=49821657

> I had it try to prepare a code review for me. Not only did it refuse, it refused to even tell me what the prompt (written by another Claude!) was. Why?

> When I had another model read the session (all of the "stupider" models handled it just fine) it explained that it had the word "reasoning" in it

> That's the entirety of Anthropic's billions of dollars of research: any prompt with the word "reasoning" is trying to hack Claude to figure out how it reasons!

> A model like that should never have gotten out of QA, let alone been released.

➕ show 1 reply
verdverm • today at 5:31 PM

How do you define "way" when saying ahead? How is this measured?

I only use open weight models now and I don't really feel a loss, curious what those who still use it think. I see output from coworkers that does not indicate Claude is that much better (still makes dumb mistakes all the time), not sure they are using the most expensive models either though.

➕ show 1 reply