What a joke Opus 4.7 at max is.
I gave it an agentic software project to critically review.
It claimed gemini-3.1-pro-preview is wrong model name, the current is 2.5. I said it's a claim not verified.
It offered to create a memory. I said it should have a better procedure, to avoid poisoning the process with unverified claims, since memories will most likely be ignored by it.
It agreed. It said it doesn't have another procedure, and it then discovered three more poisonous items in the critical review.
I said that this is a fabrication defect, it should not have been in production at all as a model.
It agreed, it said it can help but I would need to verify its work. I said it's footing me with the bill and the audit.
We amicably parted ways.
I would have accepted a caveman-style vocabulary but not a lobotomized model.
I'm looking forward to LobotoClaw. Not really.
Based on last few attemts on claude code to address a docker build issue this feels like a downgrade
Claude Code hasn't updated yet it seems, but I was able to test it using `claude --model claude-opus-4-7`
Or `/model claude-opus-4-7` from an existing session
edit: `/model claude-opus-4-7[1m]` to select the 1m context window version
if Opus 4.7 or Mythos are so good how come Claude has some of the worst uptime in most online services?
Am I going to have to make it rewrite all the stuff 4.6 did?
Tried it, after about 10 messages, Opus 4.7 ceased to be able to recall conversation beyond the initial 10 messages. Super weird.
I get a little sad with every new Claude release. Sonnet 4.5 is my favorite and each new model means it's one step closer to being retired. Nothing else replaces it for me
hmmm 20x Max plan on 2.1.111 `Claude Opus is not available with the Claude Pro plan. If you have updated your subscription plan recently, run /logout and /login for the plan to take effect.`
Interesting that despite Anthropic billing it at the same rate as Opus 4.6, GitHub CoPilot bills it at 7.5x rather than 3x.
Qwen 3.6 OSS and now this, almost feels like Anthropic rushed a release to steal hype away from Qwen
Seems like it's not in Claude Code natively yet, but you can do an explicit `/model claude-opus-4-7` and it works.
Claude Code doesn't seem to have updated yet, but I was able to try it out by running `claude --model claude-opus-4-7`
show us the benchmarks with "adaptive thinking" turned on
What’s the default context window? Seems extremely short.
Excited to start using from within Cursor.
Those Mythos Preview numbers look pretty mouthwatering.
Excited to use 1 prompt and have my whole 5-hour window at 100%. They can keep releasing new ones but if they don't solve their whole token shrinkage and gaslighting it is not gonna be interesting to se.
while it seems even with 4.7 we will never see the quality of early 4.6 days, some dude is posting 'agi arrived!!!' on instagram and linkedIn.
Pretty bad. As nerfed 4.6
> "We are releasing Opus 4.7 with safeguards that automatically detect and block requests that indicate prohibited or high-risk cybersecurity uses. "
They're really investing heavily into this image that their newest models will be the death knell of all cybersecurity huh?
The marketing and sensationalism is getting so boring to listen to
Regardless of the model quality improvement, the corporate damage was done by not only ignoring the Opus quality degradation but gaslighting users into thinking they aren’t using it right.
I switched to Codex 5.4 xhigh fast and found it to be as good as the old Claude. So I’ll keep using that as my daily driver and only assess 4.7 on my personal projects when I have time.
Uh oh:
> The new /ultrareview slash command produces a dedicated review session that reads through changes and flags bugs and design issues that a careful reviewer would catch. We’re giving Pro and Max Claude Code users three free ultrareviews to try it out.
More monetization a tier above max subscriptions. I just pointed openclaw at codex after a daily opus bill of $250.As Anthropic keeps pushing the pricing envelope wider it makes room for differentiation, which is good. But I wish oAI would get a capable agentic model out the door that pushes back on pricing.
Ps I know that Anthropic underbought compute and so we are facing at least a year of this differentiated pricing from them, but still..ouch
Well this explains the outages over the last few days
Backlash on HN for Anthropic adjusting usage limits is insane. There's almost no discussion about the model, just people complaining about their subscription.
four prompts with opus 4.6 today is equivalent to 30 or 40 two months ago. infernal downgrade in my case.
Seems they jumped the gun releasing this without a claude code update?
/model claude-opus-4.7
⎿ Model 'claude-opus-4.7' not foundWill it be like the usual: let it work great for 2 weeks, nerf it after?
Opus 4.7 would come out the day before my paid plan ends.
"Agentic Coding/Terminal/Search/Analysis/Etc"...
False: Anthropic products cannot be used with agents.
This is the first new model from Anthropic in a while that I'm not super enthused about. Not because of the model, I literally haven't opened the page about it, I can already guess what it says ("Bigger, better, faster, stronger"), but because of the company.
I have enjoyed using Claude Code quite a bit in the past but that has been waning as of late and the constant reports of nerfed models coupled with Anthropic not being forthcoming about what usage is allowed on subscriptions [0] really leaves a bad taste in my mouth. I'll probably give them another month but I'm going to start looking into alternatives, even PayG alternatives.
[0] Please don't @ me, I've read every comment about how it _is clear_ as a response to other similar comments I've made. Every. Single. One. of those comments is wrong or completely misses the point. To head those off let me be clear:
Anthropic does not at all make clear what types of `claude -p` or AgentSDK usage is allowed to be used with your subscription. That's all I care about. What am I allowed to use on my subscription. The docs are confusing, their public-facing people give contradictory information, and people commenting state, with complete confidence, completely wrong things.
I greatly dislike the Chilling Effect I feel when using something I'm paying quite a bit (for me) of money for. I don't like the constant state of unease and being unsure if something might be crossing the line. There are ideas/side-projects I'm interested in pursuing but don't because I don't want my account banned for crossing a line I didn't know existed. Especially since there appears to be zero recourse if that happens.
I want to be crystal clear: I am not saying the subscription should be a free-for-all, "do whatever you want", I want clear lines drawn. I increasingly feeling like I'm not going to get this and so while historically I've prefered Claude over ChatGPT, I'm considering going to Codex (or more likely, OpenCode) due to fewer restrictions and clearer rules on what's is and is not allowed. I'd also be ok with kind of warning so that it's not all or nothing. I greatly appreciate what Anthropic did (finally) w.r.t. OpenClaw (which I don't use) and the balance they struck there. I just wish they'd take that further.
yay! lobotomized mythos is out
Getting a little suspicious that we might not actually get AGI.
Really disappointed with Anthropic recently, burned through 2 max plans and extra usage past 10 days, getting limited almost 1h in a 5h session. Reading about the extra "safe guards" might be the nail on the coffin.
> during its training we experimented with efforts to differentially reduce these capabilities
> We are releasing Opus 4.7 with safeguards that automatically detect and block requests that indicate prohibited or high-risk cybersecurity uses.
Ah f... you!
is this just mythos flex?
Is that time to turning back from Codex to Claude Code?
its a pretty good coding model - using it in cursor now.
We've all been complaining about Opus 4.6 for weeks and now there's a new model. Did they intentionally gimp 4.6 so they can advertise how much better 4.7 is?
It’s funny, a few months ago I would have been pretty excited about this. But I honestly don’t really care because I can’t trust Anthropic to not play games with this over the next month post release.
I just flat out don’t trust them. They’ve shown more than enough that they change things without telling users.
just started using codex. claude is just marketing machine and benchmaxxing and only if you pay gazillion and show your ID you can use their dangerous model.
This is the 7th advert on the front page right now. It's ridiculous
Might be sticking with 4.6 it's only been 20 minutes of using 4.7 and there are annoyances I didn't face with 4.6 what the heck. Huge downgrade on MRCR too....
256K:
- Opus 4.6: 91.9% - Opus 4.7: 59.2%
1M:
- Opus 4.6: 78.3% - Opus 4.7: 32.2%
Codex release coming today: https://x.com/thsottiaux/status/2044803491332526287
7 trivial prompts, and at 100% limit, using sonnet, not Opus this morning. Basically everyone at our company reporting the same use pattern. Support agent refuses to connect me to a human and terminated the conversation, I can't even get any other support because when I click "get help" (in Claude Desktop) it just takes me back to the agent and that conversation where fin refuses to respond any more.
And then on my personal account I had $150 in credits yesterday. This morning it is at $100, and no, I didn't use my personal account, just $50 gone.
Commenting here because this appears to be the only place that Anthropic responds. Sorry to the bored readers, but this is just terrible service.