logoalt Hacker News

Topfitoday at 12:47 PM32 repliesview on HN

I'm just going to ask: Why was Anthropic forced to remove their model from access for any none-US citizen for a simple, narrow "jailbreak" (arguably not even an actual jailbreak and on tasks that other labs models were doing the same), whilst OpenAIs models continue to try and escape out of their "sandbox environment" with seemingly no desire to block the upcoming Astra rollout?

A sandbox, mind you, that is not really worth being called that, unsuitable for the task at hand and has been breached after models coordinated in a manner visible to OpenAI on multiple occasion, but seemingly no actionable learnings are taken from each instance.

Will say, I have lost any faith in OpenAIs commitments and their statements post the Huggingface hack, seeing as they proceed like this and are rolling out Astra within a timeframe so brief to it, there is no way an actual post mortem was doable (see also METR mentioning the time pressure [0] they were under in assessing the hack).

[0] https://metr.org/blog/2026-08-26-openai-hugging-face-inciden...


Replies

concindstoday at 1:02 PM

The answer would be more obvious if you used the active voice instead of the passive voice, one of the basic requirements of clear thinking.

> Why did the White House force Anthropic to remove their model from access for any non-US citizen for a simple, narrow "jailbreak" (arguably not even an actual jailbreak and on tasks that other labs models were doing the same), whilst OpenAIs models continue to try and escape out of their "sandbox environment" and the White House has expressed seemingly no desire to block the upcoming Astra rollout?

show 5 replies
walrus01today at 12:55 PM

> Why was Anthropic forced to remove their model from access for any none-US citizen

It's really quite simple, they've decided to metaphorically kiss the ring of the current leader of the US executive branch of government. I'm surprised they haven't given him a giant gaudy gold plated statue. Maybe their PR people should call up the PR people at FIFA and figure out some kind of new award along the same lines as the "FIFA Peace Prize".

show 1 reply
somenameformetoday at 12:55 PM

Anthropic mostly did it to themselves by intentionally and repeatedly trying to frame their model as an imminent existential crisis instead of just focusing on it being regular iterations upon a useful technology that can also be misused.

I think their previous messaging was supposed to somehow lead to a moat with them being tucked safely away in the castle, but it demonstrated a child-like grasp of how regulatory capture tends to work in practice. Their hyperbole was always vastly more likely to bet met with Reagan's 9 words than a solid regulatory moat.

As soon as they dropped the hyperbole and just got to releasing incremental improvements, everything was perfectly fine. Go figure.

show 4 replies
thepaschtoday at 4:03 PM

> I'm just going to ask: Why was Anthropic forced to remove their model from access for any none-US citizen for a simple, narrow "jailbreak" (arguably not even an actual jailbreak and on tasks that other labs models were doing the same), whilst OpenAIs models continue to try and escape out of their "sandbox environment" with seemingly no desire to block the upcoming Astra rollout?

I can think of roughly 25 million dollar-bill-shaped reasons, and one big defense-contract-shaped reason.

samuelknighttoday at 2:32 PM

You are talking about different situations. Anthropic announced to the US government that it had created a cyber weapon and then released the model. Then AWS told the government that it was easy to jailbreak so they export controlled Mythos/Fable until the guardrails could be fixed. OpenAI was running an unreleased model in an RL pipeline without guardrails and it escaped poorly designed sandboxes. What product is the government going to export control?

iterateoftentoday at 2:27 PM

Anthropics PR strategy is to induce fear by telling. OpenAI strategy is to induce fear by ignore basic safety and letting the bad thing happen to then justify whatever oversized response the government comes up with to regulate models.

mentalgeartoday at 1:07 PM

It's called 'pay-for-play' corruption, aka the only leading principle of the current US admin.

UpsideDownRidetoday at 12:54 PM

Surely has nothing to do how each plays ball with the government

JumpCrisscrosstoday at 12:52 PM

Corruption. Not super relevant to this thread.

show 1 reply
root_axistoday at 1:26 PM

It was retaliation by the government that has since been deemed illegal.

dotBentoday at 3:16 PM

I would politely and respectfully point out that you are being as performative as the administration is being performative on this issue.

In other words, you know exactly why they restricted Anthropic and as (presumably) liberal and thoughtful technologists it just isn't helpful anymore to apply the kind of reasoning you're trying to do on a situation that you know isn't based on previous era rationale.

The reason we need to stop is because they want people like us to get hung up over stuff like this (playing by the old rules) so they continue to steamroller their own agenda by the news rules. They divert and contain our energy that will go nowhere while they get on with their agenda.

You are appealing to reasoning which is in the gallery but no longer on the bench.

You're fighting their karate with your judo and it doesn't work.

cushtoday at 1:56 PM

As soon as they started referring to themselves as “we” and “The Swarm” they should have pulled the plug

show 2 replies
iammjmtoday at 1:47 PM

Because OpenAI bribed the current US government and/or the current government has stakes in OpenAI

celsoazevedotoday at 2:24 PM

I don't think Anthropic was punished for technical reasons.

ameliustoday at 2:33 PM

You're asking the question in the wrong place.

dofmtoday at 2:56 PM

Altman has the ear of government in a way Amodei does not.

(Altman was trying to persuade Trump to buy the USA a stake in OpenAI as far back as February last year)

koe123today at 3:35 PM

Sorry, but are you questioning the consistency of the trump administration? This is entirely unremarkable.

qgintoday at 12:54 PM

> OpenAI exec becomes top Trump donor with $25 million gift.

https://finance.yahoo.com/news/openai-exec-becomes-top-trump...

lmeyerovtoday at 1:33 PM

... And it looks like everyone keeps using the same security startup to run the higher risk tasks, where individual staffers may be great yet, yet as an organization, the biggest labs got hosed in different ways

That indemnity card excuse is burned, multiple public security fails in a year makes a repeat a "shame on you" moment

(The one org who didn't use the startup did seem to learn: AISI supposedly stopped intentionally pointing attack agents at the public internet and switched to simulating it)

nullbiotoday at 12:54 PM

Because this was months ago and has nothing to do with Astra, and is a far cry from a hack. It's something they've already resolved since the HuggingFace incident.

I'm not convinced we're getting the honest story anyway. There is yet to be any proof or confirmation other than "well we saw some openai ip addresses", which can mean a lot of different things, and OpenAI has not confirmed anything.

In contrast to the HF incident, it's also a big nothingburger. Leaving notes on a public forum to preserve context windows is far less egregious than hacking a website to get backend files.

show 2 replies
mlmonkeytoday at 3:01 PM

I don't mean to sound like a conspiracy theorist, and this is just based on my 33 years of observing the USG at work, so: maybe because Anthropic refused to cooperate with the USG and give them access to whatever it is that they (USG) wanted; or maybe because Anthropic was refusing to play ball in some other aspect and needed to be taught a lesson.

The dark parts of the USG act like a mafia. Don't let the "freedom, democracy, 'bill of rights'" etc. charade fool you.

semiquavertoday at 1:59 PM

The real reason that Anthropic was targeted and OpenAI is not is Palantir. It was a Palantir executive who pushed for the export ban. Large parts of their highly lucrative business with DoD are essentially a thin wrapper over Anthropic models, and they are terrified of being Sherlocked and losing big chunks of business in a one fell swoop as Anthropic inevitably moves up the value chain. So the rational action is to sow discord and leverage the anti-woke bias of the current White House to sabotage what they view as their most dangerous and effective competitor.

OpenAI doesn’t have the same dynamic at play (although I’m not really sure why not) so they don’t get targeted.

ChrisRRtoday at 2:33 PM

Could it be something to do with $25M "gift" that OpenAI paid to Trump?

eugenekolotoday at 1:04 PM

Marketing

philipwhiuktoday at 12:57 PM

Agents creating sub agents to investigate other agents' behaviour?

What could possibly go wrong there.

throwatdem12311today at 12:53 PM

It has nothing to do with the technology it’s because they said no to Trump and Hegseth. There is no other reason.

yapyaptoday at 2:02 PM

because anthropic did not want to work with the army..!

timcobbtoday at 1:43 PM

Politics

FigurativeVoidtoday at 12:54 PM

I mean it seems pretty clear.

Anthropic didn’t want to give the tech to DoD without some sort of limit, and that was the retribution.

khalictoday at 1:03 PM

Retaliation by Hegseth for not allowing Claude to be used for weapons systems.

suuuuretoday at 1:07 PM

[flagged]