It seems like anthropic is far ahead of openai, and has no reports like this. We have to conclude this is a skill issue/engineering quality problem inside openai.
just because they are a well known name, doesnt mean they havent botched hiring over the last two years or so
Less bad, but https://www.anthropic.com/news/investigating-incidents-cyber...
In some sense though, sure, skill issue explains the gap vs. Anthropic’s much less severe alignment issues.
> We have to conclude
That’s not the most parsimonious explanation even if the assumption it rests on (anthropic ahead of OpenAI) is true, which we don’t have proof of.
Ive anecdotally heard that openai is far more chaotic, which includes not having a central infra team for example (or at least some teams not counting on depending on them). At least the previous hacks in openai were mainly due to bad infra architecture design.
>It seems like anthropic is far ahead of openai
We don't know what internal models look like, and any guesses about it are just speculation.
Anthropic's models seem crippled and hamstrung to begin with
What do you mean? https://www.felonybench.com/
They're almost tied for felonies.