I think this blog post might be a "reply" of sorts to Anthropic's tweet from some days earlier?
> On ARC-AGI-3, an evaluation where AI models must solve novel problems, Opus 5’s score is three times as high as the next best model.
https://x.com/claudeai/status/2080699504576045299