logoalt Hacker News

aabhaytoday at 5:30 AM2 repliesview on HN

It’s very clear from this article (and other product features and rumors) that Anthropic is teeing up for their next model release whose breakthrough feature will be the existence of capable agent collaboration.

The irony behind this goal, which is primarily driven by agent simulation environments (gyms) where the goals require agent collaboration, is that this collaboration is still directed towards verifiable reward systems like codebase tasks. So despite being highly qualified to communicate, the model will still be “dumb” in that for unstructured and unverifiable domains the agents won’t be more intelligent or more nuanced.

Agents that might still feel dumb in “general” tasks but are increasingly sophisticated at the narrow domain of math, computer science, and AI research.


Replies

andaitoday at 7:59 AM

The RLVR has made them verifiably worse (and less rewarding!) at communication.

At least for Claude. GPT had the same problem when 5 came out but they reversed it somehow.

p1esktoday at 7:35 AM

Strictly speaking, all we need is them improving AI research.