4 ms·
It’s very clear from this article (and other product features and rumors) that Anthropic is teeing up for their next model release whose breakthrough feature wi
by aabhay 2mo ago
It’s very clear from this article (and other product features and rumors) that Anthropic is teeing up for their next model release whose breakthrough feature will be the existence of capable agent collaboration.
The irony behind this goal, which is primarily driven by agent simulation environments (gyms) where the goals require agent collaboration, is that this collaboration is still directed towards verifiable reward systems like codebase tasks. So despite being highly qualified to communicate, the model will still be “dumb” in that for unstructured and unverifiable domains the agents won’t be more intelligent or more nuanced.
Agents that might still feel dumb in “general” tasks but are increasingly sophisticated at the narrow domain of math, computer science, and AI research.
- p1esk 2mo agoStrictly speaking, all we need is them improving AI research.
- andai 2mo agoThe RLVR has made them verifiably worse (and less rewarding!) at communication. At least for Claude. GPT had the same problem when 5 came out but they reversed it somehow.
- shevy-java 2mo ago> It’s very clear from this article (and other product features and rumors) that Anthropic is teeing up for their next model release whose breakthrough feature will be the existence of capable agent collaboration. It's a promo article, aka an ad. Unsurprisingly. > Agents that might still feel dumb in “general” tasks but are increasingly sophisticated at the narrow domain of math, computer science, and AI research. I don't see any cleverness there. They just slurp up data and pretend to understand it all.
- cyanydeez 2mo agoSo they're mostly turning agents into blind solidiers. surely this is a good idea.
- deleted 2mo ago[deleted]