Anthropic Opus 5 launch: self-verification as the new model capability axis and what it means for agentic system architecture
Anthropic Just Changed the Question The benchmark race has been the dominant story in AI for the past two years. Who scores highest on MATH? Who tops the coding evals? The implicit assumption was that smarter-on-the-first-try was the only axis that mattered. Anthropic just signaled they think that assumption is wrong. When Opus 5 launched…
