Claude Opus 4.7 leads on SWE-bench and agentic reasoning, beating GPT-5.4 and Gemini 3.1 Pro

This article was published on April 16, 2026 In short: Anthropic has released Claude Opus 4.7, its most capable generally available model, with benchmark-leading scores on SWE-bench Pro (64.3% vs GPT-5.4’s 57.7%), multi-agent coordination for hours-long workflows, 3x higher image resolution, and a 14% improvement in multi-step agentic reasoning with a third of the tool errors. Priced at $5/$25 per million tokens, it is available across Claude plans and through Amazon Bedrock, Vertex AI, and Microsoft Foundry. Anthropic has released Claude Opus 4.7, its most capable generally available model to date, with benchmark-leading performance in software engineering and agentic reasoning that widens the gap between Claude and both OpenAI’s GPT-5.4 and Google’s Gemini 3.1 Pro on the tasks that matter most to developers and enterprise users. The release comes at a moment when Anthropic’s commercial momentum is difficult to overstate. The company is running at a $30 billion annualised revenue rate, has attracted investor offers at roughly $800 billion, and is in early IPO talks. Opus 4.7 is the model that has to justify those numbers, not by winning every benchmark, but by being the model that
Source: For the complete article, please visit the original source link below.