{"id":91194,"date":"2026-06-30T21:02:24","date_gmt":"2026-06-30T21:02:24","guid":{"rendered":"https:\/\/www.europesays.com\/ai\/91194\/"},"modified":"2026-06-30T21:02:24","modified_gmt":"2026-06-30T21:02:24","slug":"anthropics-new-claude-sonnet-5-closes-the-gap-to-the-pricier-opus-model-series","status":"publish","type":"post","link":"https:\/\/www.europesays.com\/ai\/91194\/","title":{"rendered":"Anthropic&#8217;s new Claude Sonnet 5 closes the gap to the pricier Opus model series"},"content":{"rendered":"<p>Anthropic released Claude Sonnet 5. In benchmarks, it closes in on the larger Opus 4.8 and even beats it in some areas. The model is available now at an introductory price.<\/p>\n<p>Anthropic calls it the most agentic Sonnet yet: it can build plans, grab tools like browsers and terminals, and work on its own at a level that just months ago only bigger, pricier models could pull off, according to the company. Sonnet 5 is meant to close that gap.<\/p>\n<p>Benchmarks show a clear jump over Sonnet 4.6<\/p>\n<p>Anthropic&#8217;s published benchmarks show Sonnet 5 beating its predecessor Sonnet 4.6 in every tested category while gaining ground on the pricier Opus 4.8. On agentic coding, Sonnet 5 hits 63.2 percent on SWE-bench Pro, up from 58.1 percent for Sonnet 4.6. Opus 4.8 sits at 69.2 percent. On Terminal-Bench 2.1, Sonnet 5 pulls 80.4 percent versus Sonnet 4.6&#8217;s 67.0 percent. For multidisciplinary reasoning (Humanity&#8217;s Last Exam), the model reaches 57.4 percent with tools, nearly matching Opus 4.8 at 57.9 percent. On computer use (OSWorld-Verified), Sonnet 5 posts 81.2 percent compared to 78.5 percent for its predecessor.<\/p>\n<p><a href=\"https:\/\/www.europesays.com\/ai\/wp-content\/uploads\/2026\/06\/sonnet_5_benchmarks-scaled-1.webp\"><img fetchpriority=\"high\" decoding=\"async\" class=\"wp-image-57999 size-full\" src=\"https:\/\/www.europesays.com\/ai\/wp-content\/uploads\/2026\/06\/sonnet_5_benchmarks-scaled-1.webp\" alt=\"\" width=\"2560\" height=\"1215\"\/><\/a>Sonnet 5 beats its predecessor, Sonnet 4.6, across every tested category and closes in on the pricier Opus 4.8. On knowledge work (GDPval-AA v2), Sonnet 5 even edges past Opus 4.8 with 1,618 points versus 1,615. | Image: Anthropic<\/p>\n<p>On the knowledge work benchmark GDPval-AA v2,\u00a0<a href=\"https:\/\/the-decoder.com\/openai-says-top-ai-models-are-reaching-expert-territory-on-real-world-knowledge-work\/\" target=\"_blank\" rel=\"noopener noreferrer nofollow\">which tests AI on real-world knowledge tasks<\/a>, Sonnet 5 actually beats the larger Opus 4.8, scoring 1,618 to Opus&#8217;s 1,615. Anthropic says feedback from early-access partners told the same story. Sonnet 5 acts far more agentically than previous versions, showing up in things like how it handles search tasks.<\/p>\n<p><a href=\"https:\/\/www.europesays.com\/ai\/wp-content\/uploads\/2026\/06\/sonnet_5_agentic_search-scaled-1.webp\"><img loading=\"lazy\" decoding=\"async\" class=\"wp-image-58000 size-full\" src=\"https:\/\/www.europesays.com\/ai\/wp-content\/uploads\/2026\/06\/sonnet_5_agentic_search-scaled-1.webp\" alt=\"\" width=\"2560\" height=\"1440\"\/><\/a>Agentic search performance on BrowseComp by effort level and cost per task. Sonnet 5 (orange) clearly outperforms Sonnet 4.6 (gray) at every level while offering cheaper entry points. Opus 4.8 (yellow) stays ahead at the highest effort settings. | Image: Anthropic<br \/>\nCybersecurity isn&#8217;t a concern this time<\/p>\n<p>Lately, Anthropic has been making news for models it\u00a0can&#8217;t\u00a0ship. The\u00a0<a href=\"https:\/\/the-decoder.com\/anthropics-fable-5-could-return-within-days-as-trump-administration-prepares-to-lift-restrictions\/\" target=\"_blank\" rel=\"noopener noreferrer nofollow\">US government is blocking the company&#8217;s two most capable models, Mythos 5 and Fable 5<\/a>, over cybersecurity concerns. That context hangs over the Sonnet 5 launch. Anthropic is clearly eager to get ahead of any similar worries. The model wasn&#8217;t trained on cybersecurity tasks, the company says, and in tests for risky capabilities like writing software exploits, it scores far below both Opus 4.8 and Mythos 5.<\/p>\n<p><img loading=\"lazy\" decoding=\"async\" class=\"wp-image-37219 size-full\" src=\"https:\/\/www.europesays.com\/ai\/wp-content\/uploads\/2026\/06\/sonnet5_firefox_exploits-3840x2160-1-scaled-1.webp\" alt=\"\" width=\"2560\" height=\"1440\"\/>Firefox 147 exploit evaluation. Like its predecessor Sonnet 4.6, Sonnet 5 couldn&#8217;t develop a fully working exploit but shows a slightly higher partial control rate at 13.2 percent. Mythos 5 and Opus 4.8 are far more capable at this task. | Image: Anthropic<\/p>\n<p>Sonnet 5 does score a bit higher than its predecessor on these tasks, though. So Anthropic has switched on\u00a0<a href=\"https:\/\/support.claude.com\/en\/articles\/14604842-real-time-cyber-safeguards-on-claude\" target=\"_blank\" rel=\"noopener noreferrer nofollow\">cyber safeguards<\/a>\u00a0by default. They flag and block risky cyber usage in real time, on par with the protections already in place for Claude Opus 4.7 and 4.8. They&#8217;re dialed back compared to Fable 5&#8217;s guardrails, which users\u00a0<a href=\"https:\/\/the-decoder.com\/claude-fable-5-anthropic-admits-wrong-tradeoff-after-invisibly-throttling-rival-ai-researchers\/\" target=\"_blank\" rel=\"noopener noreferrer nofollow\">complained about almost immediately<\/a>. Anthropic says it views the overall cybersecurity risk from Sonnet 5 as low.<\/p>\n<p>On the safety front, the model does a better job turning down malicious requests and fending off prompt injection attacks than Sonnet 4.6, according to Anthropic. Hallucinations and\u00a0<a href=\"https:\/\/the-decoder.com\/sycophantic-ai-chatbots-can-break-even-ideal-rational-thinkers-researchers-formally-prove\/\" target=\"_blank\" rel=\"noopener noreferrer nofollow\">sycophantic behavior<\/a>, the tendency to just agree with whatever the user says, are down as well. Anthropic&#8217;s full safety evaluation is in the\u00a0<a href=\"https:\/\/www.anthropic.com\/claude-sonnet-5-system-card\" target=\"_blank\" rel=\"noopener noreferrer nofollow\">Claude Sonnet 5 System Card<\/a>.<\/p>\n<p>Introductory pricing runs through August 2026<\/p>\n<p>Claude Sonnet 5 is live now on all plans. It&#8217;s the new default for Free and Pro users, and Max, Team, and Enterprise subscribers can access it too. Developers can plug it into Claude Code and the Claude Platform. On the API side, it goes by\u00a0<a href=\"https:\/\/platform.claude.com\/docs\/en\/about-claude\/models\/overview\" target=\"_blank\" rel=\"noopener noreferrer nofollow\">&#8220;claude-sonnet-5&#8221;<\/a>. The training cutoff is January 2026, with a one-million-token context window.<\/p>\n<p>Until August 31, 2026, Anthropic is charging $2 per million input tokens and $10 per million output tokens.\u00a0<a href=\"https:\/\/platform.claude.com\/docs\/en\/about-claude\/pricing\" target=\"_blank\" rel=\"noopener noreferrer nofollow\">After that<\/a>, prices jump to $3 and $15, which is what previous Sonnet models cost.<\/p>\n<p>Real-world costs might tell a different story:\u00a0Because the model works more agentically, <a href=\"https:\/\/the-decoder.com\/frontier-radar-3-how-agentic-ai-is-turning-tokens-into-a-business-metric\/\" rel=\"nofollow noopener\" target=\"_blank\">it&#8217;s likely to chew through more tokens per task<\/a>. So even at the same per-token rate, running Sonnet 5 could end up costing more than its predecessors. The same thing happened when Opus went from 4.6 to 4.7.<\/p>\n<p>\t\t\t\tAI News Without the Hype \u2013 Curated by Humans<\/p>\n<p>\n\t\t\t\t\tSubscribe to THE DECODER for ad-free reading, a weekly AI newsletter, our exclusive &#8220;AI Radar&#8221; frontier report six times a year, full archive access, and access to our comment section.\t\t\t\t<\/p>\n<p>\t\t\t\t<a href=\"https:\/\/the-decoder.com\/subscription\/\" class=\"inline-block text-white bg-(--heise-primary) mt-3 hover:bg-blue-800 focus:ring-4 focus:outline-none focus:ring-blue-300 font-medium rounded-sm w-full sm:w-auto  pl-3 pr-3 py-2.5 text-center newsletter-submit-button hover:no-underline\" rel=\"nofollow noopener\" target=\"_blank\"><br \/>\n\t\t\t\t\tSubscribe now\t\t\t\t<\/a><\/p>\n","protected":false},"excerpt":{"rendered":"Anthropic released Claude Sonnet 5. In benchmarks, it closes in on the larger Opus 4.8 and even beats&hellip;\n","protected":false},"author":2,"featured_media":91195,"comment_status":"","ping_status":"","sticky":false,"template":"","format":"standard","meta":{"footnotes":""},"categories":[8],"tags":[53,182,48038],"class_list":["post-91194","post","type-post","status-publish","format-standard","has-post-thumbnail","category-anthropic","tag-anthropic","tag-claude","tag-sonnet-5"],"_links":{"self":[{"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/posts\/91194","targetHints":{"allow":["GET"]}}],"collection":[{"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/posts"}],"about":[{"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/types\/post"}],"author":[{"embeddable":true,"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/users\/2"}],"replies":[{"embeddable":true,"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/comments?post=91194"}],"version-history":[{"count":0,"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/posts\/91194\/revisions"}],"wp:featuredmedia":[{"embeddable":true,"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/media\/91195"}],"wp:attachment":[{"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/media?parent=91194"}],"wp:term":[{"taxonomy":"category","embeddable":true,"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/categories?post=91194"},{"taxonomy":"post_tag","embeddable":true,"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/tags?post=91194"}],"curies":[{"name":"wp","href":"https:\/\/api.w.org\/{rel}","templated":true}]}}