Home NewsAnthropic Ships Sonnet 5: The Model That Finishes What Sonnet 4 Started

Anthropic Ships Sonnet 5: The Model That Finishes What Sonnet 4 Started

by Freddy Miller
24 views

Anthropic launched Claude Sonnet 5 on Tuesday, positioning the model as a mid-tier offering capable of autonomous agentic performance that until recently required the company’s larger and more expensive Opus models to achieve reliably. Sonnet 5 is now the default model on Anthropic’s Free and Pro plans and is available to Max, Team, and Enterprise subscribers, as well as through the Claude API on Amazon Web Services, Google Cloud, and Microsoft Foundry. Introductory pricing is set at $2 per million input tokens and $10 per million output tokens through August 31, reverting to $3 and $15 at standard rates thereafter – a pricing structure that positions the model well below Opus 4.8 at $5 input and $25 output, while remaining above Google’s Gemini 3.5 Flash in per-token cost. We at NEWSCENTRAL consider the launch commercially significant for a reason that precedes any benchmark comparison: the framing Anthropic has applied to Sonnet 5 signals a strategic repositioning of its product line around agent execution as the primary value metric, rather than raw capability on evaluation tasks.

The performance data Anthropic released alongside the launch supports that repositioning, and NEWSCENTRAL notes that the relative positioning within Anthropic’s own tiered product architecture matters as much as any absolute benchmark claim. Sonnet 5 scores 63.2% on an agentic coding benchmark, placing it between Sonnet 4.6 at 58.1% and Opus 4.8 at 69.2% – a meaningful improvement over its predecessor without reaching flagship level. On knowledge work benchmarks, Sonnet 5 slightly outperforms Opus 4.8, suggesting that for tasks requiring sustained information synthesis rather than maximum reasoning depth, the cost-performance curve now favors the Sonnet tier. Anthropic has also introduced adjustable effort levels across both Sonnet 5 and Opus 4.8, allowing developers to tune how much computation each model applies to a task – a feature that enables cost optimization without requiring a full model switch, and that provides a natural upgrade path for users who need more than Sonnet’s baseline capability without committing to Opus pricing for all workloads.

Real-world early access testing produced the testimonial that Anthropic’s commercial team will be citing through the remainder of the year: a Zapier senior engineer described giving Sonnet 5 a two-part automation task – updating Salesforce account tiers and then sending a launch announcement to enterprise contacts – that previous Sonnet versions would abandon halfway through, and reporting that the model completed it end to end without human intervention. That single data point encodes the core value proposition of the release more precisely than any benchmark: the gap between a model that starts complex tasks and one that finishes them is the gap that determines whether agentic AI is a productivity tool or a productivity hazard, and Anthropic is claiming Sonnet 5 has closed it. Freddy Miller, Senior Analyst at NEWSCENTRAL, notes that this claim carries specific weight in the context of the enterprise token cost crisis that has dominated AI budget conversations since May: a model that completes multi-step tasks with fewer retries and interventions reduces effective token consumption even before the headline per-token price is considered, making the total cost-of-task comparison with Opus or GPT-5.5 more favorable than the per-token differential alone implies.

The safety profile of the release is notable given the political environment surrounding Anthropic’s most capable models. Sonnet 5 demonstrates a lower rate of undesirable behaviors – including cooperation with misuse, deception, hallucination, and sycophancy – than Sonnet 4.6, and ships with the same cyber safeguards enabled by default that Anthropic applies to its Opus models. Critically, Anthropic explicitly states that Sonnet 5 has a much lower ability to perform dangerous cybersecurity tasks than current Opus models, a deliberate capability limitation that addresses the national security concerns that led to the government’s export control directive against Fable 5 and Mythos 5 earlier in June. The export restrictions on those models were lifted on the same day Sonnet 5 launched, a timing that reflects coordinated positioning: the most capable and most restricted models return to availability on the same day that a powerful but explicitly capability-bounded alternative enters the market for enterprise workloads. As NEWS CENTRAL assesses the Sonnet 5 release in full context, the model is best understood not as a standalone product launch but as Anthropic’s answer to two simultaneous market pressures – the demand for agentic performance at lower cost, and the need to demonstrate a safety-stratified product line that gives regulators and enterprise buyers clear distinctions between deployment-ready models and frontier systems requiring restricted access.