model releaseAnthropic

Anthropic releases Claude Sonnet 5 with improved agentic capabilities, $2/$10 per million tokens through August

TL;DR

Anthropic has released Claude Sonnet 5, replacing Sonnet 4.6 as its medium-sized model. The company claims improved agentic performance approaching Opus 4.8 levels while maintaining lower pricing at $2 per million input tokens and $10 per million output tokens through August 31.

2 min read
0

Claude Sonnet 5 — Quick Specs

Context window1000K tokens
Input$2/1M tokens
Output$10/1M tokens

Anthropic Releases Claude Sonnet 5 with Enhanced Agentic Capabilities

Anthropic launched Claude Sonnet 5 on June 30, 2026, replacing the four-month-old Sonnet 4.6 as its medium-sized model. The company is offering introductory pricing of $2 per million input tokens and $10 per million output tokens through August 31, after which pricing increases to $3 and $15 respectively.

Sonnet 5 is now the default model for both free and Pro Claude users across the Chat, Cowork, Claude Code, and Platform products.

Performance Claims

According to Anthropic, Sonnet 5 delivers performance "close to that of Opus 4.8" while maintaining lower pricing. The company claims "substantial improvement" over Sonnet 4.6 in reasoning, tool use, coding, and knowledge work.

The model can "make plans, use tools like browsers and terminals, and run autonomously at a level that, just a few months ago, required larger and more expensive models," Anthropic states. Specific benchmark scores have not been disclosed.

Model Lineup Context

In Anthropic's three-tier system, Sonnet sits between the smaller Haiku model and the larger Opus model. The company also operates two other model lines: Mythos 5 (with reduced guardrails for security researchers and platform owners) and Fable 5 (the safer version currently blocked by U.S. government restrictions).

Claude Opus 4.8 was released in late May 2026. Sonnet 4.6, the previous medium-sized model, launched in February 2026.

Rate Limits and Availability

Anthropic increased rate limits across its products to accommodate higher token usage at elevated effort levels. Users can select effort levels based on project requirements. The model is available immediately through all Claude interfaces.

What This Means

Sonnet 5 represents Anthropic's attempt to compress flagship-level performance into a mid-tier price point. The four-month release cycle suggests aggressive iteration on agentic capabilities. The introductory pricing—50% below standard rates—indicates either competitive pressure or an effort to build market share before the price increase. Without published benchmarks, however, the claimed performance parity with Opus 4.8 remains unverified. The August 31 pricing cliff could create migration challenges for cost-sensitive applications.

Related Articles

research

Anthropic's Claude Fable 5.1 Reportedly Solves 1653 Royalist Cipher in 44 Minutes

According to testing firm Vals AI, Anthropic's Claude Fable 5.1 independently identified and solved the 'Cyphral Distich,' a 1653 numeric cipher by Sir Thomas Urquhart that had defeated other frontier models. The AI decoded a hidden pro-royalist message by mapping each number to a word in Urquhart's original text.

product update

Anthropic Brings Background Computer Use to Claude Code and Cowork on Mac

Anthropic has enabled background computer use for Claude Code and Claude Cowork on macOS, available to Pro and Max subscribers. The feature lets Claude click, type, and open apps on a Mac without taking over the user's active cursor, following a similar launch by OpenAI's ChatGPT earlier in 2026.

model release

OpenAI's GPT-6 Astra Reportedly Automates AI Engineering Tasks at Under $6 an Hour, According to Latent Space Testing

A Latent Space report describes GPT-6 Astra, a new OpenAI model the blog says can autonomously handle AI engineering tasks—training models, labeling data, deploying systems—at an estimated cost of under $6 per hour. The claims, including 97.6% on FrontierMath and 99.9% on ARC-AGI-3, come from independent blog testing rather than an official OpenAI announcement.

model release

Alibaba Releases Qwen3.8 Max (0902), a 2.4-Trillion-Parameter MoE Model With 1M-Token Context

Alibaba's Qwen team released Qwen3.8 Max (0902), a 2.4-trillion-parameter mixture-of-experts model with a 1M-token context window that accepts text, image, and video input. The snapshot is post-trained for coding, agentic workflows, and long-horizon task execution, priced at $2/$6 per 1M input/output tokens.

Comments

Loading...