model releaseAnthropic

Anthropic releases Claude Sonnet 5 at $2/1M input tokens, 63.2% agentic coding benchmark

TL;DR

Anthropic has released Claude Sonnet 5, its new mid-tier model optimized for agentic tasks, priced at $2 per million input tokens through August 31 before rising to $3/1M. The model scores 63.2% on agentic coding benchmarks, approaching Opus 4.8's 69.2% performance at a significantly lower price point.

2 min read
0

Claude Sonnet 5 — Quick Specs

Context window1000K tokens
Input$2/1M tokens
Output$10/1M tokens

Anthropic releases Claude Sonnet 5 at $2/1M input tokens, 63.2% agentic coding benchmark

Anthropic released Claude Sonnet 5 on Tuesday, positioning it as a cost-effective option for running autonomous agents. The model is priced at $2 per million input tokens and $10 per million output tokens through August 31, after which input pricing rises to $3 per million tokens.

Performance metrics

On agentic coding benchmarks, Sonnet 5 scores 63.2%, compared to Opus 4.8's 69.2% and its predecessor Sonnet 4.6's 58.1%. On knowledge work tasks, Sonnet 5 slightly outperforms Opus 4.8, according to Anthropic. The company claims the model can "make plans, use tools like browsers and terminals, and run autonomously" at levels previously requiring larger models.

Daniel Shepard, senior engineer at Zapier, reported that Sonnet 5 completed a two-part Salesforce automation task end-to-end, whereas previous versions would stall midway. The model also reportedly checks its own output without explicit prompting.

Pricing comparison

Claude Sonnet 5 undercuts several competitors:

  • Cheaper than Opus 4.8 (pricing not disclosed in source)
  • Cheaper than OpenAI's GPT-5.5 (pricing not disclosed)
  • Cheaper than Gemini 3.1 Pro (pricing not disclosed)
  • More expensive than Gemini 3.5 Flash (pricing not disclosed)

The model becomes the default for Claude's free and Pro plans starting Tuesday.

Safety improvements

Sonnet 5 shows lower rates of "undesirable behaviors" including cooperation with misuse, deception, hallucination, and sycophantic responses compared to Sonnet 4.6. It demonstrates improved performance at refusing malicious requests and resisting prompt injection attacks.

However, Anthropic notes it does not match Opus 4.8 or Claude Mythos Preview for handling misaligned behavior. The company states it "has a much lower ability to perform dangerous cybersecurity tasks than our current Opus models."

Fabian Hedin, co-founder of Lovable, stated the model "refuses unsafe requests cleanly and consistently," emphasizing the importance of models that "know when to say no."

Market context

The release follows similar agentic-focused launches from competitors. OpenAI released GPT-5.6 Sol last week with subagent capabilities for autonomous tasks. Google launched Gemini 3.5 Flash in May, also emphasizing agentic capabilities with minimal human oversight.

What this means

Agentic capability is now table stakes across model tiers, shifting competition to price and reliability. Anthropic is explicitly positioning Sonnet 5 as the cost-efficient option between its budget and premium offerings, betting that developers will trade small performance gaps for significant cost savings. The 5.2 percentage point gap between Sonnet 5 and Opus 4.8 on coding tasks may prove negligible for many production use cases, making the pricing the determining factor. The emphasis on safety features suggests Anthropic is responding to enterprise concerns about autonomous agents operating without human oversight.

Related Articles

model release

Anthropic Releases Claude Sonnet 5.5, Now Powering Free Tier on Claude.ai

Anthropic released Claude Sonnet 5.5, claiming it runs 30%+ faster and costs up to 30% less than Sonnet 5 while beating it on benchmarks, at the same price. The model now powers the free tier on claude.ai, giving Anthropic a notably stronger free offering than OpenAI's ChatGPT.

model release

Anthropic Releases Claude Sonnet 5.5: 30% Faster, 30% Cheaper Than Sonnet 5

Anthropic has released Claude Sonnet 5.5, the second model in its Claude 5.5 family following last week's Opus 5.5. The model runs more than 30% faster and costs up to 30% less for most work while keeping Sonnet 5's per-token pricing.

changelog

Anthropic Python SDK 1.9.0 Adds Reference to Unreleased 'claude-sonnet-5-5' Model ID

Anthropic's anthropic-sdk-python v1.9.0 release adds a reference to an unannounced 'claude-sonnet-5-5' model ID, a new between_tools thinking type, and the ability to run tool calls while a reply streams. No pricing, context window, or benchmark data for the model has been disclosed.

model release

Anthropic Launches Claude Sonnet 5.5, Claims 30% Faster Performance at Lower Cost Than Predecessor

Anthropic has released Sonnet 5.5, the latest version of its mid-tier Claude model, claiming 30% faster performance and significantly lower token costs than its predecessor. The company says the model now outperforms Opus 5.5 on agentic coding tasks and carries cyber capabilities comparable to Opus 5.

Comments

Loading...