model releaseAnthropic

Anthropic Launches Claude Opus 5 (Fast) at $10/$50 per Million Tokens, 1M Context Window

TL;DR

Anthropic has released Claude Opus 5 (Fast), a higher-throughput variant of Opus 5 that carries identical capabilities but runs at roughly 2x the price of the standard model. The model ships with a 1 million token context window and is available now through OpenRouter.

2 min read
0

Anthropic Ships Speed-Optimized Opus 5 Variant

Anthropic has released Claude Opus 5 (Fast), a fast-mode variant of its Opus 5 model line, according to a listing on OpenRouter. The model is priced at $10 per 1M input tokens and $50 per 1M output tokens, and supports a 1 million token context window.

According to Anthropic's documentation, the Fast variant delivers "identical capabilities" to the standard Opus 5 model but with higher output speed, at roughly 2x the pricing of the regular version. This suggests a base Opus 5 model exists at a lower price point, though Anthropic has not published separate standard-tier pricing in the material reviewed for this article.

The listing shows a release date of July 24, 2026. OpenRouter is currently hosting the model through a single provider, meaning requests are forwarded directly without cross-provider routing decisions.

What's known about pricing and performance

OpenRouter's platform tracks "effective pricing" — the average cost customers actually pay after prompt caching is applied. For models with heavy repeated context, that discount can reportedly reach 60-80% off list price, though actual savings depend on individual caching patterns and were not broken out specifically for Opus 5 (Fast) in the source material.

No benchmark scores, parameter counts, or training cutoff date were disclosed in the available listing. Throughput, latency, and time-to-first-token (TTFT) metrics are tracked live on OpenRouter's dashboard but were not included as static figures in the source content.

Access

The model is available now via OpenRouter's API, which is OpenAI-compatible — developers can typically switch to it by swapping the base URL and model slug (anthropic/claude-opus-5-fast) in existing integrations. Uptime is monitored continuously by OpenRouter, which automatically retries requests through alternate providers if the primary provider errors out — though currently only one provider serves this model.

What this means

A "Fast" variant priced at 2x the standard rate is a familiar pattern: it lets developers pay a premium for lower latency and higher throughput on the same underlying capability set, without waiting for or requesting a distinct smaller model. This is useful for latency-sensitive applications — real-time agents, interactive coding assistants, customer-facing chat — where token generation speed matters more than cost efficiency.

The bigger open question is what "Opus 5" itself represents. Its existence implies Anthropic has already released or is preparing to release a full Opus 5 line, of which this Fast variant is one configuration. Pricing at $10/$50 per 1M tokens places it in premium territory, above most current frontier model pricing, consistent with Anthropic's historical practice of pricing Opus-tier models at the top of the market. Until Anthropic publishes benchmark data, capability claims about Opus 5 remain unverified, and the practical value of the Fast variant will depend on how much latency improvement it actually delivers in production workloads.

Related Articles

model release

OpenAI Launches GPT-6 Astra With Half the Message Allowance of GPT-5.6 Sol

OpenAI has begun rolling out GPT-6 Astra to top-tier ChatGPT plans, the API, Azure, and AWS Bedrock. The model delivers roughly half the usage allowance of GPT-5.6 Sol across comparable plans, with Plus and Business users gaining access in the coming days.

model release

InclusionAI Releases Ling 3.0 Flash Fin, a Finance-Focused MoE Model with 5.1B Active Parameters

InclusionAI has released Ling 3.0 Flash Fin, a finance-specialized mixture-of-experts model built on Ling 3.0 Flash. The model activates 5.1B of its 124B total parameters and targets long-horizon investment planning tasks while retaining general reasoning, coding, and math capabilities.

research

Anthropic's Claude Fable 5.1 Reportedly Solves 1653 Royalist Cipher in 44 Minutes

According to testing firm Vals AI, Anthropic's Claude Fable 5.1 independently identified and solved the 'Cyphral Distich,' a 1653 numeric cipher by Sir Thomas Urquhart that had defeated other frontier models. The AI decoded a hidden pro-royalist message by mapping each number to a word in Urquhart's original text.

model release

OpenAI Releases GPT-6 Astra, First Model to Cross 'Critical' Cybersecurity Threshold

OpenAI has begun rolling out GPT-6 Astra, the first model to reach the company's internal 'Critical' cybersecurity threshold. Access is being phased, with companies in OpenAI's Daybreak cybersecurity program getting priority following added safeguards after a prior model containment breach.

Comments

Loading...