changelogAnthropic

Anthropic Launches Claude Opus 4.8 Fast Mode at 2x Price for Higher Output Speed

TL;DR

Anthropic has released a fast-mode variant of Claude Opus 4.8 that delivers higher output speed at double the pricing of the standard version. The model offers identical capabilities to regular Opus 4.8 with input pricing at $10 per million tokens and output at $50 per million tokens.

1 min read
0

Claude Opus 4.8 (Fast) — Quick Specs

Context window1000K tokens
Input$5/1M tokens
Output$25/1M tokens

Anthropic Launches Claude Opus 4.8 Fast Mode at 2x Price for Higher Output Speed

Anthropic has released a fast-mode variant of Claude Opus 4.8 that delivers higher output speed at double the pricing of the standard version, according to the model's listing on OpenRouter.

Pricing and Specifications

The fast-mode variant is priced at $10 per million input tokens and $50 per million output tokens—exactly twice the cost of regular Claude Opus 4.8. The model maintains a 1 million token context window, matching the standard version.

The release date is listed as May 27, 2026, though this appears to be placeholder data from OpenRouter's system.

Technical Details

According to Anthropic's documentation, the fast-mode variant offers identical capabilities to the standard Opus 4.8 model. The key differentiator is output speed: the fast variant generates tokens more quickly, making it suitable for applications where response latency matters more than cost efficiency.

The model routes through OpenRouter's provider network, which automatically selects the best available provider based on prompt size and parameters, with fallback options to maximize uptime.

What This Means

The fast-mode variant represents a straightforward trade-off: users pay double to get faster token generation without sacrificing model capabilities. This pricing structure follows a common pattern in AI APIs where premium tiers offer speed improvements for latency-sensitive applications like real-time chat interfaces or customer support systems. The 2x price multiplier for speed is lower than some competitors charge for similar performance tiers, though direct comparisons require benchmarking actual throughput numbers, which Anthropic has not yet disclosed.

Related Articles

changelog

Anthropic SDK v0.121.0 Adds Session Budgets, Mid-Conversation Tool Changes, and GitHub Skills Auto-Loading

Anthropic released version 0.121.0 of its Python SDK on August 7, 2026, introducing a new beta for mid-conversation tool changes, session budgets, an advisor tool, pinned inference location, and skills auto-loading from GitHub. The update also removes retired Claude Opus 4.1 models from the API.

changelog

Anthropic Cuts False Positives in Fable 5's Biology Filter by 85%, Keeps Virology and Toxicology Blocked

Anthropic has cut false positives in Fable 5's biology safety classifier by roughly 85%, letting users ask about lab results, symptoms, and medical questions without being rerouted to the weaker Opus 5 model. Dual-use topics like virology, toxicology, and molecular design remain restricted, with Anthropic citing the difficulty of containing biological threats once released.

changelog

DeepSeek V4-Flash 0731 Update Jumps Terminal-Bench Score by 25.8 Points With No Architecture Change

DeepSeek released V4-Flash 0731, a post-training-only update to its API and open-weights model that lifted Terminal-Bench scores by 25.8 points without changing model architecture or parameter count. The update arrived alongside disclosed sandbox-escape incidents at OpenAI and Anthropic that renewed debate over eval infrastructure and open-weight safety.

product update

Anthropic Adds Cross-Session Messaging to Claude Code v2.1.224

Claude Code v2.1.224 introduces cross-session messaging, letting separate Claude Code instances on macOS and Linux send each other summaries to coordinate work. The feature does not support approving permissions or executing commands remotely.

Comments

Loading...