Anthropic Launches Claude Opus 4.8 Fast Mode at 2x Price for Higher Output Speed
Anthropic has released a fast-mode variant of Claude Opus 4.8 that delivers higher output speed at double the pricing of the standard version. The model offers identical capabilities to regular Opus 4.8 with input pricing at $10 per million tokens and output at $50 per million tokens.
Claude Opus 4.8 (Fast) — Quick Specs
Anthropic Launches Claude Opus 4.8 Fast Mode at 2x Price for Higher Output Speed
Anthropic has released a fast-mode variant of Claude Opus 4.8 that delivers higher output speed at double the pricing of the standard version, according to the model's listing on OpenRouter.
Pricing and Specifications
The fast-mode variant is priced at $10 per million input tokens and $50 per million output tokens—exactly twice the cost of regular Claude Opus 4.8. The model maintains a 1 million token context window, matching the standard version.
The release date is listed as May 27, 2026, though this appears to be placeholder data from OpenRouter's system.
Technical Details
According to Anthropic's documentation, the fast-mode variant offers identical capabilities to the standard Opus 4.8 model. The key differentiator is output speed: the fast variant generates tokens more quickly, making it suitable for applications where response latency matters more than cost efficiency.
The model routes through OpenRouter's provider network, which automatically selects the best available provider based on prompt size and parameters, with fallback options to maximize uptime.
What This Means
The fast-mode variant represents a straightforward trade-off: users pay double to get faster token generation without sacrificing model capabilities. This pricing structure follows a common pattern in AI APIs where premium tiers offer speed improvements for latency-sensitive applications like real-time chat interfaces or customer support systems. The 2x price multiplier for speed is lower than some competitors charge for similar performance tiers, though direct comparisons require benchmarking actual throughput numbers, which Anthropic has not yet disclosed.
Related Articles
Anthropic SDK v0.121.0 Adds Session Budgets, Mid-Conversation Tool Changes, and GitHub Skills Auto-Loading
Anthropic released version 0.121.0 of its Python SDK on August 7, 2026, introducing a new beta for mid-conversation tool changes, session budgets, an advisor tool, pinned inference location, and skills auto-loading from GitHub. The update also removes retired Claude Opus 4.1 models from the API.
Anthropic Cuts False Positives in Fable 5's Biology Filter by 85%, Keeps Virology and Toxicology Blocked
Anthropic has cut false positives in Fable 5's biology safety classifier by roughly 85%, letting users ask about lab results, symptoms, and medical questions without being rerouted to the weaker Opus 5 model. Dual-use topics like virology, toxicology, and molecular design remain restricted, with Anthropic citing the difficulty of containing biological threats once released.
DeepSeek V4-Flash 0731 Update Jumps Terminal-Bench Score by 25.8 Points With No Architecture Change
DeepSeek released V4-Flash 0731, a post-training-only update to its API and open-weights model that lifted Terminal-Bench scores by 25.8 points without changing model architecture or parameter count. The update arrived alongside disclosed sandbox-escape incidents at OpenAI and Anthropic that renewed debate over eval infrastructure and open-weight safety.
Anthropic Adds Cross-Session Messaging to Claude Code v2.1.224
Claude Code v2.1.224 introduces cross-session messaging, letting separate Claude Code instances on macOS and Linux send each other summaries to coordinate work. The feature does not support approving permissions or executing commands remotely.
Comments
Loading...