Anthropic's Claude Opus 4.8 launches on AWS Bedrock in four regions
Anthropic's Claude Opus 4.8 is now available on Amazon Bedrock and Claude Platform on AWS. The model is designed for autonomous multi-stage tasks, agentic coding, and long-running workflows with reduced supervision.
Claude Opus 4.8 — Quick Specs
Claude Opus 4.8 launches on AWS Bedrock in four regions
Anthropic's Claude Opus 4.8 is now available on Amazon Bedrock in four regions: US East (N. Virginia), Asia Pacific (Tokyo), Europe (Ireland), and Europe (Stockholm). The model is also available on Claude Platform on AWS across North America, South America, Europe, and Asia Pacific.
What's new in Opus 4.8
According to Anthropic and AWS, Claude Opus 4.8 is designed for autonomous multi-stage tasks that can run for hours without human intervention. The model can "hold a plan across stages, better track what it has done and what remains, and adjust course when something breaks rather than surfacing an error and stopping," according to the announcement.
Key improvements cited by the companies include:
- Coding: Navigation of production codebases, planning before editing, maintaining context across long sessions
- Agentic workflows: Handling complex dependency chains and multi-step tool use with reduced oversight
- Professional work: Synthesizing long documents into structured deliverables like briefs and reports
- Consistency: Lower output variance and fewer review cycles compared to previous versions
The announcement emphasizes use cases in financial services (investment research, earnings analysis), legal (contract review, due diligence), life sciences (literature review, regulatory submissions), and cybersecurity (threat intelligence, vulnerability assessment).
AWS integration
Claude Opus 4.8 is accessible through multiple AWS interfaces:
- Amazon Bedrock console Playground
- Anthropic Messages API via bedrock-runtime endpoints
- AWS Bedrock Converse API for multi-model workflows
- Anthropic SDK and AWS SDK (Boto3)
The model ID is us.anthropic.claude-opus-4-8 on Bedrock.
AWS integration provides enterprise security features, regional data residency options, and scalable inference within existing AWS environments. Developers using Claude Platform on AWS get Anthropic's native platform experience when regional data residency isn't required.
Pricing and specifications
The announcement does not disclose pricing, context window size, parameter count, or benchmark scores for Claude Opus 4.8. The code examples show a max_tokens parameter of 4096 for output, but input context limits are not specified.
What this means
This launch represents Anthropic's continued expansion of AWS distribution beyond existing Claude 3 models. The emphasis on "autonomous" and "multi-stage" capabilities suggests positioning against OpenAI's o1 series and other reasoning-focused models, though without published benchmarks, direct performance comparisons remain unavailable. The multi-region Bedrock availability addresses enterprise requirements for data sovereignty and low-latency inference in key markets. For AWS customers already using Bedrock, the integration path is straightforward through existing APIs.
Related Articles
Anthropic's Claude Fable 5.1 Reportedly Solves 1653 Royalist Cipher in 44 Minutes
According to testing firm Vals AI, Anthropic's Claude Fable 5.1 independently identified and solved the 'Cyphral Distich,' a 1653 numeric cipher by Sir Thomas Urquhart that had defeated other frontier models. The AI decoded a hidden pro-royalist message by mapping each number to a word in Urquhart's original text.
Anthropic Brings Background Computer Use to Claude Code and Cowork on Mac
Anthropic has enabled background computer use for Claude Code and Claude Cowork on macOS, available to Pro and Max subscribers. The feature lets Claude click, type, and open apps on a Mac without taking over the user's active cursor, following a similar launch by OpenAI's ChatGPT earlier in 2026.
Alibaba Releases Qwen3.8 Max (0902), a 2.4-Trillion-Parameter MoE Model With 1M-Token Context
Alibaba's Qwen team released Qwen3.8 Max (0902), a 2.4-trillion-parameter mixture-of-experts model with a 1M-token context window that accepts text, image, and video input. The snapshot is post-trained for coding, agentic workflows, and long-horizon task execution, priced at $2/$6 per 1M input/output tokens.
OpenAI's GPT-6 Astra Cuts Hallucinations, But Indirect Prompt Injection Attacks Still Succeed 8.5% of the Time
OpenAI's new GPT-6 Astra model shows major improvements in hallucination rates and jailbreak resistance over predecessor GPT-5.6 Sol, according to OpenAI's system card. However, indirect prompt injection attacks hidden in documents still succeed 8.5% of the time in external testing by Gray Swan, down from 27% but still above rival Claude Opus 5's 4.8% rate.
Comments
Loading...