API pricing

9 articles tagged with API pricing

August 14, 2026
product updateOpenAI

OpenAI Launches 'Ultrafast' Mode for GPT-5.6 Sol, Hitting 750 Tokens/Second via Cerebras

OpenAI has launched a preview of 'Ultrafast' mode for GPT-5.6 Sol, delivering up to 750 output tokens per second through Cerebras inference hardware. The feature is initially limited to select API customers as part of a tiered speed pricing structure.

changelogDeepSeek

DeepSeek to Quadruple API Prices for V4 Pro and V4 Flash Starting August 16

DeepSeek will raise API output token pricing roughly fourfold starting August 16, introducing peak and off-peak rates for its V4 Pro and V4 Flash models. Despite the increase, DeepSeek remains cheaper than competitors like OpenAI's GPT-5.6 Sol and Moonshot's Kimi K3.

August 13, 2026
changelogDeepSeek

DeepSeek Ships V4-Pro-0813, Open-Sources Agent Harness, Raises API Prices Up to 52%

DeepSeek released build V4-Pro-0813 with major agent benchmark gains, open-sourced its Deepseek Harness agent framework under MIT license, and announced API price increases of up to 52% effective August 16.

August 12, 2026
model release

xAI Launches Grok 4.6, Claims Parity with GPT-5.6 Sol and Near-Parity with Claude Fable 5

xAI released Grok 4.6, claiming intelligence on par with OpenAI's GPT-5.6 Sol and just one point behind Anthropic's Claude Fable 5 Max on the Artificial Analysis Intelligence Index. The model is priced at $2 per million input tokens and $6 per million output tokens, with a faster variant at double that rate.

product update

Mistral Adds EU/US Regional Routing and Paid Priority Queue, Both With Coverage Gaps

Mistral has made regional inference generally available, letting customers route requests through EU or US servers for a 10 percent surcharge, while also launching a paid Priority Tier that charges 1.75x standard pricing for faster processing during peak traffic. Both offerings carry significant limitations on what data and features they actually cover.

July 30, 2026
changelogOpenAI+1

OpenAI Slashes GPT-5.6 Luna Pricing by 80%, Cuts Terra by 20%

OpenAI cut GPT-5.6 Luna pricing by 80 percent to $0.20 per million input tokens and $1.20 per million output tokens, while Terra dropped 20 percent to $2/$12. The company attributes the cuts to infrastructure efficiency gains and mounting price competition, particularly from Chinese providers.

July 24, 2026
model releaseAnthropic

Anthropic Ships Claude Opus 5, Claims It Matches Flagship Fable 5 on Coding at Half the Cost

Anthropic released Claude Opus 5 on July 24, its fourth model launch in under two months, priced at $5 per million input tokens and $25 per million output tokens. The company claims the model matches or beats its flagship Fable 5 on most coding and knowledge-work benchmarks while posting the lowest deception rate of any model it has shipped.

model releaseAnthropic

Anthropic Releases Claude Opus 5, Claims Near-Fable 5 Intelligence at Half the Price

Anthropic has released Claude Opus 5, upgrading from Opus 4.8, with pricing held at $5 per million input tokens and $25 per million output tokens. The company claims the model approaches the intelligence of its flagship Fable 5 model at half the cost.

April 29, 2026
model releaseOpenAI

OpenAI releases GPT-5.5 with 82.7% Terminal-Bench score, API priced at $5/$30 per million tokens

OpenAI released GPT-5.5 on April 23, its first retrained base model since GPT-4.5, scoring 82.7% on Terminal-Bench 2.0 versus GPT-5.4's 75.1% and Claude Opus 4.7's 69.4%. API pricing is set at $5 per million input tokens and $30 per million output tokens, exactly double GPT-5.4 rates.