model releaseOpenAI

OpenAI's GPT-5.4 mini now available in GitHub Copilot

TL;DR

OpenAI has released GPT-5.4 mini, the lightweight variant of its agentic coding model GPT-5.4, in GitHub Copilot. The model represents OpenAI's highest-performing mini offering to date for code generation and completion tasks.

1 min read
0

GPT-5.4 mini — Quick Specs

Context window400K tokens
Input$0.75/1M tokens
Output$4.5/1M tokens

OpenAI's GPT-5.4 mini Now Generally Available in GitHub Copilot

OpenAI's GPT-5.4 mini has begun rolling out to GitHub Copilot users as a generally available option. According to GitHub, the model is the latest fast-optimized version of GPT-5.4, OpenAI's agentic coding model designed for code generation and completion.

Performance Claims

GitHub states that GPT-5.4 mini represents OpenAI's highest-performing mini model to date in early internal testing. The company has not yet disclosed specific benchmark scores, latency metrics, or token pricing for the model variant.

What's Included

GPT-5.4 mini joins GitHub Copilot's existing model roster, which already includes access to various OpenAI models. The release appears focused on providing developers with a faster, more cost-efficient option for real-time code completion while maintaining competitive performance compared to previous mini variants.

No specific context window size, input/output pricing per million tokens, or parameter count has been disclosed in the announcement.

What This Means

GPT-5.4 mini positions GitHub Copilot users to access a presumably faster inference speed relative to full GPT-5.4, which could reduce latency in IDE integration. However, without published benchmarks or pricing, the concrete advantages over existing models remain unclear. The release follows the industry pattern of offering tiered model variants—full-scale and mini editions—to balance performance with speed and cost across different use cases. Developers using GitHub Copilot should expect this as an additional option within their existing subscription tier.

Related Articles

changelog

OpenAI Cuts GPT-5.6 Prices Up to 80%, Says Model's Own Self-Optimization Work Drove the Savings

OpenAI cut GPT-5.6 Luna pricing by 80% to $0.20/$1.20 per million tokens and GPT-5.6 Terra by 20% to $2/$12, while adding a 2.5x-faster mode for Sol at double the price. The company says GPT-5.6 itself rewrote production inference kernels and tuned its own speculative decoding pipeline to enable the cuts.

changelog

OpenAI Slashes GPT-5.6 Luna Pricing by 80%, Cuts Terra by 20%

OpenAI cut GPT-5.6 Luna pricing by 80 percent to $0.20 per million input tokens and $1.20 per million output tokens, while Terra dropped 20 percent to $2/$12. The company attributes the cuts to infrastructure efficiency gains and mounting price competition, particularly from Chinese providers.

changelog

OpenAI Cuts GPT-5.6 Terra Price 20%, Luna Price 80% Across API and ChatGPT

OpenAI is cutting API prices for its GPT-5.6 Terra and Luna models by 20% and 80%, respectively, compared to prices set earlier this month. The company says the lower costs are also reflected in usage limits for ChatGPT Work and Codex subscribers, though subscription prices remain unchanged.

changelog

OpenAI Cuts GPT-5.6 Luna Price 80%, Terra 20%, as Enterprise Cost Pressure Mounts

OpenAI is cutting the price of GPT-5.6 Luna by 80% and GPT-5.6 Terra by 20%, just three weeks after launching the models. The move comes as enterprises grow more cost-conscious and rivals including Anthropic, Google, and Moonshot AI push cheaper alternatives.

Comments

Loading...