OpenAI Releases GPT-5.6 Luna: $1/$6 Per 1M Tokens With 1M Context Window
OpenAI has released GPT-5.6 Luna, a fast and cost-efficient model in its GPT-5.6 series. The model features a 1 million token context window and is priced at $1 per 1M input tokens and $6 per 1M output tokens, with a knowledge cutoff of February 2026.
GPT-5.6 Luna — Quick Specs
OpenAI Releases GPT-5.6 Luna: $1/$6 Per 1M Tokens With 1M Context Window
OpenAI has released GPT-5.6 Luna, a cost-efficient model in its GPT-5.6 series. The model is priced at $1 per 1M input tokens and $6 per 1M output tokens, with a 1 million token context window and a knowledge cutoff of February 2026.
According to OpenAI, GPT-5.6 Luna is designed for high-volume, latency-sensitive tasks including chat applications, classification, and lightweight agentic workflows. The company claims the model provides "capable reasoning for its price tier."
Pricing and Specifications
GPT-5.6 Luna's pricing positions it as one of OpenAI's more economical offerings:
- Input: $1 per 1M tokens
- Output: $6 per 1M tokens
- Context window: 1M tokens
- Released: July 9, 2026
- Knowledge cutoff: February 2026
The model is currently available through OpenRouter, which forwards requests directly to OpenAI without routing decisions. OpenRouter data indicates that prompt caching can reduce effective costs by 60-80% below list prices for workloads with repeated context.
Technical Details
The model is listed as part of the GPT-5.6 series, though OpenAI has not disclosed specific details about parameter count, architecture changes from previous GPT-5 variants, or benchmark performance scores. The company describes it as "fast" and suited for latency-sensitive applications, though specific throughput and time-to-first-token metrics were not provided at launch.
GPT-5.6 Luna is accessible through OpenAI-compatible APIs, allowing developers to integrate it by changing only the model slug in existing codebases.
What This Means
GPT-5.6 Luna appears positioned as OpenAI's answer to demand for cost-effective models that can handle large context windows without premium pricing. At $1/$6 per 1M tokens with 1M context, it competes directly with mid-tier offerings from Anthropic and other providers targeting high-volume production workloads. The lack of disclosed benchmark scores makes it difficult to assess its capabilities relative to competing models, though its designation as suitable for "lightweight agentic workflows" suggests it may not match the reasoning performance of OpenAI's flagship models. The February 2026 knowledge cutoff indicates relatively recent training data for deployment in mid-2026.
Related Articles
OpenAI Removes Text Chat Limits for ChatGPT Free and Go Users, Upgrades GPT-5.6 Sol for Plus and Pro
OpenAI will remove text chat rate limits for ChatGPT Free and Go users starting next week and add a 'Think' button for deeper reasoning. Plus and Pro subscribers get an updated GPT-5.6 Sol model that OpenAI claims is more accurate with facts, dates, and sourcing.
OpenAI Removes Text Chat Limits for Free ChatGPT Users, Launches GPT-5.6 Luna
OpenAI is removing text chat limits for Free and Go ChatGPT users, powered by a new GPT-5.6 Luna model with a 'Think' button for harder questions. The company also upgraded GPT-5.6 Sol for Plus and Pro users, claiming a 68% reduction in factual errors versus GPT-5.5-Instant.
OpenAI Refines GPT-5.6 Sol for ChatGPT, Unifies Instant/Thinking Modes, Makes Free Text Chat Unlimited
OpenAI is rolling out a ChatGPT-specific tuning of GPT-5.6 Sol that merges Instant and Thinking modes behind a new reasoning slider for Plus and Pro subscribers. Free users now get unlimited text chats with GPT-5.6 Luna and a new Think button.
OpenAI Removes Text Message Rate Limits for Free ChatGPT Accounts
OpenAI is removing rate limits on text-only prompts for Free and Go tier ChatGPT accounts starting next week. Image generation, file uploads, and voice mode will still be capped, and GPT-5.6 Luna becomes the new default model for those tiers.
Comments
Loading...