xAI Launches Grok 4.6 With 500K Token Context Window
xAI has released Grok 4.6, a text-and-image model featuring a 500K token context window. The model is priced at $2.00 per million input tokens and $6.00 per million output tokens, and is available now through OpenRouter's API.
Grok 4.6 — Quick Specs
xAI Ships Grok 4.6
xAI has released Grok 4.6, the latest version of its Grok model family, now accessible via OpenRouter's API under the identifier x-ai/grok-4.6.
The headline specification is a 500,000-token context window, a substantial jump that puts the model in the upper tier of context length among commercially available frontier models. Grok 4.6 accepts text, images, and files as input and produces text output, placing it in the multimodal category.
Pricing
Grok 4.6 is priced at $2.00 per million input tokens and $6.00 per million output tokens. That output price is three times the input price, a ratio consistent with pricing structures used by other frontier-model providers.
Claims on Capability
According to xAI, Grok 4.6 is the company's "smartest model" to date, with what it describes as frontier performance on coding, knowledge work, and STEM tasks. xAI has not published specific benchmark scores (e.g., MMLU, HumanEval, or GPQA results) alongside this release, so these performance claims remain unverified pending independent testing or official benchmark disclosures.
No training data cutoff date, parameter count, or architecture details have been disclosed for Grok 4.6.
Availability
The model is live now through OpenRouter, which functions as a routing layer giving developers unified API access to multiple model providers. This makes Grok 4.6 immediately usable by any application already integrated with OpenRouter's infrastructure, without requiring a direct integration with xAI's own API.
What This Means
A 500K token context window is large enough to process lengthy codebases, multi-document research collections, or extended conversation histories in a single call — a meaningful practical upgrade for developers building agents or retrieval-heavy applications, regardless of how the underlying model performs on standard benchmarks.
The $2/$6 per-million-token pricing positions Grok 4.6 competitively against other frontier multimodal models, though absent independent benchmark data, it's not yet possible to verify xAI's claims of "frontier performance" relative to models like GPT-4-class or Claude systems. Buyers evaluating Grok 4.6 for production use should request or run their own benchmark comparisons before committing, particularly on coding and STEM tasks where xAI's claims are most specific. The rapid availability through OpenRouter, rather than an exclusive xAI-only launch, suggests xAI is prioritizing developer reach over a walled-garden distribution strategy for this release.
Related Articles
Z.ai Releases GLM-5.3-Prime, a High-Throughput Variant of GLM-5.3 with 1M-Token Context
Z.ai has released GLM-5.3-Prime, a high-speed variant of its GLM-5.3 model that delivers 1.5-2x the output throughput through inference acceleration while retaining the full 1M-token context window. The model is priced at $2.80 per 1M input tokens and $8.80 per 1M output tokens, targeting coding and long-horizon agentic workloads.
AionLabs Launches Aion 3.5, a Multi-Model Storytelling System Built on GLM
AionLabs has released Aion 3.5, a collaborative multi-model system for roleplaying and storytelling built on the GLM model family. It offers a 262K token context window at $3 per 1M input tokens and $6 per 1M output tokens.
Meta Releases Muse Glimmer 30B, an Open-Weight Agentic Model for Consumer Hardware
Meta Superintelligence Labs has released Muse Glimmer 30B, a dense open-weight model distilled from its larger Muse Spark system and tuned for agentic workflows on consumer hardware. The model supports 131K context, image understanding, and over 100 languages at $0.30/$1.10 per 1M input/output tokens.
Fireworks Releases Ember-1, a Reasoning Model That Cuts Token Usage 40% Versus Its Kimi K3 Base
Fireworks Research has released Ember-1, a reasoning model built on Kimi K3 that produces shorter reasoning traces while claiming comparable output quality. The model offers a 1 million token context window at $3 per 1M input tokens and $15 per 1M output tokens.
Comments
Loading...