xAI releases Grok 4.3 reasoning model with 1M token context at $1.25/M input tokens
xAI has released Grok 4.3, a reasoning model with a 1 million token context window and no output token limit. The model accepts text and image inputs, has always-on reasoning that cannot be disabled, and uses tiered pricing starting at $1.25 per million input tokens and $2.50 per million output tokens.
Grok 4.3 — Quick Specs
xAI releases Grok 4.3 reasoning model with 1M token context at $1.25/M input tokens
xAI has released Grok 4.3, a multimodal reasoning model with a 1 million token context window and no output token limit. Released on April 30, 2026, the model is now available through OpenRouter.
Specifications and pricing
Grok 4.3 processes text and image inputs with text output. Input tokens are priced at $1.25 per million tokens, while output tokens cost $2.50 per million tokens. According to xAI, requests exceeding 200,000 total tokens are billed at a higher rate, though the elevated pricing tier has not been disclosed.
The model features always-on reasoning that cannot be disabled or configured by effort level. This distinguishes it from other reasoning models that allow users to adjust computational intensity.
Technical capabilities
xAI positions Grok 4.3 for agentic workflows, instruction-following tasks, and applications requiring high factual accuracy. The absence of an output token limit, combined with the 1 million token context window, enables the model to handle long-document analysis and multi-step agentic tasks without truncation.
The model supports multimodal input, accepting both text and images, but outputs text only.
API access
Grok 4.3 is accessible through OpenRouter's API, which normalizes requests and responses across providers. The platform routes requests to available providers and includes fallback mechanisms for uptime.
OpenRouter's API supports accessing the model's reasoning process through a reasoning_details array in responses. The platform requires preserving complete reasoning details when passing messages back to the model for continued conversations.
What this means
Grok 4.3 enters a competitive reasoning model market where always-on reasoning represents a trade-off: consistent step-by-step thinking for all queries, but no ability to reduce computational cost for simpler tasks. The 1M token context window and unlimited output position it for enterprise document analysis and complex multi-turn interactions. The tiered pricing structure for requests over 200K tokens suggests xAI expects the model to be used for extended contexts, though the lack of disclosed upper-tier pricing creates uncertainty for budget planning at scale.
Related Articles
Unverified 'GPT Astra' Model Appears on OpenRouter With 1.05M Token Context, No OpenAI Confirmation
OpenRouter is listing a model called 'OpenAI GPT Astra Latest' with a 1.05 million token context window and $10/$50 per-million-token pricing. OpenAI has made no public announcement, and the listing's own description says it is an auto-redirecting alias rather than a fixed model.
OpenRouter Lists 'GPT Sol Latest' — An Alias Pointer to OpenAI's Newest Sol-Family Model, Not a Standalone Release
OpenRouter has added a listing called '~openai/gpt-sol-latest,' described as an alias that always points to the newest model in an undisclosed 'GPT Sol' family from OpenAI. The listing shows a 1050K token context window and pricing of $2.00 per million input tokens and $10.00 per million output tokens, but OpenAI has not publicly confirmed a model line by this name.
DeepSeek V4.1-Flash Cuts KV Cache Memory by Up to 8x, Matches Opus 5 on Coding Benchmark
DeepSeek released V4.1-Flash, a 552-billion-parameter model built to slash the memory overhead of long-context AI agents. The model cuts GPU cache needs to roughly a quarter of its predecessor's and matches closed models from OpenAI and Anthropic on select coding benchmarks.
DeepSeek Launches V4.1 Flash: Low-Cost MoE Model Claims to Beat V4 Pro
DeepSeek has released V4.1 Flash, a sparse mixture-of-experts model priced at $0.30 per 1M input tokens and $1.20 per 1M output tokens with a 1 million token context window. DeepSeek claims the model exceeds the larger V4 Pro on performance, speed, and task completion time.
Comments
Loading...