model release

Anonymous 'Ox Alpha' Reasoning Model Appears on OpenRouter with Free 1M-Token Context

TL;DR

A stealth model called Ox Alpha has appeared on OpenRouter, offering a 1 million token context window at no cost during its preview period. The model's developer remains anonymous, and OpenRouter says it is acting only as a router, not the model's owner or provider.

2 min read
0

Stealth Model Ox Alpha Launches Free on OpenRouter

An unidentified AI developer has released a model called Ox Alpha through OpenRouter, positioning it as a reasoning system built for coding, sustained agentic work, and production workloads. The model is currently free to use and offers a 1 million token context window, according to its OpenRouter listing.

Ox Alpha appeared on the platform on August 20, 2026, though details about its architecture, parameter count, and training data remain undisclosed. OpenRouter explicitly states it is not the model's developer, owner, or provider — it is only routing requests to a third-party operator who has chosen to remain anonymous during this preview phase.

What's Known

According to OpenRouter's listing, Ox Alpha is designed for:

  • Long-horizon software engineering tasks
  • Complex reasoning workflows
  • Multimodal workflows that combine text with visual context

The model supports a 1,000,000 token context window and is currently offered at no cost — no input or output pricing has been set, consistent with a preview or evaluation release. Latency, throughput, and uptime metrics are not yet available due to limited usage data, though OpenRouter reports 100% availability over the last 24 hours across all locations, with three days of uptime history so far.

No benchmark scores, parameter counts, or training cutoff dates have been disclosed. The provider has not confirmed its identity, and no company in the AI industry has yet claimed Ox Alpha as its own.

Data Handling

OpenRouter notes that prompts and completions sent to Ox Alpha are retained by the anonymous provider but are not used for training. All other usage is governed by OpenRouter's Stealth Model Terms, a framework the platform applies to unidentified models during preview windows before a public launch.

What This Means

Stealth releases on OpenRouter have become a common way for AI labs — established or new — to gather real-world usage data and community feedback before a model's public unveiling and commercial pricing. Free access paired with a large context window is a strong incentive for developers to test the model on agentic coding tasks, generating exactly the kind of feedback a lab needs before deciding on final pricing and positioning.

Until the developer is identified, treat any capability claims about Ox Alpha as unverified. The "reasoning model" label and stated focus on coding and agentic workloads mirror positioning used by several major labs for their latest generation of models, but without benchmark data or architectural details, there's no way to independently verify performance claims. Enterprises evaluating the model for production workloads should note that data retention policies, while described as training-free, come from an unnamed party bound only by OpenRouter's Stealth Model Terms rather than a named company's standard enterprise agreement.

Related Articles

model release

Z.ai Releases GLM-5.3 with 1M-Token Context and Always-On Reasoning

Z.ai has released GLM-5.3, a large-scale reasoning model aimed at software engineering and long-horizon agent tasks, featuring a 1M-token context window and mandatory reasoning that cannot be disabled. The model is priced at $1.40 per 1M input tokens and $4.40 per 1M output tokens on OpenRouter.

model release

Qwen Launches Qwen3.8 27B, an Open-Weight Vision-Language Model with 262K Context

Qwen has released Qwen3.8 27B, a 27-billion-parameter dense vision-language model with a 262K token context window, available now via OpenRouter at $0.45 per million input tokens and $3.20 per million output tokens.

model release

Z.ai Releases GLM-5.3, Claims Frontier Coding Scores From a 750B-Parameter Model

Z.ai released GLM-5.3, a coding-focused model built on the same base as GLM-5.2 but with substantially extended post-training, and claims it surpasses Moonshot AI's Kimi K3 on many agentic coding benchmarks despite having roughly a third of the parameters. The model is live in Z.ai's coding plan now, with API and open-weight Hugging Face access expected within two weeks.

model release

Z.ai Releases GLM-5.3, Claims Frontier Agentic Coding Performance from 750B-Parameter Model via Post-Training Alone

Z.ai released GLM-5.3, available now in its coding plan, with API access and open weights on Hugging Face to follow within two weeks. The company says the model matches or beats larger frontier systems on agentic coding benchmarks using the same base checkpoint as GLM-5.2, with all gains coming from expanded post-training.

Comments

Loading...