Anonymous 'Ox Alpha' Reasoning Model Appears on OpenRouter with Free 1M-Token Context
A stealth model called Ox Alpha has appeared on OpenRouter, offering a 1 million token context window at no cost during its preview period. The model's developer remains anonymous, and OpenRouter says it is acting only as a router, not the model's owner or provider.
Stealth Model Ox Alpha Launches Free on OpenRouter
An unidentified AI developer has released a model called Ox Alpha through OpenRouter, positioning it as a reasoning system built for coding, sustained agentic work, and production workloads. The model is currently free to use and offers a 1 million token context window, according to its OpenRouter listing.
Ox Alpha appeared on the platform on August 20, 2026, though details about its architecture, parameter count, and training data remain undisclosed. OpenRouter explicitly states it is not the model's developer, owner, or provider — it is only routing requests to a third-party operator who has chosen to remain anonymous during this preview phase.
What's Known
According to OpenRouter's listing, Ox Alpha is designed for:
- Long-horizon software engineering tasks
- Complex reasoning workflows
- Multimodal workflows that combine text with visual context
The model supports a 1,000,000 token context window and is currently offered at no cost — no input or output pricing has been set, consistent with a preview or evaluation release. Latency, throughput, and uptime metrics are not yet available due to limited usage data, though OpenRouter reports 100% availability over the last 24 hours across all locations, with three days of uptime history so far.
No benchmark scores, parameter counts, or training cutoff dates have been disclosed. The provider has not confirmed its identity, and no company in the AI industry has yet claimed Ox Alpha as its own.
Data Handling
OpenRouter notes that prompts and completions sent to Ox Alpha are retained by the anonymous provider but are not used for training. All other usage is governed by OpenRouter's Stealth Model Terms, a framework the platform applies to unidentified models during preview windows before a public launch.
What This Means
Stealth releases on OpenRouter have become a common way for AI labs — established or new — to gather real-world usage data and community feedback before a model's public unveiling and commercial pricing. Free access paired with a large context window is a strong incentive for developers to test the model on agentic coding tasks, generating exactly the kind of feedback a lab needs before deciding on final pricing and positioning.
Until the developer is identified, treat any capability claims about Ox Alpha as unverified. The "reasoning model" label and stated focus on coding and agentic workloads mirror positioning used by several major labs for their latest generation of models, but without benchmark data or architectural details, there's no way to independently verify performance claims. Enterprises evaluating the model for production workloads should note that data retention policies, while described as training-free, come from an unnamed party bound only by OpenRouter's Stealth Model Terms rather than a named company's standard enterprise agreement.
Related Articles
inclusionAI releases Ling 3.1 Flash: 560B MoE, 25B active, 262K context, free on OpenRouter
inclusionAI has released Ling 3.1 Flash, a hybrid reasoning mixture-of-experts model with 560B total and 25B active parameters and a 262K-token context window. It is listed as free on OpenRouter through NovitaAI. No benchmark scores have been published on the listing.
Unbiased releases Pareto 26.10 Preview: 1M context, $0.80/$3.20 per 1M tokens on OpenRouter
Unbiased has listed Pareto 26.10 Preview on OpenRouter, a multimodal composite model with a 1.0M-token context window priced at $0.80 input and $3.20 output per 1M tokens. The company says it targets research, coding, and agentic workflows, and warns the preview may change without notice. No benchmark scores have been published.
China Telecom's Xing4.0-29B-A4B: 29B MoE, 4B Active, 256K Context, Trained Fully on Ascend NPUs
China Telecom AI's Xing4.0-29B-A4B (formerly the TeleChat line) is a mixture-of-experts model with 29B total and 4B active parameters and a native 256K context window, extensible to 512K. The company claims it is the first model of this scale trained entirely on Ascend NPUs with MindSpore. Community GGUF quantizations from Venastine-Research are already available.
Cloudflare releases Clef decision models, claims 39 ms median latency vs. 524 ms for TypeSafe's Jev
Cloudflare has released Clef and Clef-flash, two open-weight decision models that return probabilities over predefined answer options instead of generating text. The company claims median latencies of 39 ms and 209 ms, against just over 524 ms for TypeSafe AI's Jev. Both support text and images and are API-compatible with Jev.
Comments
Loading...