model release

Alibaba Releases Qwen3.8 Max, a Multimodal Reasoning Model with 1M Token Context

TL;DR

Alibaba has moved Qwen3.8 Max out of preview into general availability, positioning it as the flagship of the Qwen3.8 series with a 1 million token context window and multimodal input support. The model is priced at $2.00 per million input tokens and $6.00 per million output tokens via OpenRouter.

2 min read
0

Alibaba's Qwen team has released Qwen3.8 Max, the general-availability successor to the Qwen3.8 Max Preview model. The release is now accessible through OpenRouter's API under the identifier qwen/qwen3.8-max.

Key Specifications

Qwen3.8 Max ships with a 1,000K (1 million) token context window, according to Alibaba. The model accepts text, image, and video inputs and produces text outputs, making it a multimodal reasoning model rather than a text-only system.

Pricing on OpenRouter is set at $2.00 per million input tokens and $6.00 per million output tokens.

Alibaba describes Qwen3.8 Max as the flagship model in the Qwen3.8 series, intended according to the company for complex reasoning and visual understanding tasks. The company has not disclosed the model's parameter count, training data cutoff date, or benchmark scores at time of publication.

What Changed From Preview

This release supersedes the Qwen3.8 Max Preview version. Alibaba has not published a detailed changelog specifying architectural or capability differences between the preview and general-availability builds. Without disclosed benchmark comparisons, the practical improvements over the preview remain according to Alibaba's release notes rather than independently verified.

Availability

The model is live now through OpenRouter, giving developers immediate API access without a separate preview or waitlist process. This mirrors Alibaba's pattern with prior Qwen releases, where preview models are used to gather feedback before a stable, generally available version ships with finalized pricing.

What this means

A 1 million token context window puts Qwen3.8 Max in the same tier as long-context leaders like Gemini's largest context configurations, at least on paper. Combined with multimodal input (text, image, video) and reasoning-oriented positioning, Alibaba is clearly targeting enterprise and developer use cases that require processing large documents, codebases, or video alongside complex reasoning chains.

The pricing — $2.00/M input and $6.00/M output — sits in a competitive but not aggressively cheap tier compared to other frontier-adjacent models available on OpenRouter. Without published benchmark scores (MMLU, GPQA, or coding evaluations), buyers have no independent way to verify Alibaba's reasoning and visual-understanding claims against competitors like GPT-4o, Claude, or Gemini 2.0 series models.

The bigger open question is whether the jump from "Preview" to general availability reflects meaningful capability gains or primarily a stability and SLA commitment for production use. Until Alibaba or third parties publish comparative benchmarks, developers evaluating Qwen3.8 Max should treat the reasoning and multimodal claims as unverified and test against their own workloads before committing.

Related Articles

model release

Anonymous Provider Launches Union Alpha, a Free 262K-Context Multimodal Model on OpenRouter

A third-party provider using the alias 'Stealth' has released Union Alpha on OpenRouter, a multimodal model with a 262K context window, currently free to use during its preview period. The model's developer remains anonymous, and OpenRouter states it is not the model's owner or operator.

model release

Unverified 'GPT Astra' Model Appears on OpenRouter With 1.05M Token Context, No OpenAI Confirmation

OpenRouter is listing a model called 'OpenAI GPT Astra Latest' with a 1.05 million token context window and $10/$50 per-million-token pricing. OpenAI has made no public announcement, and the listing's own description says it is an auto-redirecting alias rather than a fixed model.

model release

OpenRouter Lists 'GPT Sol Latest' — An Alias Pointer to OpenAI's Newest Sol-Family Model, Not a Standalone Release

OpenRouter has added a listing called '~openai/gpt-sol-latest,' described as an alias that always points to the newest model in an undisclosed 'GPT Sol' family from OpenAI. The listing shows a 1050K token context window and pricing of $2.00 per million input tokens and $10.00 per million output tokens, but OpenAI has not publicly confirmed a model line by this name.

model release

Unverified 'GPT Terra' Model Surfaces on OpenRouter With 1.05M-Token Context, No OpenAI Confirmation

OpenRouter's catalog lists '~openai/gpt-terra-latest,' an alias pointing to what it describes as the newest model in an unannounced 'GPT Terra' family, with a 1.05 million token context window and $2/$12 per-million-token pricing. OpenAI has made no public statement confirming the model's existence.

Comments

Loading...