Alibaba releases Qwen3.5 Plus with 1M token context window at $0.40 per million input tokens
Alibaba released an updated version of Qwen3.5 Plus on April 27, 2026, with a 1 million token context window. The multimodal model accepts text, image, and video input and is priced at $0.40 per million input tokens and $2.40 per million output tokens, with tiered pricing above 256K tokens.
Qwen3.5 Plus — Quick Specs
Alibaba releases Qwen3.5 Plus with 1M token context window at $0.40 per million input tokens
Alibaba released an updated version of Qwen3.5 Plus on April 27, 2026, expanding its context window to 1 million tokens. The multimodal model accepts text, image, and video input and produces text output.
Pricing and specifications
The model is priced at $0.40 per million input tokens and $2.40 per million output tokens. According to the model page, tiered pricing applies above 256K tokens, though specific tier rates are not disclosed.
Qwen3.5 Plus supports three input modalities:
- Text
- Image
- Video
The model outputs text only.
Context window expansion
The 1 million token context window represents a significant expansion for the Qwen series. This capacity allows the model to process substantially longer documents, codebases, or multi-turn conversations within a single context.
The model is currently available through OpenRouter, which routes requests across multiple providers with automatic fallbacks. No other distribution channels were disclosed in the announcement.
What this means
Qwen3.5 Plus's pricing positions it in the mid-range of multimodal models with extended context windows. At $0.40 per million input tokens, it's more expensive than text-only models but competitive for multimodal capabilities. The tiered pricing structure suggests Alibaba is targeting use cases that require long context while managing compute costs for shorter inputs. The lack of disclosed benchmark scores or technical specifications makes direct performance comparison with competing models difficult.
Related Articles
OpenAI Cuts GPT-6 Sol and Luna Prices in Half, but Independent Benchmarks Show Flat Performance
OpenAI's GPT-6 Sol and Luna cut input/output token prices in half versus GPT-5.6, with Sol now at $2/$10 per million tokens and Luna at $0.10/$0.50. Independent testing from Artificial Analysis shows intelligence scores barely moved, with regressions on some knowledge-work benchmarks.
OpenRouter Adds DeepSeek Flash Latest Alias With 1M-Token Context Window
OpenRouter has launched deepseek-flash-latest, a persistent endpoint that always points to the current DeepSeek Flash model. It offers a 1,049K token context window, text-and-image input, and pricing of $0.15 per 1M input tokens and $0.60 per 1M output tokens.
OpenAI Python SDK v3.19.0 Adds GCP Storage Support and References Unreleased 'GPT-Rosalind' Model
OpenAI released v3.19.0 of its Python SDK on September 22, 2026, adding GCP external storage support and a code reference to an unannounced research model called GPT-Rosalind. The release also ships five bug fixes covering WebSocket handling, retry logic, and async compatibility.
OpenAI Rolls Out Improved Prompt Caching for GPT-6
OpenAI has updated its prompt caching system for GPT-6, adding explicit cache breakpoints, new diagnostic tools, and finer-grained controls. The company claims the changes improve cache hit rates and reduce both latency and cost for repeated-context API calls.
Comments
Loading...