model release

Alibaba Qwen Releases Qwen3.6 Flash with 1M Context Window at $0.25 per 1M Input Tokens

TL;DR

Alibaba's Qwen team has released Qwen3.6 Flash, a multimodal language model supporting text, image, and video input with a 1 million token context window. The model is priced at $0.25 per 1M input tokens and $1.50 per 1M output tokens, with tiered pricing above 256K tokens.

2 min read
0

Qwen3.6 Flash — Quick Specs

Context window1000K tokens
Input$0.25/1M tokens
Output$1.5/1M tokens

Alibaba Qwen Releases Qwen3.6 Flash with 1M Context Window

Alibaba's Qwen team has released Qwen3.6 Flash, a multimodal language model that processes text, image, and video inputs with a 1 million token context window. Released on April 27, 2026, the model is positioned as a fast, efficient option in the Qwen 3.6 series.

Pricing and Technical Specifications

The model is priced at $0.25 per 1 million input tokens and $1.50 per 1 million output tokens for prompts up to 256K tokens. According to the release information, tiered pricing applies for requests exceeding 256K tokens, though specific rates for higher tiers were not disclosed.

The 1M token context window places Qwen3.6 Flash among models with extended context capabilities, though it falls short of some competitors offering 2M+ token windows. The model supports prompt caching with separate pricing for cache read and cache creation operations.

Multimodal Capabilities

Qwen3.6 Flash handles three input modalities: text, images, and video. This positions it as a general-purpose multimodal model, though specific benchmark scores and performance metrics were not provided in the release announcement.

The model is available through OpenRouter, which routes requests across multiple providers with automatic fallback for uptime optimization. OpenRouter's implementation supports reasoning-enabled features, allowing the model to display step-by-step thinking processes through a reasoning_details array in API responses.

API Integration

Developers can access Qwen3.6 Flash through OpenRouter's normalized API, which maintains compatibility with OpenAI SDK conventions. The platform provides request routing to optimize for prompt size and parameters, with provider fallbacks to maintain service availability.

What This Means

Qwen3.6 Flash represents Alibaba's continued push into competitive AI model pricing while expanding multimodal capabilities. The $0.25 per 1M input tokens rate undercuts several major competitors, though direct performance comparisons remain unclear without published benchmark scores. The tiered pricing structure for larger contexts suggests the model is optimized for shorter interactions, with the 256K threshold marking a significant cost increase point. Video input support is notable, as this capability remains relatively uncommon among broadly available language models.

Related Articles

model release

Microsoft Releases Fara1.5-27B, a 27B Vision-Only Web Browsing Agent with 262K Context

Microsoft Research AI Frontiers has released Fara1.5-27B, a 27-billion-parameter multimodal agent that completes web tasks by reading screenshots and emitting click/type/scroll commands. The model, fine-tuned from Qwen3.5-27B, ships under MIT license with a 262K-token context window and is designed to run alongside Microsoft's MagenticLite sandbox.

model release

Anthropic Launches Claude Opus 5 (Fast) at $10/$50 per Million Tokens, 1M Context Window

Anthropic has released Claude Opus 5 (Fast), a higher-throughput variant of Opus 5 that carries identical capabilities but runs at roughly 2x the price of the standard model. The model ships with a 1 million token context window and is available now through OpenRouter.

model release

Anthropic's Claude Opus 5 Hits 0% Prompt Injection Success Rate in Browser Agent Tests, With Defenses Enabled

Anthropic's system card for Claude Opus 5 reports a 0% prompt injection success rate across 129 browser agent test scenarios when Auto Mode is enabled. On Gray Swan's broader indirect prompt injection benchmark, Opus 5 posted a 2.0% attacker success rate after 15 attempts, the lowest among tested frontier models.

model release

Anthropic Ships Claude Opus 5, Claims Near-Fable Performance at Half the Price

Anthropic released Claude Opus 5 on July 24, 2026, positioning it as a lower-cost alternative to its more expensive Claude Fable 5 model. Independent evaluators Epoch AI and Artificial Analysis report mixed but largely favorable results, with Opus 5 nearly matching Fable 5 on coding benchmarks while cutting cost-per-task by roughly 20%.

Comments

Loading...

Qwen3.6 Flash: 1M Context Multimodal Model Released by Alibaba | TPS