Alibaba Qwen Releases Qwen3.6 Flash with 1M Context Window at $0.25 per 1M Input Tokens
Alibaba's Qwen team has released Qwen3.6 Flash, a multimodal language model supporting text, image, and video input with a 1 million token context window. The model is priced at $0.25 per 1M input tokens and $1.50 per 1M output tokens, with tiered pricing above 256K tokens.
Qwen3.6 Flash — Quick Specs
Alibaba Qwen Releases Qwen3.6 Flash with 1M Context Window
Alibaba's Qwen team has released Qwen3.6 Flash, a multimodal language model that processes text, image, and video inputs with a 1 million token context window. Released on April 27, 2026, the model is positioned as a fast, efficient option in the Qwen 3.6 series.
Pricing and Technical Specifications
The model is priced at $0.25 per 1 million input tokens and $1.50 per 1 million output tokens for prompts up to 256K tokens. According to the release information, tiered pricing applies for requests exceeding 256K tokens, though specific rates for higher tiers were not disclosed.
The 1M token context window places Qwen3.6 Flash among models with extended context capabilities, though it falls short of some competitors offering 2M+ token windows. The model supports prompt caching with separate pricing for cache read and cache creation operations.
Multimodal Capabilities
Qwen3.6 Flash handles three input modalities: text, images, and video. This positions it as a general-purpose multimodal model, though specific benchmark scores and performance metrics were not provided in the release announcement.
The model is available through OpenRouter, which routes requests across multiple providers with automatic fallback for uptime optimization. OpenRouter's implementation supports reasoning-enabled features, allowing the model to display step-by-step thinking processes through a reasoning_details array in API responses.
API Integration
Developers can access Qwen3.6 Flash through OpenRouter's normalized API, which maintains compatibility with OpenAI SDK conventions. The platform provides request routing to optimize for prompt size and parameters, with provider fallbacks to maintain service availability.
What This Means
Qwen3.6 Flash represents Alibaba's continued push into competitive AI model pricing while expanding multimodal capabilities. The $0.25 per 1M input tokens rate undercuts several major competitors, though direct performance comparisons remain unclear without published benchmark scores. The tiered pricing structure for larger contexts suggests the model is optimized for shorter interactions, with the 256K threshold marking a significant cost increase point. Video input support is notable, as this capability remains relatively uncommon among broadly available language models.
Related Articles
Alibaba Releases Qwen3.8 Max (0902), a 2.4-Trillion-Parameter MoE Model With 1M-Token Context
Alibaba's Qwen team released Qwen3.8 Max (0902), a 2.4-trillion-parameter mixture-of-experts model with a 1M-token context window that accepts text, image, and video input. The snapshot is post-trained for coding, agentic workflows, and long-horizon task execution, priced at $2/$6 per 1M input/output tokens.
Alibaba's Qwen Releases Qwen-Drive-1.0-4B, a Unified VLM for Autonomous Driving Perception and Planning
Alibaba's Qwen team has released Qwen-Drive-1.0-4B, a 4B-parameter vision-language model built on Qwen3.5 that unifies 3D perception, driving question answering, and motion planning in one framework. The model reports strong open-loop, pseudo-closed-loop, and closed-loop driving benchmark results while claiming minimal loss of general vision-language ability.
Alibaba Open-Sources Qwen3.8-2.4T-A95B, Its First Qwen-Max-Class Model With Public Weights
Alibaba's Qwen team released Qwen3.8-2.4T-A95B on August 12, 2026, the open-weight version of Qwen3.8-Max and the first Qwen-Max-class model made publicly available. The 2.4 trillion-parameter mixture-of-experts model activates only 95 billion parameters per token and supports context windows up to 1 million tokens.
Alibaba Releases Qwen-Drive 1.0, an Open Driving Model That Explains Its Own Decisions
Alibaba has released Qwen-Drive 1.0, a driving model built on Qwen3.5-4B that handles spatial perception, route planning, and cockpit dialogue in a single system. Reinforcement learning cut the rate of off-road driving errors in simulation from 24 percent to 12 percent, though the model's stated reasoning doesn't always match its actual maneuvers.
Comments
Loading...