model release

Alibaba Launches Qwen3.7 Flash: 1M-Context Vision-Language Model at $0.03/$0.13 per 1M Tokens

TL;DR

Alibaba has released Qwen3.7 Flash, a vision-language reasoning model with a 1 million token context window aimed at multimodal agents, visual coding, and computer-use tasks. The model is priced at $0.03 per 1M input tokens and $0.13 per 1M output tokens and is available through OpenRouter.

2 min read
0

Qwen3.7 Flash — Quick Specs

Context window1000K tokens
Input$0.03/1M tokens
Output$0.13/1M tokens

Alibaba's Qwen team has released Qwen3.7 Flash, a vision-language reasoning model built for multimodal agents, visual coding, search, and computer interaction tasks. The model is now listed on OpenRouter with a 1 million token context window and pricing of $0.03 per 1M input tokens and $0.13 per 1M output tokens.

What Qwen3.7 Flash Is

According to the OpenRouter listing, Qwen3.7 Flash is designed to handle both text and visual inputs, with stated strengths in object recognition, spatial understanding, and real-world visual perception. This positions the model for use cases that combine language reasoning with vision tasks — such as agents that need to interpret screenshots, navigate user interfaces, or reason about physical layouts.

The model is currently hosted by a single provider on OpenRouter, meaning requests are forwarded directly without multi-provider routing. OpenRouter's platform tracks effective pricing after prompt caching, throughput, latency, time-to-first-token, and uptime, though granular performance numbers were not included in the source listing.

Pricing and Context

At $0.03 per 1M input tokens and $0.13 per 1M output tokens, Qwen3.7 Flash is priced in the budget tier relative to frontier multimodal models from competitors. The 1 million token context window puts it in the same class as long-context models like Gemini 1.5 Pro and Qwen's own long-context variants, enabling use cases such as processing lengthy documents, multi-turn agent sessions, or large codebases alongside visual inputs.

An Unusual Release Date

The OpenRouter listing shows a release date of July 27, 2026 — a date that has not yet occurred. This discrepancy could not be independently verified at time of publication. It may reflect a placeholder, a scheduling artifact in OpenRouter's system, or an error in how the listing was generated. Readers should treat the specific release date with caution until Alibaba or OpenRouter confirms it directly.

No benchmark scores, parameter count, or training data cutoff were disclosed in the available listing. Alibaba has not published a separate model card or technical report for Qwen3.7 Flash as of this writing.

What This Means

Qwen3.7 Flash extends Alibaba's Qwen family further into low-cost, high-context multimodal territory, continuing a pattern of aggressive pricing that has characterized the Qwen line against U.S. competitors. The $0.03/$0.13 per-1M pricing undercuts many comparable vision-language models, which could make it attractive for high-volume agent and automation workloads where cost per request matters more than peak benchmark performance.

However, the lack of disclosed benchmark scores, parameter count, and a verifiable release date leaves significant gaps. Until Alibaba publishes a technical report or the release date discrepancy is resolved, developers evaluating Qwen3.7 Flash for production use should treat performance claims — object recognition, spatial understanding, visual perception — as marketing descriptions rather than validated capabilities. The model's real-world value will depend on independent benchmarking once broader access and documentation become available.

Related Articles

model release

Alibaba's Qwen Releases Qwen-Drive-1.0-4B, a Unified VLM for Autonomous Driving Perception and Planning

Alibaba's Qwen team has released Qwen-Drive-1.0-4B, a 4B-parameter vision-language model built on Qwen3.5 that unifies 3D perception, driving question answering, and motion planning in one framework. The model reports strong open-loop, pseudo-closed-loop, and closed-loop driving benchmark results while claiming minimal loss of general vision-language ability.

model release

InclusionAI Releases Ling 3.0 Flash VL, Adding Vision to Its 124B MoE Model

InclusionAI has released Ling 3.0 Flash VL, a vision-language extension of its 124B total-parameter, 5.5B active Mixture-of-Experts model. The model adds native image and video understanding, supports a 131K token context window, and is priced at $0.06 per 1M input tokens and $0.18 per 1M output tokens via OpenRouter.

model release

Alibaba Open-Sources Qwen3.8-2.4T-A95B, Its First Qwen-Max-Class Model With Public Weights

Alibaba's Qwen team released Qwen3.8-2.4T-A95B on August 12, 2026, the open-weight version of Qwen3.8-Max and the first Qwen-Max-class model made publicly available. The 2.4 trillion-parameter mixture-of-experts model activates only 95 billion parameters per token and supports context windows up to 1 million tokens.

model release

Unverified 'GPT Astra' Model Appears on OpenRouter With 1.05M Token Context, No OpenAI Confirmation

OpenRouter is listing a model called 'OpenAI GPT Astra Latest' with a 1.05 million token context window and $10/$50 per-million-token pricing. OpenAI has made no public announcement, and the listing's own description says it is an auto-redirecting alias rather than a fixed model.

Comments

Loading...