Google Releases Gemini 3.7 Flash, Cuts Price in Half Versus 3.6 Flash
Google has released Gemini 3.7 Flash, just three weeks after Gemini 3.6 Flash, claiming substantial gains in coding, web development, and document reasoning. The model launches at an introductory price of $0.75 per 1M input tokens and $3.75 per 1M output tokens — half the cost of its predecessor.
Google DeepMind released Gemini 3.7 Flash on August 13, 2026, positioning it as its "most intelligent workhorse model yet for coding and agents." The release comes just three weeks after Gemini 3.6 Flash, continuing Google's rapid iteration cadence on the Flash line.
The model launches at an introductory price of $0.75 per 1M input tokens and $3.75 per 1M output tokens — half the per-token cost of 3.6 Flash, according to Google. That pricing is confirmed as an introductory rate available through the end of the year, with future pricing unspecified.
Benchmark claims
Google reports the following gains for 3.7 Flash over 3.6 Flash, based on internal and third-party benchmarks:
- FrontierCode 1.1 Main: 43.6% vs 34.4%
- DeepSWE v1.1: 65.3% vs 49.0%
- WebDev Arena (Arena.ai) Elo: 1588 vs 1538
- GDP.pdf (complex document processing): 34.0% vs 22.0%
- AutomationBench (real-world business workflows): 30.4% vs 17.0%
Google claims the model shows stronger first-pass code accuracy, better production-ready code generation, and improved design adherence when generating UIs from screenshots or design systems. In knowledge-dense domains like finance, law, and biosciences, the company says 3.7 Flash delivers materially better reasoning and document-processing accuracy than its predecessor. None of these benchmark figures have been independently verified by third parties as of publication.
Developer experience and rollout
According to Google, 3.7 Flash better handles roadblocks, clarifies ambiguous intent, and follows instructions with higher fidelity than 3.6 Flash, while dedicating more computation to multi-step planning and tool calls. The company frames this as reducing manual oversight and retries in agentic coding workflows.
The model is available immediately through the Gemini API in Google AI Studio, Android Studio, and Google Antigravity for developers; through the Gemini Enterprise Agent Platform and Gemini Enterprise app for businesses; and through Gemini Spark — Google's 24/7 personal AI agent — for Google AI Pro and Ultra subscribers in more than 160 countries. Spark, launched at Google I/O, is being upgraded to run on 3.7 Flash starting today, with Google citing improved tool use across Google Workspace apps for tasks like drafting emails and consolidating files.
Google also states that 3.7 Flash ships with updated Frontier Safety safeguards targeting misuse in CBRN (chemical, biological, radiological, nuclear) domains and cyber offense, consistent with its existing bioresilience and cyber safety programs. Full technical specifications, including exact context window size and parameter count, were not disclosed in the announcement; the company points to a separate model card for additional detail.
What this means
A three-week gap between Flash releases signals Google is now iterating on its lower-cost model tier at a pace closer to weekly software updates than traditional model-release cycles. Halving the price while claiming double-digit benchmark improvements on coding and business-automation tasks is an aggressive move against competitors' cheaper tiers, particularly for developers running high-volume agentic workloads where token costs compound quickly. The benchmark gains, if they hold up under independent testing, would matter most for teams already building coding agents and document-processing pipelines on Gemini Flash — but until third-party evaluation confirms the FrontierCode, DeepSWE, and AutomationBench numbers, these remain Google's own claims. The tight release cadence also raises questions about how much genuine architectural change versus fine-tuning is driving each successive Flash version.
Related Articles
Google Launches Gemini 3.8 Flash TTS and Flash-Lite TTS with Voice Creation from Text Prompts
Google DeepMind has released Gemini 3.8 Flash TTS and Gemini 3.8 Flash-Lite TTS, text-to-speech models that generate custom voices from natural language prompts and support line-by-line performance direction. The models top Hume AI's Voice Design Benchmark at 71.4 and claim first and second place on its Overall Quality Index.
Google DeepMind's New Chief Prioritizes Fast Gemini 4 Release Over AGI Debate
Google DeepMind's new head Koray Kavukcuoglu says Gemini 4 is in early post-training and could ship well before year-end, following the quiet cancellation of Gemini 3.5 Pro. He downplayed the AGI question that defined predecessor Demis Hassabis's tenure, calling it 'not the right conversation.'
Google Launches Gemini 3.8 Flash TTS: Voice Cloning and Text-Described Voices for $9-18 per Million Audio Tokens
Google has released Gemini 3.8 Flash TTS and Flash-Lite TTS, two speech generation models that let users design voices from text descriptions or clone a voice from a 30-second sample. Both support over 100 languages and roll out now through the Gemini API and Google AI Studio.
Anthropic Launches Claude Opus 5.5 at 20% Lower List Price, Claims Parity with Claude Fable 5.1
Anthropic released Claude Opus 5.5, the first model in its new 5.5 family, cutting list pricing 20% to $4/$20 per 1M input/output tokens while claiming performance on par with Claude Fable 5.1. Independent analysis shows the cost savings largely disappear at maximum reasoning effort due to higher token consumption.
Comments
Loading...