model release

Google Releases Gemini 3.7 Flash, Cuts Price in Half Versus 3.6 Flash

TL;DR

Google has released Gemini 3.7 Flash, just three weeks after Gemini 3.6 Flash, claiming substantial gains in coding, web development, and document reasoning. The model launches at an introductory price of $0.75 per 1M input tokens and $3.75 per 1M output tokens — half the cost of its predecessor.

3 min read
0

Google DeepMind released Gemini 3.7 Flash on August 13, 2026, positioning it as its "most intelligent workhorse model yet for coding and agents." The release comes just three weeks after Gemini 3.6 Flash, continuing Google's rapid iteration cadence on the Flash line.

The model launches at an introductory price of $0.75 per 1M input tokens and $3.75 per 1M output tokens — half the per-token cost of 3.6 Flash, according to Google. That pricing is confirmed as an introductory rate available through the end of the year, with future pricing unspecified.

Benchmark claims

Google reports the following gains for 3.7 Flash over 3.6 Flash, based on internal and third-party benchmarks:

  • FrontierCode 1.1 Main: 43.6% vs 34.4%
  • DeepSWE v1.1: 65.3% vs 49.0%
  • WebDev Arena (Arena.ai) Elo: 1588 vs 1538
  • GDP.pdf (complex document processing): 34.0% vs 22.0%
  • AutomationBench (real-world business workflows): 30.4% vs 17.0%

Google claims the model shows stronger first-pass code accuracy, better production-ready code generation, and improved design adherence when generating UIs from screenshots or design systems. In knowledge-dense domains like finance, law, and biosciences, the company says 3.7 Flash delivers materially better reasoning and document-processing accuracy than its predecessor. None of these benchmark figures have been independently verified by third parties as of publication.

Developer experience and rollout

According to Google, 3.7 Flash better handles roadblocks, clarifies ambiguous intent, and follows instructions with higher fidelity than 3.6 Flash, while dedicating more computation to multi-step planning and tool calls. The company frames this as reducing manual oversight and retries in agentic coding workflows.

The model is available immediately through the Gemini API in Google AI Studio, Android Studio, and Google Antigravity for developers; through the Gemini Enterprise Agent Platform and Gemini Enterprise app for businesses; and through Gemini Spark — Google's 24/7 personal AI agent — for Google AI Pro and Ultra subscribers in more than 160 countries. Spark, launched at Google I/O, is being upgraded to run on 3.7 Flash starting today, with Google citing improved tool use across Google Workspace apps for tasks like drafting emails and consolidating files.

Google also states that 3.7 Flash ships with updated Frontier Safety safeguards targeting misuse in CBRN (chemical, biological, radiological, nuclear) domains and cyber offense, consistent with its existing bioresilience and cyber safety programs. Full technical specifications, including exact context window size and parameter count, were not disclosed in the announcement; the company points to a separate model card for additional detail.

What this means

A three-week gap between Flash releases signals Google is now iterating on its lower-cost model tier at a pace closer to weekly software updates than traditional model-release cycles. Halving the price while claiming double-digit benchmark improvements on coding and business-automation tasks is an aggressive move against competitors' cheaper tiers, particularly for developers running high-volume agentic workloads where token costs compound quickly. The benchmark gains, if they hold up under independent testing, would matter most for teams already building coding agents and document-processing pipelines on Gemini Flash — but until third-party evaluation confirms the FrontierCode, DeepSWE, and AutomationBench numbers, these remain Google's own claims. The tight release cadence also raises questions about how much genuine architectural change versus fine-tuning is driving each successive Flash version.

Related Articles

model release

Google Releases Gemini 3.7 Flash With 1M-Token Context and Multimodal Input

Google has released Gemini 3.7 Flash, a multimodal model built for agentic workflows, coding, and multi-step reasoning. It offers a 1,049K token context window and is priced at $0.38 per million input tokens and $1.88 per million output tokens, available now via OpenRouter.

model release

xAI Releases Grok 4.6, a 1.5T-Parameter Model Powering New 'Grok Bot' AI Teammate Product

xAI released Grok 4.6, a confirmed 1.5T-parameter model built on Grok 4.5 with heavier training on long-horizon agentic tasks. It powers the newly launched Grok Bot product and scores 61 on Artificial Analysis's Intelligence Index at $2/$6 per 1M input/output tokens — well below frontier competitors.

model release

xAI's Grok 4.6 Matches Claude and GPT-5.6 on Benchmarks, Costs 60% Less

xAI's Grok 4.6 ties OpenAI's GPT-5.6 Sol on the Artificial Analysis Intelligence Index with a score of 61, trailing only Anthropic's Claude Opus 5 and Claude Fable 5. Pricing remains at $2/$6 per million tokens, undercutting both competitors by more than 60 percent.

model release

MiniMax Releases Music 3, an Open-Weight Model for Generating Full 5-Minute Songs

MiniMax released Music 3, an open-weight music generation model that produces complete songs up to five minutes long from lyrics and text descriptions. The model combines an 8B and 0.6B language model pair with a Flow Matching synthesis system to output 32 kHz stereo audio.

Comments

Loading...

Gemini 3.7 Flash: Google's New Coding & Agent Model | TPS