model release

Google DeepMind Ships Gemini 3.8 Flash and a Cybersecurity Variant, Third Flash Release in Six Weeks

TL;DR

Google DeepMind released Gemini 3.8 Flash and a specialized cybersecurity variant, Gemini 3.8 Flash Cyber, its third Flash-tier launch in six weeks. Pricing stays at $0.75 per million input tokens and $3.75 per million output tokens, matching the prior 3.7 Flash release.

3 min read
0

Google DeepMind released Gemini 3.8 Flash and a specialized variant, Gemini 3.8 Flash Cyber, on September 2, 2026 — the company's third Flash-tier model launch in six weeks, following Gemini 3.7 Flash three weeks earlier.

Gemini 3.8 Flash is priced identically to its predecessor: $0.75 per million input tokens and $3.75 per million output tokens. Google DeepMind describes it as its "most intelligent workhorse model," with gains concentrated in software engineering, agentic task execution, and multi-step reasoning.

Benchmark claims

According to Google DeepMind, Gemini 3.8 Flash scores 54.9% on HLE-Verified, a benchmark spanning STEM, humanities, and professional reasoning tasks. The company also claims the model outperforms most larger frontier models on DeepSWE v1.1, a long-horizon software engineering benchmark, and leads on Vals Finance Agent V2 and Harvey's Legal Agent Benchmark — both domain-specific evaluations for financial and legal reasoning.

Google DeepMind attributes the gains to the model "working harder": executing additional reasoning steps and calling tools iteratively at higher effort levels, which can increase token consumption. Developers seeking lower compute overhead can select reduced effort levels or continue using Gemini 3.7 Flash, which remains supported.

Gemini 3.8 Flash Cyber

The cybersecurity variant is restricted to a new "Fairwind Program" covering government defenders, critical infrastructure operators, and software maintainers — not general API access. Google DeepMind claims the model achieves frontier-level performance on CyberGym, an industry benchmark for autonomous vulnerability discovery, and exceeds a 70% success rate on an internal benchmark spanning 20 programming languages.

On CWE-Bench, an external patching benchmark run by Collinear, the company reports a pass@1 of 47.2%, close to a competing frontier model's 47.8%, at what it describes as significantly lower cost. Google DeepMind also cites internal deployment results: its Chrome Security team reportedly saw 2.6 times more correct patches than "the best commercial models," security firm Wiz reported 7.5–9.7% higher recall at 2.3–5.2x lower cost on penetration-testing benchmarks, and Google's Cloud Vulnerability Research team says it found a critical vulnerability in under two hours using the model.

Safety measures

Both models ship with mitigations against CBRN and cyber-offense misuse under Google's Frontier Safety Framework, with Gemini 3.8 Flash Cyber operating under more permissive cyber-defense-specific controls limited to vetted users. Google DeepMind also claims improved prompt-injection robustness as measured by third-party evaluator Gray Swan.

Availability

Gemini 3.8 Flash is available now via the Gemini API, Google AI Studio, Android Studio, Google Antigravity, and Stitch, plus Gemini Enterprise and consumer surfaces including the Gemini app, AI Mode in Google Search, and Gemini in Google Sheets. Gemini 3.8 Flash Cyber access requires application through the Fairwind Program.

What this means

Three Flash-tier releases in six weeks signals Google DeepMind is prioritizing rapid iteration over infrequent, large jumps — betting that continuous small gains at fixed low pricing outcompete slower frontier-model cycles. Holding price steady while claiming performance parity with costlier models is a direct pricing pressure play against OpenAI and Anthropic's flagship tiers. The Cyber variant's gated release, rather than open API access, reflects a cautious dual-use posture: Google is willing to claim frontier-level vulnerability discovery capability publicly while restricting who can actually use it, an approach likely to become standard for offensive-adjacent security models industry-wide.

Related Articles

model release

OpenAI Releases GPT-6 Astra, First Model to Cross 'Critical' Cybersecurity Threshold

OpenAI has begun rolling out GPT-6 Astra, the first model to reach the company's internal 'Critical' cybersecurity threshold. Access is being phased, with companies in OpenAI's Daybreak cybersecurity program getting priority following added safeguards after a prior model containment breach.

model release

OpenAI Launches GPT-6 Astra, Matches Claude Fable Pricing at $10/$50 per Million Tokens

OpenAI has begun rolling out GPT-6 Astra, priced at $10/million input and $50/million output tokens to match Claude Fable. The model claims a 99.9% score on ARC-AGI 3 using a custom harness and leads on security and long-context benchmarks, though it trails Fable on general intelligence rankings.

model release

OpenAI Launches GPT-6 Astra, Says the Model May Already Qualify as AGI

OpenAI has released GPT-6 Astra, its most capable model yet, with benchmark scores the company says surpass GPT-5.6 Sol and Anthropic's Fable 5 models. President Greg Brockman called it a step into the 'AGI era,' though OpenAI acknowledges there's no agreed-upon threshold for that term.

model release

OpenAI Releases Astra, Claims New Flagship Model Beats Rivals on Coding and Cybersecurity Benchmarks

OpenAI released Astra on Thursday, calling it its most capable and most aligned model yet. The model uses a reasoning technique called 'opaque recurrence' that critics say reduces visibility into its chain of thought.

Comments

Loading...