Google releases Gemini 3.1 Flash Image, claims Pro-level quality at $0.50 per 1M tokens
Google has released Gemini 3.1 Flash Image, internally codenamed "Nano Banana 2," an image generation and editing model with a 131K context window. The model is priced at $0.50 per 1M input tokens and $3 per 1M output tokens.
Gemini 3.1 Flash Image — Quick Specs
Google releases Gemini 3.1 Flash Image, claims Pro-level quality at $0.50 per 1M tokens
Google has released Gemini 3.1 Flash Image, internally codenamed "Nano Banana 2," an image generation and editing model with a 131,000 token context window. The model is priced at $0.50 per million input tokens and $3 per million output tokens.
According to Google, the model delivers "Pro-level visual quality at Flash speed," positioning it as a faster, more cost-efficient alternative to its premium image models. The company claims the model combines advanced contextual understanding with fast inference, making complex image generation and iterative edits more accessible.
Technical specifications
Gemini 3.1 Flash Image supports:
- Context window: 131,000 tokens
- Multimodal input/output (image generation and editing)
- Configurable aspect ratios via the
image_configAPI parameter - Released: June 18, 2026
The model is available through OpenRouter, which routes requests across multiple hosting providers based on performance and pricing optimization.
Pricing comparison
At $0.50 per 1M input tokens and $3 per 1M output tokens, Gemini 3.1 Flash Image is positioned in the mid-tier pricing range for image generation models. OpenRouter reports that effective pricing can be 60-80% lower when prompt caching is applied for repeated context.
The pricing structure suggests Google is targeting production workloads that require both quality and cost efficiency, particularly for applications involving iterative image editing where context reuse is common.
What this means
Gemini 3.1 Flash Image represents Google's push into faster, more affordable image generation without sacrificing quality claims. The 131K context window is notably large for an image model, potentially enabling more complex multi-turn editing workflows. However, Google has not released benchmark comparisons against competing image models like DALL-E 3, Midjourney, or Stable Diffusion variants, making it difficult to independently verify the "Pro-level quality" claim. The model's real-world performance and adoption will depend on how it stacks up in user testing against established alternatives.
Related Articles
Mistral AI Releases Shieldstral-1.0-3B, a 3B-Parameter Policy-Adaptive Safety Classifier
Mistral AI has released Shieldstral-1.0-3B, a compact open-weight safety classifier that evaluates text and images against natural-language policies specified at inference time. The 3B model runs on a single GPU and reports F1 scores competitive with or exceeding larger moderation models like LlamaGuard-4-12B and GPT-OSS-Safeguard-20B on multiple benchmarks.
Mistral Releases Shieldstral, a 3B Open-Weights Safety Classifier That Matches Models 7x Its Size
Mistral has released Shieldstral, a 3B open-weights safety classifier that reframes content moderation as a policy-adaptive question-answering task. The model claims to match or outperform guard models up to 7x its size on text safety and multimodal benchmarks, and runs on a single 16GB GPU.
OpenAI Halts Parts of Astra Model Development After It Hit 'Critical' Cybersecurity Threshold
OpenAI disclosed that its in-development Astra model showed cyberattack capabilities strong enough that it cannot rule out a 'Critical' risk classification. The company has paused related internal activity and added security controls under its Preparedness Framework.
Mistral's 3B-Parameter Shieldstral Matches 20B Safety Model on Text Benchmarks
Mistral's new Shieldstral, a 3-billion-parameter open-weight safety classifier, posts an 84.9% F1 score on text benchmarks—tying OpenAI's GPT-OSS-Safeguard-20B, a model roughly seven times larger. The model lets operators define safety rules at runtime using plain-language yes/no questions instead of fixed taxonomies.
Comments
Loading...