Google releases Gemini 3.5 Flash at half the price of frontier models, announces Omni world model
Google released Gemini 3.5 Flash, priced at half to one-third the cost of comparable frontier models, and announced it will become the default model in the Gemini app globally. The company also unveiled Omni, a world model for simulating physical environments, and Gemini Spark, an AI agent in beta testing.
Gemini 3.5 Flash — Quick Specs
Google releases Gemini 3.5 Flash at half the price of frontier models, announces Omni world model
Google released Gemini 3.5 Flash at its I/O developer conference on May 19, 2026, positioning the model at half to one-third the price of comparable frontier models according to CEO Sundar Pichai. The model will immediately become the default in the Gemini app and AI mode in search globally.
Pricing and performance
Google claims Gemini 3.5 Flash delivers "cutting-edge capabilities" at 50-67% lower cost than competing frontier models. Specific pricing per million tokens was not disclosed. Pichai described the model as "remarkably fast" in a pre-event briefing, with Google stating users "no longer have to trade quality for latency."
The company said it strengthened cybersecurity defenses to make 3.5 Flash "less likely to generate harmful content and mistakenly refuse to answer safe queries." No benchmark scores were provided.
Gemini 3.5 Pro delayed
Gemini 3.5 Pro, the heavier-weight version, is currently in internal use but won't launch publicly until June 2026. Google did not provide technical specifications or performance comparisons for the Pro version.
Omni world model for physical simulation
Google announced Omni, a world model designed to simulate physical environments and predict outcomes based on user actions. The model supports image and audio input and will be integrated into Flash, the Gemini App, Google Flow, and YouTube Shorts.
According to Google, Omni can "edit the action, add in new characters or objects" in user-uploaded videos. World models are primarily used in robotics and gaming applications. DeepMind has researched world models extensively, though Google did not specify how much of that research informed Omni's development.
Gemini Spark agent in beta
Google introduced Gemini Spark, an AI agent that can "reason across information in connected apps" and take actions on behalf of users. The agent launches in beta next week for trusted testers and Google AI Ultra subscribers only. No timeline for general availability was provided.
The agent represents Google's push into agentic AI as the company attempts to create deeper integrations across its product suite.
Market context
The announcements come as OpenAI and Anthropic, both reportedly preparing for IPOs in 2026, dominate AI market attention. Anthropic recently released its Mythos model, which the company claims discovered thousands of previously unknown software vulnerabilities.
Google's capital expenditure on AI infrastructure has increased significantly, with Wall Street expecting the company to demonstrate return on investment through product integrations.
What this means
Google's aggressive pricing on Gemini 3.5 Flash directly challenges OpenAI and Anthropic on cost, a critical factor for developers building at scale. The lack of disclosed benchmarks makes it impossible to verify performance claims against GPT-4 or Claude 3.5 Sonnet. Omni's world model capabilities could differentiate Google in robotics and video editing applications, but commercial viability depends on accuracy and reliability metrics not yet public. The staggered rollout of Spark and delayed Pro model suggest Google is still validating core capabilities before broad deployment.
Related Articles
OpenAI's GPT-6 Astra Reportedly Automates AI Engineering Tasks at Under $6 an Hour, According to Latent Space Testing
A Latent Space report describes GPT-6 Astra, a new OpenAI model the blog says can autonomously handle AI engineering tasks—training models, labeling data, deploying systems—at an estimated cost of under $6 per hour. The claims, including 97.6% on FrontierMath and 99.9% on ARC-AGI-3, come from independent blog testing rather than an official OpenAI announcement.
Google Launches WeatherNext 3, Claims 50% More Accurate Precipitation Forecasts
Google DeepMind and Google Research released WeatherNext 3, a weather AI model trained on real-time geostationary satellite data instead of lagging numerical weather prediction outputs. Google claims up to 50% more accurate day-ahead precipitation forecasts, now rolling out to Search, Maps, and the Gemini app.
Google Launches Gemini 3.8 Flash With Same Pricing as Predecessor, Its Third Flash Model in Six Weeks
Google released Gemini 3.8 Flash, its third Flash model in six weeks, pricing it identically to its predecessor at $0.75 per million input tokens and $3.75 per million output tokens. The launch coincides with a favorable antitrust ruling and public praise from Berkshire Hathaway's Greg Abel, giving Alphabet a stronger narrative after a four-month stock slide.
Google Lists Gemini 3.8 Flash on OpenRouter With 1M-Token Context, September 2026 Release Date
Google's Gemini 3.8 Flash has surfaced on OpenRouter with a 1-million-token context window and discounted pricing of $0.75 per 1M input tokens and $3.75 per 1M output tokens. Google has not issued a separate public announcement, and the listed release date of September 2, 2026 is unusually far out, leaving key details unconfirmed.
Comments
Loading...