model release

Google releases Gemma 4 E2B, optimized to run natively on Pixel 10's Tensor G5 TPU

TL;DR

Google has released Gemma 4 E2B for TPU, a variant of its open-source Gemma 4 model optimized to run natively on the Tensor G5 chip in Pixel 10 devices. The multimodal model enables completely offline AI chat, image recognition, and audio transcription on Pixel 10, 10 Pro, 10 Pro XL, and 10 Pro Fold.

2 min read
0

Google releases Gemma 4 E2B, optimized to run natively on Pixel 10's Tensor G5 TPU

Google announced Gemma 4 E2B for TPU today, a variant of its open-source Gemma 4 model designed to run natively on the Tensor Processing Unit in Pixel 10 devices. The announcement came at I/O Connect India, following a similar satellite event in Berlin last week.

Model specifications

Gemma 4 E2B runs on the Tensor G5's TPU and is supported on four devices: Pixel 10, 10 Pro, 10 Pro XL, and 10 Pro Fold. Google describes it as "state-of-the-art, powerful, yet remarkably lightweight," though specific parameter counts and benchmark scores were not disclosed.

The model is based on Gemma 4, which Google first introduced in April as the foundation for the upcoming Gemini Nano 4. Gemma is Google's series of open models designed for on-device execution.

Multimodal capabilities

Gemma 4 E2B supports three fully offline modes:

  • AI Chat: On-device conversations with no internet connection required
  • Ask Image: Object, plant, and issue identification from photos
  • Ask Audio: Private audio transcription for lectures and notes

Google demonstrated "Mobile Actions" that allow users to control core phone functions like WiFi and maps through voice or text commands, all processed locally.

Real-world applications

Google highlighted two specific use cases:

Retail: Converting recipe ideas into localized in-store shopping maps completely offline, allowing customers to navigate stores without internet connectivity.

Automotive: Providing mechanics with immediate visual diagnostics from photos of faulty parts, enabling on-the-spot troubleshooting.

What this means

This release represents Google's push to run increasingly capable AI models entirely on-device, addressing privacy concerns and enabling functionality in areas with poor connectivity. By optimizing specifically for the Tensor G5's TPU architecture, Google is differentiating its Pixel hardware through exclusive AI capabilities that competitors cannot easily replicate. The retail and automotive examples suggest Google is targeting enterprise deployments where offline operation is critical, not just consumer use cases. However, without published benchmarks or comparisons to cloud-based alternatives, the actual performance trade-offs of this on-device approach remain unclear.

Related Articles

model release

Google Releases TimesFM-3, a 330M-Parameter Model That Forecasts Sales Using Weather and Discount Data

Google Research has released TimesFM-3, a 330-million-parameter time series forecasting model that predicts outcomes like sales by combining related variables, historical data, and known future events such as discounts or weather. The model claims top rankings on three benchmarks against Amazon's Chronos-2 and the Toto-2.0 family.

model release

Unverified 'GPT Astra' Model Appears on OpenRouter With 1.05M Token Context, No OpenAI Confirmation

OpenRouter is listing a model called 'OpenAI GPT Astra Latest' with a 1.05 million token context window and $10/$50 per-million-token pricing. OpenAI has made no public announcement, and the listing's own description says it is an auto-redirecting alias rather than a fixed model.

model release

OpenRouter Lists 'GPT Sol Latest' — An Alias Pointer to OpenAI's Newest Sol-Family Model, Not a Standalone Release

OpenRouter has added a listing called '~openai/gpt-sol-latest,' described as an alias that always points to the newest model in an undisclosed 'GPT Sol' family from OpenAI. The listing shows a 1050K token context window and pricing of $2.00 per million input tokens and $10.00 per million output tokens, but OpenAI has not publicly confirmed a model line by this name.

model release

AllSpark's Iris-mini and Iris-pro Top Open-Weight Search Agent Benchmarks

Chinese lab AllSpark has released Iris-mini and Iris-pro, two open-weight search agents built on Qwen3 models that claim the top spot among open-weight systems in their size classes on four research benchmarks. The release includes model weights, an agent harness, and evaluation code, with training pipelines to follow.

Comments

Loading...