Gemini
50 articles tagged with Gemini
Google Research's WikiSkill Framework Boosts AI Agent Performance Up to 23 Points by Building Persistent Memory of Past
Google Research has introduced WikiSkill, a framework that lets AI agents build a persistent, growing knowledge base from past task attempts instead of discarding what they learn after each run. Tested across five models and five benchmarks, WikiSkill lifted average scores by 14 to 24 percentage points over baseline agents with no skill memory.
Google Launches Gemini 3.5 Transcribe with 4.0% Word Error Rate Across 85 Languages
Google has released Gemini 3.5 Transcribe, a speech-to-text model that automatically detects 85 languages, removes filler words, and corrects misspoken phrases. The company claims a 4.0 percent word error rate for streaming audio and 70 percent lower latency than its predecessor, Chirp 3.
Google Launches Gemini 3.5 Transcribe, a Speech-to-Text Model That Cleans Up Rambling Speech
Google has released Gemini 3.5 Transcribe, a new speech-to-text model that automatically detects over 85 languages, removes filler words, and structures unstructured speech into clean text. The model powers Android's Rambler feature and is rolling out to Chrome, Docs, Gmail, and other Google products.
Google DeepMind Launches Gemini 3.5 Transcribe, Claims 2.6% Word Error Rate in Testing
Google DeepMind has released Gemini 3.5 Transcribe, a speech-to-text model available via two APIs for real-time streaming and pre-recorded audio. According to Artificial Analysis benchmarks cited by Google, the model achieves a 2.6% word error rate for non-streaming transcription and 4.0% for streaming.
Google Launches Gemini 3.5 Transcribe with 2.6% Word Error Rate, Powers Gboard Rambler
Google has released Gemini 3.5 Transcribe, a speech-to-text model claiming a 4.0% word error rate in streaming mode and 2.6% in non-streaming mode, according to benchmarks from Artificial Analysis. The model already powers Gboard Rambler on Android and the Gemini app for macOS, with Chrome support coming next.
Google Launches Gemini Enterprise for Legal, Connecting AI to Contract and Research Platforms
Google Cloud has launched Gemini Enterprise for Legal, a preview product that connects its Gemini models to legal software like iManage, DocuSign, Everlaw, RelativityOne, and Harvey via MCP connectors. The tool automates contract review, legal research, and regulatory tracking while respecting existing permission structures.
Google Tests 'Device Help' Gemini Tool Exclusively on Pixel 11 Pro
A new 'Device Help' tool has appeared in the Gemini app's plus menu on Pixel 11 Pro devices running Google app beta 17.52. The Labs-badged feature offers conversational assistance for settings, troubleshooting, and device management, but is not available on the base Pixel 11 or older phones.
Pixel 11's Gboard Adds Gemini-Powered 'Personalized Suggestions' That Read Screen Context
Google has updated Gboard's Writing tools on the Pixel 11 with a new 'Suggested' tab powered by Gemini Intelligence. The feature drafts personalized message suggestions using typing history and on-screen context, processed via Google's Private AI Compute.
Google Adds AI Study Tools with Interactive 3D Simulations to Search and Gemini
Google announced new study features across Search and Gemini, including AI-generated interactive visuals, 3D simulations, customized practice quizzes, and a dedicated learning hub. The rollout intensifies competition with OpenAI and education startups like Knowt and Gauth for student users.
Google Workspace Grants Gemini Default Access to Gmail, Docs, Calendar, and Chat Data
Google Workspace ships with Gemini's access to Gmail, Docs, Calendar, Chat, Drive, and Meet turned on by default, using real-time retrieval-augmented generation rather than stored training data. Admins can disable these 'Workspace Intelligence Sources' organization-wide through the admin.google.com console, though per-user controls remain limited.
Google to Let Users Remove Visible AI Watermarks From Nano Banana, Omni, Lyria Content
Google VP Josh Woodward announced a new toggle that lets users remove visible sparkle-icon watermarks from AI-generated content made with Nano Banana, Omni, and Lyria. The invisible SynthID watermark and C2PA metadata will remain unaffected, and the toggle won't roll out in the EU or South Korea where visible labeling is legally required.
Google Lets Users Turn Off Visible Watermarks on Nano Banana, Omni, and Lyria Outputs
Google announced users can now toggle off visible watermarks on AI-generated images, video, and songs from its Nano Banana, Omni, and Lyria models. Invisible SynthID watermarks and C2PA metadata remain in place for transparency.
Google Rolls Out Gemini 3.7 Flash to Chat Interface, Adds Watermark Toggle
Gemini 3.7 Flash has replaced 3.6 Flash in the Gemini app's model picker on Android, iOS, web, and macOS, following its debut in the Spark agent a day earlier. The app also now lets users turn off visible corner watermarks on AI-generated images, video, and music.
Vercel AI SDK Patch Adds Support for Unreleased 'gemini-3.7-flash' Model ID
Vercel shipped a patch release of @ai-sdk/google-vertex (v4.0.182) that adds support for a model identifier called 'gemini-3.7-flash.' Google has not publicly announced this model, and no official details on pricing, context window, or benchmarks exist yet.
Google DeepMind Ships Gemini 3.7 Flash, Closing Gap With Claude 4.8 and GPT-5.5
Google DeepMind has released Gemini 3.7 Flash, a new entry in its fast-tier model line that reportedly closes a performance gap that opened up under Gemini 3.5 and 3.6 Flash against Anthropic's Claude 4.8+ and OpenAI's GPT-5.5+ series. Full pricing and benchmark details have not yet been disclosed.
Google Ships Gemini 3.7 Flash, Cuts Price 50% Just Three Weeks After 3.6 Flash
Google released Gemini 3.7 Flash just three weeks after its predecessor, posting sharp gains on coding benchmarks while cutting launch pricing in half to $0.75 per million input tokens and $3.75 per million output tokens.
Google Releases Gemini 3.7 Flash, Cuts Price in Half Versus 3.6 Flash
Google has released Gemini 3.7 Flash, just three weeks after Gemini 3.6 Flash, claiming substantial gains in coding, web development, and document reasoning. The model launches at an introductory price of $0.75 per 1M input tokens and $3.75 per 1M output tokens — half the cost of its predecessor.
Google Releases Gemini 3.7 Flash With 1M-Token Context and Multimodal Input
Google has released Gemini 3.7 Flash, a multimodal model built for agentic workflows, coding, and multi-step reasoning. It offers a 1,049K token context window and is priced at $0.38 per million input tokens and $1.88 per million output tokens, available now via OpenRouter.
Gemini's Market Share Reportedly Collapses to 1.9%, According to New Third-Party Data
Three independent data sources — Pangram, OpenRouter, and Similarweb — indicate Google's Gemini is losing market share to ChatGPT and Claude. Pangram reports Gemini's share fell from 12% to 1.9% between tracking periods, while Anthropic climbed from 4.3% to 14.9%.
Google Expands Gemini-Powered Ask Maps Globally With Personal Intelligence, Real-Time Transit, Agentic Ordering
Google Maps' Gemini-powered Ask Maps chat is rolling out globally to English speakers in Australia, Brazil, Canada, Indonesia, Japan, Mexico, and over 150 other countries and territories. The update adds Gmail-based Personal Intelligence, real-time transit data, conversation memory, and agentic capabilities for ordering food and booking hotels.
Google Assistant Shuts Down on Android and Wear OS Starting September 4, 2026
Google has set September 4, 2026 as the start date for removing Google Assistant access on Android phones, tablets, Wear OS watches, headphones, and Android Auto. Gemini becomes the sole assistant on these platforms, though Google Assistant remains active in cars with Google built-in.
Google's Gemini-Based AI Agents Found and Fixed 1,072 Chrome Security Bugs in Two Release Cycles
Google's Chrome security team used a Gemini-based agentic harness to find and fix 1,072 vulnerabilities across two release milestones, surpassing the combined total of the previous 23 releases. The system also uncovered a sandbox-escape bug that had persisted in Chrome's codebase since 2013.
Google Cancels Standalone AI Studio Mobile App, Shifts App-Building Into Gemini App Instead
Google has canceled the standalone AI Studio app for Android and iOS that it teased at I/O 2026, despite 800,000 pre-orders. Instead, app-building capabilities will be integrated directly into the Gemini app for mobile and desktop.
Oracle Adds Google's Gemini to Fusion Apps and NetSuite; Shares Jump 8.4%
Oracle is embedding Google's Gemini 3.1 Flash-Lite and Gemini 3.5 Flash models into its Fusion Applications and NetSuite software, expanding a partnership with its cloud rival. Oracle shares rose as much as 8.4% to $127.64 on the news.
Google's Gemini Spark Gains Chrome Auto-Browse Control, Expands to 160+ Countries
Google's Gemini Spark personal agent can now control desktop Chrome directly, using logged-in accounts and saved passwords to complete web tasks. The feature launches in the US first, alongside a Google AI Pro expansion bringing Spark to more than 160 additional countries.
Waymo Adds Gemini Voice Assistant and Redesigns In-Cabin Interface for Ojai Vehicles
Waymo has added Gemini to its Ojai vehicles, letting riders ask questions and adjust cabin controls like AC and pullover requests via voice. The assistant operates independently from the Waymo Driver and has no access to vehicle movement or routing.
Replit Launches 'Replit Design,' an AI Design Suite Powered by Claude, GPT-5, Gemini, Kimi, and GLM
Replit has launched Replit Design, a browser-based AI design suite that lets users generate apps, sites, and brand assets using models including Claude, GPT-5, Gemini, Kimi, and GLM. The product replaces Replit's earlier Canvas tool and integrates the Mobbin UI reference library directly into the workflow.
Gemini for macOS Rolls Out Voice Control With Screen-Aware Dictation and Editing
Google is rolling out advanced voice control for Gemini on macOS, version 1.88, letting users dictate cleaned-up text and issue voice commands that reference on-screen content. The feature, previewed at I/O 2026, requires opt-in Gemini reasoning to unlock summarization, rewriting, and image editing tasks.
Gemini for macOS Rolls Out Voice Control With Screen-Aware Task Execution
Google is rolling out advanced voice control to Gemini for macOS version 1.88, combining Gboard Rambler-style dictation with a screen-aware assistant that can summarize files, rewrite text, and generate images by voice. The feature, previewed at I/O 2026, activates via a long-press of the Fn key and requires Gemini reasoning to be enabled for the advanced capabilities.
Google Simplifies Gemini App's Thinking Level Picker, Adds Notification Controls
Google is simplifying the Gemini app's model picker by collapsing the two-stage 'Standard' vs 'Extended thinking' selector into a single toggle. The company is also rolling out new notification settings on Android and reorganizing the Gemini Spark task interface.
Google Expands Gemini Spark Agentic Assistant to All AI Pro and Ultra Subscribers
Google is expanding access to Gemini Spark, its agentic AI assistant built on Gemini 3.5, to all Google AI Pro subscribers in the US and Google AI Ultra subscribers globally. The rollout excludes free-tier users and, for Ultra, customers in the EEA, Switzerland, the UK, and Nigeria.
Google Expands Gemini Task Automation to 40+ Apps, Debuts Screen Reasoning at Galaxy Unpacked
Google announced Gemini task automation now covers over 40 apps, up from six at beta launch, alongside new screen-reasoning capabilities for the Gemini app. The updates arrived at Samsung's Galaxy Unpacked event alongside the Galaxy Z Fold 8, Fold 8 Ultra, and Flip 8, which also showcased new Android XR glasses designs from Gentle Monster and Warby Parker.
Google delays Gemini 3.5 Pro release after disappointing coding performance in June training update
Google has delayed the release of Gemini 3.5 Pro past its June deadline due to coding performance issues. The company retrained the model in late June with new data but saw disappointing results, according to Bloomberg. An upgraded Flash model is now in testing with partners.
Google prepares voice customization for Gemini with speed, energy, formality, and warmth controls
Google is preparing to let users customize Gemini's voice output across four parameters: speed, energy, formality, and warmth, according to code discovered in the Google app 17.41.12 beta. The controls will apply to both Gemini Live and standard chat interactions.
Google Rebrands NotebookLM to Gemini Notebook, Brings Gemini 3.5 and Antigravity to AI Pro
Google renamed NotebookLM to Gemini Notebook and announced that the Gemini 3.5 model with Antigravity code execution capability will roll out to AI Pro subscribers in the coming weeks. The research tool now has over 30 million users and 600,000+ organizations.
Google releases Magic Pointer app for unreleased Googlebook device to Play Store
Google has released Magic Pointer to the Play Store, an app designed for its yet-to-be-announced Googlebook device. The app allows users to select on-screen content to receive contextual AI suggestions powered by Gemini, including search, image creation, and shopping features.
Google redesigns Gemini visual responses on Nest Hub with Material 3 weather cards
Google has updated Gemini for Home on Nest Hub and Smart Displays with redesigned visual responses using Material 3 design language. The update includes refreshed weather forecast cards, more reliable sports information, and improved Continued Conversation that no longer requires mid-conversation voice verification.
Google Drive's Ask Gemini AI assistant launches on Android and iOS for AI Pro subscribers
Google is rolling out Ask Gemini and AI Overviews to Google Drive's Android and iOS apps. The features enable multi-turn conversations across Drive, Gmail, Chat, Calendar, and web search, available to AI Pro, Ultra, Business Standard/Plus, and Enterprise Standard/Plus subscribers in English plus 28 additional languages.
Google launches Nano Banana 2 Lite image model at 4 seconds per image, $0.04 per 1,000 generations
Google released Nano Banana 2 Lite, an image generation model that produces images in four seconds at under four cents per thousand images. The model prioritizes speed and cost over quality, targeting developers building high-volume image pipelines.
Google launches Gemini 3.1 Flash Lite Image with 4-second generation time, $0.25 per 1M input tokens
Google has released Gemini 3.1 Flash Lite Image, a text-to-image model that generates 1K resolution images in approximately 4 seconds — 2.7× faster than Gemini 3.1 Flash Image. The model is priced at $0.25 per 1M input tokens and $1.50 per 1M output tokens, with a 66K context window and knowledge cutoff of January 2025.
Google releases Nano Banana 2 Lite: 4-second image generation at $0.034 per 1,000 images
Google released Nano Banana 2 Lite, an AI image generator that produces images in 4 seconds and costs $0.034 per 1,000 images. The model is optimized for high-volume workflows and replaces the original Nano Banana as Google's entry-level image generation offering.
Gmail's Gemini Flows adds AI-powered email filtering with 2,000 message monthly limit on Pro tier
Google's Workspace Studio Flows is now available to Google AI Pro ($20/month) and Ultra ($100/month) subscribers, bringing AI-powered email filtering to Gmail. The service processes up to 2,000 emails monthly on Pro tier and 10,000 on Ultra, potentially limiting utility for high-volume users receiving thousands of messages weekly.
Google releases Gemini 3.1 Flash Image, claims Pro-level quality at $0.50 per 1M tokens
Google has released Gemini 3.1 Flash Image, internally codenamed "Nano Banana 2," an image generation and editing model with a 131K context window. The model is priced at $0.50 per 1M input tokens and $3 per 1M output tokens.
Gemini for Android adds persistent chat bubbles for multitasking across apps
Google is rolling out Android Bubbles support for the Gemini overlay, allowing users to minimize active conversations into a floating spark logo icon. The feature prevents chat loss when switching away from the overlay, similar to Gemini Live's floating waveform.
Epic Games ships Model Context Protocol plugin for Unreal Engine 5.8, plans gen AI integration for UE6
Epic Games released Unreal Engine 5.8 with an experimental Model Context Protocol plugin that allows developers to connect AI models including Claude and Gemini to the game engine. The company plans to make MCP integral to Unreal Engine 6, expected in late 2027.
Google Pixel Drop to add Screen Reactions, Gemini Omni music generation in upcoming Android 17 update
Google's upcoming Pixel Drop will bring Screen Reactions and Gemini Omni-powered features to Pixel phones, according to promotional videos discovered on Amazon. The update, overdue based on Google's typical quarterly schedule following March 2026's release, appears to coincide with Android 17's stable release.
Google adds Business Profile integration to Gemini app with automated review responses, performance analysis
Google is integrating Google Business Profile data into the Gemini app, giving business owners AI-powered review management, performance analytics, and profile updates. The integration includes Business notebooks for organizing business data and generating brand-matched content.
Google DeepMind Releases Gemini 3.5 Live Translate for Real-Time Speech Translation Across 70+ Languages
Google DeepMind released Gemini 3.5 Live Translate, an audio model that provides near real-time speech-to-speech translation across 70+ languages. The model automatically detects languages, preserves speaker intonation and pacing, and maintains a few seconds of latency while generating continuous speech output.
Apple deploys 1.2T-parameter Gemini model on Nvidia Blackwell GPUs for rebuilt Siri
Apple announced at WWDC 2026 that the rebuilt Siri runs on a custom 1.2-trillion-parameter model based on Google's Gemini technology, hosted on Google Cloud servers powered by Nvidia Blackwell B200 GPUs. The company unveiled a three-tier privacy architecture and five new Apple Foundation Models to handle queries across device, private cloud, and Google Cloud infrastructure.
Google cuts AI Plus subscription to $5/month, doubles storage to 400GB
Google lowered its AI Plus subscription from $8 to $5 per month and doubled included storage from 200GB to 400GB. The plan includes access to Gemini 3 Pro, Nano Banana Pro, Deep Research, and the newly announced Gemini Omni video generation model.