Google adds voice prompting to Docs, Keep, and Gmail via Gemini AI
Google unveiled voice-based prompting for Docs, Keep, and Gmail at I/O 2026, powered by Gemini AI. The features enable document creation, note organization, and email search through spoken commands, launching this summer for Google AI Premium subscribers and Workspace business users.
Google announced voice-based prompting features for Docs, Keep, and Gmail at its I/O 2026 developer conference on Monday, enabling users to create documents, organize notes, and search email through spoken commands instead of typing.
Docs Live enables voice-driven document creation
The headline feature, Docs Live, allows users to create and edit documents entirely by speaking. In a demonstration, Google showed a user verbally instructing the tool to pull résumé details from Drive, incorporate event logistics from an email thread, and add anecdotes in a single unscripted stream of speech.
According to Google, voice enables longer and more complex prompts than most users would type, while current models can follow along even when speakers change direction mid-sentence.
CEO Sundar Pichai claimed users will soon create and edit documents using voice as a matter of course. The company recently launched Rambler, a standalone dictation product built into its Gboard keyboard that removes filler words and handles multilingual code-switching. Rambler shipped earlier this month for Samsung Galaxy and Google Pixel devices.
Keep adds voice-to-structured-notes
Keep is gaining voice capabilities that allow users to dump unstructured thoughts—from gift ideas to grocery lists to home renovation plans—which the AI then sorts into separate, organized notes.
Similar functionality exists in apps like Voicenotes and AudioPen, and desktop dictation tools like Wispr Flow, Monologue, and Aqua Voice. Google's advantage is scale: Keep integrates with the broader Workspace ecosystem, allowing voice notes to flow directly into Docs, Sheets, and other Workspace tools.
Gmail Live provides conversational inbox search
Gmail is gaining Gmail Live, a conversational voice interface for email. Users can ask Gmail to surface specific details—flight confirmation codes, Airbnb check-in instructions, or school schedules—and receive spoken answers drawn from messages. The system handles multi-step requests and understands context.
Broader AI voice integration trend
Google's Cloud Next conference last month showcased agentic AI features across Workspace. Competitors including OpenAI and Apple are embedding voice-first AI into their productivity tools.
The new voice features will roll out this summer for Google AI Premium subscribers and Google Workspace business users. Pricing for the features has not been disclosed separately from existing subscription tiers.
What this means
Google is positioning voice as the primary interface for complex, multi-step AI interactions in productivity software, leveraging Gemini's capabilities to handle unstructured spoken input. The integration across Workspace gives Google distribution advantages over standalone voice note apps, though adoption will depend on whether users prefer speaking to their documents over typing—a behavioral shift that remains unproven at scale. The move signals that major tech companies view voice as critical to the next generation of AI-powered productivity tools.
Related Articles
Google Workspace Grants Gemini Default Access to Gmail, Docs, Calendar, and Chat Data
Google Workspace ships with Gemini's access to Gmail, Docs, Calendar, Chat, Drive, and Meet turned on by default, using real-time retrieval-augmented generation rather than stored training data. Admins can disable these 'Workspace Intelligence Sources' organization-wide through the admin.google.com console, though per-user controls remain limited.
Google to Let Users Remove Visible AI Watermarks From Nano Banana, Omni, Lyria Content
Google VP Josh Woodward announced a new toggle that lets users remove visible sparkle-icon watermarks from AI-generated content made with Nano Banana, Omni, and Lyria. The invisible SynthID watermark and C2PA metadata will remain unaffected, and the toggle won't roll out in the EU or South Korea where visible labeling is legally required.
Google Lets Users Turn Off Visible Watermarks on Nano Banana, Omni, and Lyria Outputs
Google announced users can now toggle off visible watermarks on AI-generated images, video, and songs from its Nano Banana, Omni, and Lyria models. Invisible SynthID watermarks and C2PA metadata remain in place for transparency.
GitHub Adds Enterprise Managed Settings to Copilot for JetBrains
GitHub Copilot for JetBrains now supports enterprise managed settings, letting administrators enforce consistent policies for plugin governance, MCP server access, OpenTelemetry, and permission modes across their organization.
Comments
Loading...