Google Launches Gemini 3.5 Transcribe, a Speech-to-Text Model That Cleans Up Rambling Speech
Google has released Gemini 3.5 Transcribe, a new speech-to-text model that automatically detects over 85 languages, removes filler words, and structures unstructured speech into clean text. The model powers Android's Rambler feature and is rolling out to Chrome, Docs, Gmail, and other Google products.
Google has released Gemini 3.5 Transcribe, a new speech-to-text model the company says can turn unstructured, rambling speech into clean, formatted text. The model joins Gemini 3.5 Live and Gemini 3.5 Live Experimental to round out Google's Gemini Audio family.
According to Google, Gemini 3.5 Transcribe improves on earlier transcription models with greater accuracy and automatic detection of more than 85 languages. The model is designed to handle natural speech patterns rather than requiring users to dictate cleanly — Google claims it can seamlessly process self-corrections, understand a speaker's natural intent, and strip out filler words to produce polished output.
The model also supports voice-based editing commands, letting users revise text they've already dictated using spoken instructions rather than typing. Google says Gemini 3.5 Transcribe can learn custom vocabulary and unique spellings, and is particularly capable at capturing alphanumeric strings like order numbers and postal codes — a detail that points toward enterprise and customer-service use cases.
For multi-speaker audio, Google claims the model can attribute speech to up to three speakers with word-level timestamps when working from pre-recorded audio, a feature suited to transcribing podcasts, meetings, or interviews. No benchmark scores, context window size, or pricing have been disclosed for Gemini 3.5 Transcribe.
The model is already live in two places: it powers the Rambler dictation feature on Android devices, including the Pixel 11 series, and it runs the Gemini app on macOS, where Google says it can work alongside other Gemini models to complete agentic tasks.
Google says broader distribution is coming soon. Chrome users will reportedly be able to use Gemini 3.5 Transcribe for speech-to-text in any web field — dictating replies, social posts, or prompts to Gemini directly from a browser. The model is also available now in Google Antigravity, the company's agentic development platform, and Google says it is coming to Search Live, Gemini Live, Docs, Keep, and Gmail. Developers will get API access to build the model into their own products.
What this means
Gemini 3.5 Transcribe is less about raw transcription accuracy — a race Google, OpenAI, and others have run for years — and more about post-processing: turning messy spoken input into text a person could actually publish or send without editing. That positions it as an assistant-layer feature rather than a standalone product, embedded directly into Android, macOS, Chrome, and Google's productivity suite rather than sold as a separate API-first offering.
The alphanumeric-capture and custom-vocabulary claims suggest Google is targeting business and support workflows, not just casual dictation. But with no benchmark scores or pricing disclosed, independent verification of the model's claimed language coverage and speaker-attribution accuracy will have to wait until developers get broader API access.
Related Articles
Google Lists Gemini 3.8 Flash on OpenRouter With 1M-Token Context, September 2026 Release Date
Google's Gemini 3.8 Flash has surfaced on OpenRouter with a 1-million-token context window and discounted pricing of $0.75 per 1M input tokens and $3.75 per 1M output tokens. Google has not issued a separate public announcement, and the listed release date of September 2, 2026 is unusually far out, leaving key details unconfirmed.
Google Launches WeatherNext 3, Claims 50% More Accurate Precipitation Forecasts
Google DeepMind and Google Research released WeatherNext 3, a weather AI model trained on real-time geostationary satellite data instead of lagging numerical weather prediction outputs. Google claims up to 50% more accurate day-ahead precipitation forecasts, now rolling out to Search, Maps, and the Gemini app.
Google Launches Gemini 3.8 Flash With Same Pricing as Predecessor, Its Third Flash Model in Six Weeks
Google released Gemini 3.8 Flash, its third Flash model in six weeks, pricing it identically to its predecessor at $0.75 per million input tokens and $3.75 per million output tokens. The launch coincides with a favorable antitrust ruling and public praise from Berkshire Hathaway's Greg Abel, giving Alphabet a stronger narrative after a four-month stock slide.
OpenAI's GPT-6 Astra Cuts Hallucinations, But Indirect Prompt Injection Attacks Still Succeed 8.5% of the Time
OpenAI's new GPT-6 Astra model shows major improvements in hallucination rates and jailbreak resistance over predecessor GPT-5.6 Sol, according to OpenAI's system card. However, indirect prompt injection attacks hidden in documents still succeed 8.5% of the time in external testing by Gray Swan, down from 27% but still above rival Claude Opus 5's 4.8% rate.
Comments
Loading...