Today we’re introducing Gemini 3.5 Transcribe, our latest transcription model built for incredibly precise, smart dictation across your favorite apps and devices. Remember when traditional speech-to-text meant shouting over background noise, constantly hitting backspace to fix misspelled words, and manually deleting every "um" and "uh"? Those days are over. Gemini 3.5 Transcribe isn't just dictation — it’s active intelligence with precise, context-aware speech-to-text support in 85+ languages. The model automatically filters out filler words, formats unstructured speech, and even pairs with your screen context to execute voice commands. Watch as Gemini 3.5 Transcribe removes filler words and uses multimodal capabilities to seamlessly turn messy voice input and local files into a polished email draft.
— Try it out in the @Geminiapp on macOS and Gboard on @Android — Build in the Gemini API via @googleaistudio, @antigravity, and the Gemini Enterprise Agent Platform (public preview) — Coming soon to @googlechrome and Gemini Enterprise for Customer Experience Learn more: https://t.co/mGPWSdloR8
Transcription that formats and cleans as it goes is directly useful for prompting coding agents by voice, and it is reachable through the Gemini API rather than only the consumer app.
Checking sign-in…
Loading comments…