Google launches Gemini 3.5 Transcribe, a real-time speech-to-text model in the macOS app
GeminiApp · x · 2026-09-09
Google announced Gemini 3.5 Transcribe, its latest speech-to-text model built for precise, intelligent real-time transcription.
- It ships in the Gemini app for macOS, combining voice input with on-screen context
- Users can summarize local files, plan, research, and even create images via free-form natural language
- Positioned as transcription plus understanding, not just dictation
Related event: Google Launches Gemini 3.5 Transcribe on macOS with Screen Context(2 posts)→
More from Multimodal
- First samples shared from ChatGPT Images 2.5, OpenAI's newly released image model — DeryaTR_ · 2026-09-09
- New Sketch feature in GPT Image 2.5 praised for outsized productivity gains — lukaszkaiser · 2026-09-09
- ComfyUI adds Sol Attention in latest update, demoed in video — DemolitionMan32 · 2026-09-09
- Redditor crafts low-res Star Trek AI mini-episode with ComfyUI and MiniMax H3 — Perfect-Campaign9551 · 2026-09-09
- Image 2.5 showing off its text rendering proficiency — NickPassig · 2026-09-09
- GPT-Image-2.5 appears in ChatGPT; classic noise artifact test shows mixed results — mark_k · 2026-09-09