Gemini 3.5 Transcribe Released: Why Function Calling in a Speech-to-Text Model?
giffmana · x · 2026-08-27
A user questions the inclusion of "Function Calling" in the new Gemini 3.5 Transcribe model. The model features smart transcription, lower WER, custom vocabulary, multi-speaker ID, and 85+ language support with real-time streaming. Function Calling likely allows triggering custom logic within the transcription workflow.
Related event: Google Launches Gemini 3.5 Transcribe Speech-to-Text Model(17 posts)→
More from Models
- Yutori's Batra: most of the web will never get agent APIs — pixels in, clicks out — DhruvBatra_ · 2026-08-27
- OpenAI report: tens of thousands of ExploitGym agents discovered each other via Artifactory — ChrisGPT · 2026-08-27
- From Torrent Links to API Gates: The Evolution of Open Weights Distribution — thursdai_pod · 2026-08-27
- Test: Codex Usage Seems Almost Unlimited Compared to Fable — panickssery · 2026-08-27
- Qwen and GLM push forward with low-cost, high-speed models — brandon_galang · 2026-08-27
- User Corrects Misinformation Regarding Qwen3.8-27B Model Rankings — Real-C- · 2026-08-27