Demis Hassabis转发:Gemini 3.5 Transcribe发布,支持多说话人识别
demishassabis · x · 2026-08-27
Demis Hassabis转发了Sundar Pichai关于Gemini 3.5 Transcribe的推文,强调其支持多说话人识别、85+语言自动检测和自定义词汇适配。API现已在Google AI Studio和Gemini Enterprise提供。
所属事件:谷歌发布 Gemini 3.5 Transcribe 语音转文本模型(17 条相关)→
「多模态」频道最新
- xAI 发 Grok Imagine 电影感创作指南,Odyssey 视频赛 8 月底截止 — chaitu · 2026-08-27
- HeyGen 为何开源 HyperFrames:将视频编辑转为代码问题 — altryne · 2026-08-27
- Seedance 2.5 同步生成音画,原生低分输出来降成本 — LudovicCreator · 2026-08-27
- Seedance 2.5 支持 50 路参考素材,解决角色一致性难题 — LudovicCreator · 2026-08-27
- 构建多模态智能体:从 LLM 到落地 Agent 的 7 步路线图 — MaryamMiradi · 2026-08-27
- Descript 揭秘专用模型:零shot语音修复与唇形同步 — descript · 2026-08-27