跳到正文
原文
Google DeepMind·· 2026-08-27精选AI 评分75

Google DeepMind 发布 Gemini 3.5 Transcribe 语音转文字模型

Intelligent transcription with Gemini 3.5 Transcribe

AI 导读

Google DeepMind 推出 Gemini 3.5 Transcribe,这是其最新的语音转文字模型,支持实时流式与预录音频处理两种 API。据 Artificial Analysis 测量,该模型流式词错误率为 4.0%,非流式为 2.6%,相比上一代 Chirp 3 最终转录时间提升 70%,并支持超过 85 种语言及最多三人的说话人识别。

推荐理由

原文给出了新模型的词错误率、延迟改进和开放入口,读者可以据此判断它相比上一代产品的实际提升。

来源:Google DeepMind · deepmind.google