跳到正文
原文
Google DeepMind·· 2026-08-27

Google DeepMind 发布 Gemini 3.5 Transcribe 语音转文本模型

Intelligent transcription with Gemini 3.5 Transcribe

AI 导读

Google DeepMind 发布 Gemini 3.5 Transcribe,称其为目前最精确的语音转文本模型,可直接把原始音频转成准确、经过格式化的文本。

入选记录

入选两次之和 145 ≥ 门槛的两倍 120

信源等级
官方一手,这一级的门槛是 60 分
预筛
通过:Google发布Gemini语音转文字模型
推荐理由
官方给出 WER、延迟与多语言等量化指标,并说明两套 API 的接入方式,便于开发者评估语音链路选型。

评分由模型按同一份标准独立打两次,满分 100;两次之和达到门槛的两倍才入选。门槛按信源等级设定。

来源:Google DeepMind · deepmind.google