# Google DeepMind 发布 Gemini 3.5 Live Translate 实时语音翻译模型

> 原标题：Fluid, natural voice translation with Gemini 3.5 Live Translate

- 来源：Google DeepMind
- 发布时间：2026-06-09T15:16:25.000Z
- AX AI 日报：https://ai-daily.ax0x.ai/items/uvql0w3s3gkqk1o1xkgonyrdp
- 原文：https://deepmind.google/blog/fluid-natural-voice-translation-with-gemini-35-live-translate/

## 摘要

Google DeepMind 发布 Gemini 3.5 Live Translate，一款支持 70 多种语言的近实时语音到语音翻译音频模型，可自动检测语言并保留说话人的语调、节奏和音高。

## 推荐理由

原文给出模型的语言覆盖、实时翻译机制与多产品上线节奏，可据此判断语音翻译的可用边界。

## 锐评

Google 这次把实时语音翻译从"说完再翻"改成"边说边翻"，Gemini 3.5 Live Translate 只落后说话人几秒，覆盖 70 多种语言，还能保留语调、节奏和音高。对比 Google Meet 之前只支持 5 种语言、只能中英互转，这次直接跳到 2000 多种语言组合，跨度不小。

但正文没给延迟的具体秒数和翻译准确率，"几秒"这个说法太模糊。Grab 那边每月 1000 万通电话是真实场景，但只是"测试中"，没有量化结果。我会先打个折：等开发者用 Gemini Live API 跑出实际延迟数据再说。
