Google / DeepMind / Model Card
Gemini 3.5 Audio (Live Translate, Transcribe, Transcribe Live) Model Card
概要原文
保留原文 · 保留原始语言Description · 页码 2
Gemini 3.5 Audio (Live Translate, Transcribe, Transcribe Live) is an addition to the Gemini 3 series of highly-capable, natively multimodal, reasoning models. The models are cost-efficient and fast, optimized for high-volume, latency-sensitive tasks like dialogue and translation. This model card describes the native audio capabilities as additional outputs of Gemini. Information specific to these modalities is specified in-line and referred to as Gemini 3.5 Live Translate, Gemini 3.5 Transcribe, and Gemini 3.5 Transcribe Live referred to collectively as Gemini 3.5 Audio.
核心图片
点击放大查看,下载原图获取完整细节。
此报告暂无选取的核心配图,可直接阅读原始 PDF。