Google / DeepMind / Model Card
Gemini 3.5 Audio (Live Translate, Transcribe, Transcribe Live) Model Card
Source summary
Original wording · Original languageDescription · Page 2
Gemini 3.5 Audio (Live Translate, Transcribe, Transcribe Live) is an addition to the Gemini 3 series of highly-capable, natively multimodal, reasoning models. The models are cost-efficient and fast, optimized for high-volume, latency-sensitive tasks like dialogue and translation. This model card describes the native audio capabilities as additional outputs of Gemini. Information specific to these modalities is specified in-line and referred to as Gemini 3.5 Live Translate, Gemini 3.5 Transcribe, and Gemini 3.5 Transcribe Live referred to collectively as Gemini 3.5 Audio.
Core figures
Enlarge to explore. Download the original for full detail.
No core figure selected for this report. The original PDF remains available.