AI Frontier
← Browse this publisher

Google / DeepMind / Model Card

Gemini 3.5 Audio (Live Translate, Transcribe, Transcribe Live) Model Card

Gemini 3.5 Transcribe / Transcribe Live · Date unconfirmed

Source summary

Original wording · Original language

Description · Page 2

Gemini 3.5 Audio (Live Translate, Transcribe, Transcribe Live) is an addition to the Gemini 3 series of highly-capable, natively multimodal, reasoning models. The models are cost-efficient and fast, optimized for high-volume, latency-sensitive tasks like dialogue and translation. This model card describes the native audio capabilities as additional outputs of Gemini. Information specific to these modalities is specified in-line and referred to as Gemini 3.5 Live Translate, Gemini 3.5 Transcribe, and Gemini 3.5 Transcribe Live referred to collectively as Gemini 3.5 Audio.

Core figures

Enlarge to explore. Download the original for full detail.

No core figure selected for this report. The original PDF remains available.

Click the image to zoom. Press Esc to close. Full-resolution files are available below each figure.