Audio modelActive
Gemini 3.5 Transcribe
- Released
- -
- Data date
- September 22, 2026
- Model class
- Access
Profile
Specifications and access
Published information about this model. Existing estimates are explicitly labeled.
| Specification | Value and source |
|---|---|
| Model class | Speech recognitionSource |
| Access | APISource |
| API model ID | gemini-3.5-transcribeSource |
| Input | AudioSource |
| Output | TextSource |
| Task | TranscriptionSource |
| Supported details | Language detection, diarization, and timestampsSource |
| Languages | Over 85; code-switching within an audio fileSource |
| Audio duration | Up to 1 hour; up to 30 minutes with diarization or word timestampsSource |
| Speaker diarization | Up to 8 speakers; experimental for 3 or moreSource |
Evidence
Sources and data date
Every statement links to its underlying documentation or leaderboard.
| Type | Evidence and data date |
|---|---|
| Additional source | Google Gemini API model catalog (retrieved September 22, 2026) |