GE
Gemini 3.5 Transcribe
Release date: 2026-08-26Updated: 2026-09-18Views: 180
Parameters
No data
Context length
96K
Multilingual
85+ languages
Reasoning ability
1/5
Gemini 3.5 Transcribe is an AI model published by Google Deep Mind, released on 2026-08-26, for Audio model, and 96K context length.
Data sourced primarily from official releases (GitHub, Hugging Face, papers), then benchmark leaderboards, then third-party evaluators. Learn about our data methodology
Gemini 3.5 Transcribe
Model basics
Reasoning traces
No data
Thinking modes
Thinking modes not supported
Context length
96K tokens
Max output length
32K tokens
Model type
Audio model
Modality (in / out)
Text, Audio → Text
Supported languages
Afrikaans (South Africa), Amharic (Ethiopia), Arabic (Egypt), Armenian (Armenia), Assamese (India), Azerbaijani (Azerbaijan), Belarusian (Belarus), Bangla (Bangladesh) (85+ languages)
Release date
2026-08-26
Model file size
No data
MoE architecture
No
Total params / Active params
No data / Not applicable
Knowledge cutoff
No data
Gemini 3.5 Transcribe
Open source & experience
Code license
Proprietary
Weights license
Proprietary
GitHub repo
N/A
Hugging Face
N/A
Live demo
Gemini 3.5 Transcribe
Official resources
Paper
DataLearnerAI blog
N/A
Gemini 3.5 Transcribe
API details
API speed
5/5
💡Default unit: $/1M tokens. If vendors use other units, follow their published pricing.
Standard
| Type | Condition | Input | Output |
|---|---|---|---|
| Text | - | — | $12.00/ 1M tokens |
| Audio | - | $2.00/ 1M tokens | — |
“—” means the modality is not billed in that direction, or the vendor has not published a price for it.
Gemini 3.5 Transcribe
Publisher
Google Deep Mind
View publisher details Gemini 3.5 Transcribe
Model Overview
Gemini 3.5 Transcribe is an AI model published by Google Deep Mind, released on 2026-08-26, for Audio model, and 96K context length.
DataLearner on WeChat
Follow DataLearner on WeChat for AI model updates and research notes.
