Skip to content
GoogleVoice

Gemini 3.5 Transcribe

gemini-3.5-transcribe

Google Gemini 3.5 Transcribe speech-to-text. Billed per input and output token.

Context
98.3K tokens
Endpoint

Get API KeyCompare

Pricing

Input$0 / 1M
Output$0 / 1M
Cache Write (5m)Not applicable
Cache Write (1h)Not applicable
Cache ReadNot applicable
Web Search$0 / 1M

Customer audio pricing

USD

Base rates shown at GroupRatio=1. Your account's group multiplier may change the final charge.

/v1/audio/transcriptionsSpeech to text
RateUSD per unit
InputAudio$0.000002 USD per token
OutputText$0.000012 USD per token

Cached input tokens are charged at the full configured customer input rate.

Actual settled usage is deducted from your Apertis credit quota.

Quick Start

Select an endpoint and copy a working example for this model.

Endpoint
python
from openai import OpenAI client = OpenAI(    api_key="YOUR_API_KEY",    base_url="https://api.apertis.ai/v1") response = client.chat.completions.create(    model="gemini-3.5-transcribe",    messages=[        {"role": "user", "content": "Hello!"}    ],    max_tokens=1024,    temperature=0.7) print(response.choices[0].message.content) # Optional: Enable context compression to reduce token usage# response = client.chat.completions.create(#     model="gemini-3.5-transcribe",#     messages=[{"role": "user", "content": "Hello!"}],#     extra_body={"compression": {"enabled": True, "model": "gpt-4.1-mini"}}# )

Supported Parameters

API docs
Common5 params
modelfilelanguagepromptresponse_format
Extended2 params
temperaturetimestamp_granularities

Cursor IDE Model IDs

Use these namespaced identifiers in Cursor IDE to avoid conflicts with built-in models.

gemini-3.5-transcribe

Compare with Other Models

See how this model compares to others from the same provider.