Calculate the cost of using gemini-3.5-transcribe from Google Gemini for your AI applications
Pricing data last updated:
Mode: Audio Transcription
Built for enterprise-grade reliability, governance, and scale. Deploy in seconds.
gemini-3.5-transcribe is an audio transcription model from Google Gemini, one of 2 audio transcription models they offer. It is priced at $2.00 per 1M input tokens and $12.00 per 1M output tokens, ranking 5 out of 12 audio transcription models by cost and cheaper than 58% of models in this category. gemini-3.5-transcribe supports audio input.
Note: Use the interactive calculator above to estimate costs for your specific usage patterns.
At $2.00 per 1M input tokens and $12.00 per 1M output tokens, gemini-3.5-transcribe ranks 5 out of 12 audio transcription models by input cost. It is more affordable compared to the median of $2.50 for audio transcription models, and is cheaper than 58% of models in this category.
gemini-3.5-transcribe is one of 2 Google Gemini audio transcription models, with support for audio input.
| Model | Provider | Input / 1M tokens | Output / 1M tokens | vs gemini-3.5-transcribe |
|---|---|---|---|---|
| gemini-3.5-transcribe-preview | Vertex AI | $2.50 | $12.00 | +25% |
| gemini-3.5-transcribe-live-preview | Vertex AI | $3.50 | $21.00 | +75% |
| gpt-4o-mini-transcribe | Azure | $1.25 | $5.00 | -37% |
Similar audio transcription models from other providers
Yes. gemini-3.5-transcribe costs $2.00 per 1M input tokens compared to gemini-3.5-transcribe-preview's $2.50 per 1M input tokens, making it 25% more affordable. Both are audio transcription models and share support for audio input.
gemini-3.5-transcribe input pricing is $2.00 per 1M tokens, which is 20% below the median of $2.50 for audio transcription models. It ranks 5 out of 12 audio transcription models by input cost, making it cheaper than 58% of models in this category. For output, it costs $12.00 per 1M tokens compared to the median of $10.00.
Among Google Gemini's 2 audio transcription models, gemini-3.5-transcribe ranks 1 by input cost. It is the most affordable Google Gemini model with audio input support.
The most comparable audio transcription models to gemini-3.5-transcribe are: gemini-3.5-transcribe-preview from Vertex AI ($2.50/1M input tokens); gemini-3.5-transcribe-live-preview from Vertex AI ($3.50/1M input tokens); gpt-4o-mini-transcribe from Azure ($1.25/1M input tokens); gpt-4o-transcribe from Azure ($2.50/1M input tokens). These alternatives were selected based on similar capabilities, pricing, and provider diversity. You can compare any of these models in detail using the Bifrost Model Library.
Yes. At $2.00 per 1M input tokens, gemini-3.5-transcribe is the most affordable audio transcription model that supports audio input. This makes it a strong option for audio input-dependent workloads where cost efficiency matters.
gemini-3.5-transcribe is priced based on input and output tokens. Use the interactive calculator at the top of this page to estimate costs for your specific workload. Enter your expected input and output tokens volume and the calculator will show the total cost breakdown. For reference, processing 1M input tokens costs $2.00 and generating 1M output tokens costs $12.00.