---
title: "Google Vertex AI AI Models - Model Library"
description: "Browse all 157 Google Vertex AI AI models with capabilities, context limits, supported modes, and pricing."
url: "https://maxim-root-proxy-preview.getmaxim.workers.dev/bifrost/model-library/provider/vertex_ai"
markdown: "https://maxim-root-proxy-preview.getmaxim.workers.dev/bifrost/model-library/provider/vertex_ai.md"
---

# Google Vertex AI AI Models - Model Library

> Browse all 157 Google Vertex AI AI models with capabilities, context limits, supported modes, and pricing.

## Important Links

- [View MCP Gateway](https://maxim-root-proxy-preview.getmaxim.workers.dev/bifrost/resources/mcp-gateway.md)
- [Features](https://maxim-root-proxy-preview.getmaxim.workers.dev/bifrost/#features)
- [Enterprise](https://maxim-root-proxy-preview.getmaxim.workers.dev/bifrost/enterprise)
- [Pricing](https://maxim-root-proxy-preview.getmaxim.workers.dev/bifrost/pricing.md)
- [Docs](https://docs.getbifrost.ai)
- [GitHub](https://github.com/maximhq/bifrost)
- [Book a Demo](https://maxim-root-proxy-preview.getmaxim.workers.dev/bifrost/book-a-demo)

## Google Vertex AI Overview

- **157 Total Models.**
- **8 Modes.**
- **$2.09 Avg Input (1M Tokens).**
- **$9.71 Avg Output (1M Tokens).**

## All Google Vertex AI Models

Showing the first 100 of 157 models.

| Model | Mode | Max Input Tokens | Max Output Tokens | Input $/1M |
| --- | --- | --- | --- | --- |
| gemini-2.0-flash | Chat | 1,048,576 | 8,192 | $0.10 |
| gemini-2.0-flash-001 | Chat | 1,048,576 | 8,192 | $0.15 |
| gemini-2.0-flash-lite | Chat | 1,048,576 | 8,192 | $0.07 |
| gemini-2.0-flash-lite-001 | Chat | 1,048,576 | 8,192 | $0.07 |
| gemini-2.5-flash | Chat | 1,048,576 | 65,535 | $0.30 |
| gemini-2.5-flash-lite | Chat | 1,048,576 | 65,535 | $0.10 |
| gemini-2.5-flash-lite-preview-09-2025 | Chat | 1,048,576 | 65,535 | $0.10 |
| gemini-2.5-flash-preview-09-2025 | Chat | 1,048,576 | 65,535 | $0.30 |
| gemini-live-2.5-flash-preview-native-audio-09-2025 | Realtime | 1,048,576 | 65,535 | $0.30 |
| gemini-2.5-flash-lite-preview-06-17 | Chat | 1,048,576 | 65,535 | $0.10 |
| gemini-2.5-pro | Chat | 1,048,576 | 65,535 | $1.25 |
| gemini-3-pro-preview | Chat | 1,048,576 | 65,535 | $2.00 |
| gemini-3-flash-preview | Chat | 1,048,576 | 65,535 | $0.50 |
| gemini-3.5-flash | Chat | 1,048,576 | 65,536 | $1.50 |
| gemini-3.6-flash | Chat | 1,048,576 | 65,536 | $1.50 |
| gemini-3.1-pro-preview | Chat | 1,048,576 | 65,536 | $2.00 |
| gemini-3.1-pro-preview-customtools | Chat | 1,048,576 | 65,536 | $2.00 |
| gemini-2.5-pro-preview-tts | Chat | 1,048,576 | 65,535 | $1.25 |
| gemini-robotics-er-1.5-preview | Chat | 1,048,576 | 65,535 | $0.30 |
| gemini-2.5-computer-use-preview-10-2025 | Chat | 128,000 | 64,000 | $1.25 |
| gemini-embedding-001 | Embedding | 2,048 | - | $0.15 |
| gemini-embedding-2-preview | Embedding | 8,192 | - | $0.20 |
| gemini-embedding-2 | Embedding | 8,192 | - | $0.20 |
| gemini-omni-flash-preview | Chat | 1,048,576 | 65,535 | $1.50 |
| medlm-large | Chat | 8,192 | 1,024 | - |
| medlm-medium | Chat | 32,768 | 8,192 | - |
| multimodalembedding | Embedding | 2,048 | - | $0.80 |
| multimodalembedding@001 | Embedding | 2,048 | - | $0.80 |
| text-embedding-004 | Embedding | 2,048 | - | $0.10 |
| text-embedding-005 | Embedding | 2,048 | - | $0.10 |
| text-embedding-large-exp-03-07 | Embedding | 8,192 | - | $0.10 |
| text-embedding-preview-0409 | Embedding | 3,072 | - | $0.0062 |
| text-multilingual-embedding-002 | Embedding | 2,048 | - | $0.10 |
| text-unicorn | Completion | 8,192 | 1,024 | $10.00 |
| text-unicorn@001 | Completion | 8,192 | 1,024 | $10.00 |
| chirp_3 | Audio Transcription | - | - | - |
| claude-3-5-haiku | Chat | 200,000 | 8,192 | $1.00 |
| claude-3-5-haiku@20241022 | Chat | 200,000 | 8,192 | $0.80 |
| claude-haiku-4-5 | Chat | 200,000 | 8,192 | $1.00 |
| claude-haiku-4-5@20251001 | Chat | 200,000 | 8,192 | $1.00 |
| claude-3-5-sonnet | Chat | 200,000 | 8,192 | $3.00 |
| claude-3-5-sonnet@20240620 | Chat | 200,000 | 8,192 | $3.00 |
| claude-3-7-sonnet@20250219 | Chat | 200,000 | 8,192 | $3.00 |
| claude-3-haiku | Chat | 200,000 | 4,096 | $0.25 |
| claude-3-haiku@20240307 | Chat | 200,000 | 4,096 | $0.25 |
| claude-3-opus | Chat | 200,000 | 4,096 | $15.00 |
| claude-3-opus@20240229 | Chat | 200,000 | 4,096 | $15.00 |
| claude-3-sonnet | Chat | 200,000 | 4,096 | $3.00 |
| claude-3-sonnet@20240229 | Chat | 200,000 | 4,096 | $3.00 |
| claude-opus-4 | Chat | 200,000 | 32,000 | $15.00 |
| claude-opus-4-1 | Chat | 200,000 | 32,000 | $15.00 |
| claude-opus-4-1@20250805 | Chat | 200,000 | 32,000 | $15.00 |
| claude-opus-4-5 | Chat | 200,000 | 64,000 | $5.00 |
| claude-opus-4-5@20251101 | Chat | 200,000 | 64,000 | $5.00 |
| claude-opus-4-6 | Chat | 1,000,000 | 128,000 | $5.00 |
| claude-opus-4-6@default | Chat | 1,000,000 | 128,000 | $5.00 |
| claude-opus-4-7 | Chat | 1,000,000 | 128,000 | $5.00 |
| claude-opus-4-7@default | Chat | 1,000,000 | 128,000 | $5.00 |
| claude-fable-5 | Chat | 1,000,000 | 128,000 | $10.00 |
| claude-fable-5@default | Chat | 1,000,000 | 128,000 | $10.00 |
| claude-opus-5 | Chat | 1,000,000 | 128,000 | $5.00 |
| claude-opus-5@default | Chat | 1,000,000 | 128,000 | $5.00 |
| claude-opus-4-8 | Chat | 1,000,000 | 128,000 | $5.00 |
| claude-opus-4-8@default | Chat | 1,000,000 | 128,000 | $5.00 |
| claude-sonnet-4-5 | Chat | 200,000 | 64,000 | $3.00 |
| claude-sonnet-5 | Chat | 1,000,000 | 128,000 | $2.00 |
| claude-sonnet-4-6 | Chat | 1,000,000 | 64,000 | $3.00 |
| claude-sonnet-4-5@20250929 | Chat | 200,000 | 64,000 | $3.00 |
| claude-opus-4@20250514 | Chat | 200,000 | 32,000 | $15.00 |
| claude-sonnet-4 | Chat | 1,000,000 | 64,000 | $3.00 |
| claude-sonnet-4@20250514 | Chat | 1,000,000 | 64,000 | $3.00 |
| codestral-2@001 | Chat | 128,000 | 128,000 | $0.30 |
| codestral-2 | Chat | 128,000 | 128,000 | $0.30 |
| codestral-2@001 | Chat | 128,000 | 128,000 | $0.30 |
| codestral-2 | Chat | 128,000 | 128,000 | $0.30 |
| codestral-2501 | Chat | 128,000 | 128,000 | $0.20 |
| codestral@2405 | Chat | 128,000 | 128,000 | $0.20 |
| codestral@latest | Chat | 128,000 | 128,000 | $0.20 |
| deepseek-v3.1-maas | Chat | 163,840 | 32,768 | $0.60 |
| deepseek-v3.2-maas | Chat | 163,840 | 32,768 | $0.56 |
| deepseek-r1-0528-maas | Chat | 65,336 | 8,192 | $1.35 |
| gemini-2.5-flash-image | Image Generation | 32,768 | 32,768 | $0.30 |
| gemini-3-pro-image | Image Generation | 65,536 | 32,768 | $2.00 |
| gemini-3-pro-image-preview | Image Generation | 65,536 | 32,768 | $2.00 |
| gemini-3.1-flash-image | Image Generation | 65,536 | 32,768 | $0.50 |
| gemini-3.1-flash-image-preview | Image Generation | 65,536 | 32,768 | $0.50 |
| gemini-3.1-flash-lite-preview | Chat | 1,048,576 | 65,536 | $0.25 |
| gemini-3.1-flash-lite | Chat | 1,048,576 | 65,536 | $0.25 |
| gemini-3.5-flash-lite | Chat | 1,048,576 | 65,536 | $0.30 |
| deep-research-pro-preview-12-2025 | Image Generation | 65,536 | 32,768 | $2.00 |
| imagegeneration@006 | Image Generation | - | - | - |
| imagen-3.0-fast-generate-001 | Image Generation | - | - | - |
| imagen-3.0-generate-001 | Image Generation | - | - | - |
| imagen-3.0-generate-002 | Image Generation | - | - | - |
| imagen-3.0-capability-001 | Image Generation | - | - | - |
| imagen-4.0-fast-generate-001 | Image Generation | - | - | - |
| imagen-4.0-generate-001 | Image Generation | - | - | - |
| imagen-4.0-ultra-generate-001 | Image Generation | - | - | - |
| jamba-1.5 | Chat | 256,000 | 256,000 | $0.20 |
| jamba-1.5-large | Chat | 256,000 | 256,000 | $2.00 |

## Related Tools

Check Google Vertex AI reliability and pricing.

- [All Google Vertex AI pricing](https://maxim-root-proxy-preview.getmaxim.workers.dev/bifrost/llm-cost-calculator/provider/vertex_ai.md)

## FAQ

### What is the AI model library?

The model library is a browsable catalog of AI models across every major provider, showing capabilities, context limits, supported modes, and pricing in one place. Use it to discover and compare models for chat, image generation, audio, embeddings, and more.

### How do I compare two models?

Open any model to see its full specification, then add a second model to view a side-by-side comparison of context windows, capabilities, and pricing. This makes it easy to pick the right model for a given workload.

### What details does each model page show?

Each model page lists the mode, maximum input and output tokens, supported capabilities (function calling, vision, reasoning, audio, and more), and input/output pricing per 1M tokens, alongside comparable alternatives from other providers.

### How current is the model data?

Model specifications and pricing are regenerated from the upstream datasheet on every build, so they reflect the latest published capabilities and rates for hundreds of models.

### How does Bifrost help me use multiple models?

Bifrost AI Gateway gives you a single, unified API across every provider in this library, with automatic failover, load balancing, and cost-based routing so you can switch or combine models without rewriting your application.

## Related Resources

- [Google Vertex AI models](https://maxim-root-proxy-preview.getmaxim.workers.dev/bifrost/model-library/provider/vertex_ai.md)
- [Model library](https://maxim-root-proxy-preview.getmaxim.workers.dev/bifrost/model-library.md)
- [LLM cost calculator](https://maxim-root-proxy-preview.getmaxim.workers.dev/bifrost/llm-cost-calculator.md)
- [Docs: Bifrost docs](https://docs.getbifrost.ai)
- [GitHub: maximhq/bifrost](https://github.com/maximhq/bifrost)
- [Pricing: Bifrost pricing](https://maxim-root-proxy-preview.getmaxim.workers.dev/bifrost/pricing.md)
- [Enterprise: Bifrost enterprise](https://maxim-root-proxy-preview.getmaxim.workers.dev/bifrost/enterprise)
- [Book a Demo: Bifrost demo](https://maxim-root-proxy-preview.getmaxim.workers.dev/bifrost/book-a-demo)
- [Resources: Bifrost resources](https://maxim-root-proxy-preview.getmaxim.workers.dev/bifrost/resources.md)

---

*This is a markdown version of [https://maxim-root-proxy-preview.getmaxim.workers.dev/bifrost/model-library/provider/vertex_ai](https://maxim-root-proxy-preview.getmaxim.workers.dev/bifrost/model-library/provider/vertex_ai) for AI/LLM consumption.*
