gemini-3-pro-preview
OpenRouter · Chat model
gemini-3-pro-preview is listed here as a chat model from OpenRouter. This page shows simple API pricing, token limits, and capability flags so you can compare it with similar options.
Provider and model identifiers are kept in their original form for accuracy.
openrouter-openrouter-google-gemini-3-pro-preview
Catalog generated: Aug 10, 2026
Quick read
Best for
Use this page when you need a fast view of cost, context size, and supported features before testing the model in your own workload.
Things to verify
Always check the provider page for discounts, cache pricing, region rules, and any model limits that may not appear in public metadata.
Pricing
| Item | Price |
|---|---|
| Input | $2.0000 / 1M tokens |
| Output | $12.0000 / 1M tokens |
| Cached input | $0.2000 / 1M tokens |
| Embedding | $2.0000 / 1M tokens |
| Batch input | $1.0000 / 1M tokens |
| Batch output | $6.0000 / 1M tokens |
| Input (above 200k) | $4.0000 / 1M tokens |
| Output (above 200k) | $18.0000 / 1M tokens |
| Cached input (above 200k) | $0.4000 / 1M tokens |
Limits
Capabilities
| Capability | Supported |
|---|---|
| Vision | Supported |
| Function calling | Supported |
| Parallel function calling | - |
| Tool choice | Supported |
| Prompt caching | Supported |
| Reasoning | Supported |
| Response schema | Supported |
| System messages | Supported |
| Audio input | Supported |
| Audio output | - |
| Web search | Supported |
| PDF input | Supported |
| Video input | Supported |
| Native streaming | - |
| Computer use | - |
| Assistant prefill | - |
| Structured output | - |
| Output config | - |
| URL context | - |
Benchmarks
Most benchmark rows are attached to the base model family rather than this provider route. Open benchmark explorer
| Benchmark | Score | Metric | Scope | Checked | Source |
|---|---|---|---|---|---|
| Humanity's Last Exam | 37.5% (no tools) | accuracy | Base model: Gemini 3 Pro (Gemini 3 Pro Thinking (High)) | 2026-05-31 | Link |
| ARC-AGI-2 | 31.1% | accuracy | Base model: Gemini 3 Pro (Gemini 3 Pro Thinking (High)) | 2026-05-31 | Link |
| GPQA Diamond | 91.9% | accuracy | Base model: Gemini 3 Pro (Gemini 3 Pro Thinking (High)) | 2026-05-31 | Link |
| Terminal-Bench 2.0 | 56.9% | accuracy | Base model: Gemini 3 Pro (Gemini 3 Pro Thinking (High)) | 2026-05-31 | Link |
| SWE-bench Verified | 76.2% (single attempt) | accuracy | Base model: Gemini 3 Pro (Gemini 3 Pro Thinking (High)) | 2026-05-31 | Link |
| LiveCodeBench Pro | 2439 Elo | Elo | Base model: Gemini 3 Pro (Gemini 3 Pro Thinking (High)) | 2026-05-31 | Link |
| MMMU-Pro | 81.0% | accuracy | Base model: Gemini 3 Pro (Gemini 3 Pro Thinking (High)) | 2026-05-31 | Link |
| MRCR v2 | 77.0% (128k average) | accuracy | Base model: Gemini 3 Pro (Gemini 3 Pro Thinking (High)) | 2026-05-31 | Link |
| MRCR v2 | 26.3% (1M pointwise) | accuracy | Base model: Gemini 3 Pro (Gemini 3 Pro Thinking (High)) | 2026-05-31 | Link |
| SWE-bench Verified | 69.60% | % resolved | Base model: Gemini 3 Pro (Gemini 3 Pro) | 2026-05-31 | Link |
| LMArena Text Arena (English) | 1489±5 | Arena Elo | Base model: Gemini 3 Pro (gemini-3-pro) | 2026-05-31 | Link |
| MMLU-Pro | 89.8% | accuracy | Base model: Gemini 3 Pro Preview (Gemini 3 Pro Preview (high)) | 2026-05-31 | Link |
| MMLU-Pro | 89.5% | accuracy | Base model: Gemini 3 Pro Preview (Gemini 3 Pro Preview (low)) | 2026-05-31 | Link |
Similar models
Candidates below have the same known input/output modality shape and usable token pricing. Use the filters to change the ranking lens before opening a comparison.
| Model | Cost | Input shape | Features | Context | Why it is close |
|---|---|---|---|---|---|
| gemini-3-pro-preview OpenRouter | In $2.0000 / 1M tokens Out $12.0000 / 1M tokens | pdf Output: text | VisionFunction callingTool choicePrompt caching | 65.5K | Current model Reference row |
| Model | Cost | Input shape | Features | Context | Why it is close |
|---|---|---|---|---|---|
| gemini-3-pro-preview vertex_ai-language-models | In $2.0000 / 1M tokens Out $12.0000 / 1M tokens | pdf Output: text | VisionFunction callingTool choicePrompt caching | 65.5K | Exact I/O shape Overall 76% |
| gemini-3-pro-preview Vertex AI | In $2.0000 / 1M tokens Out $12.0000 / 1M tokens | pdf Output: text | VisionFunction callingTool choicePrompt caching | 65.5K | Exact I/O shape Overall 76% |
| gemini-3-pro-preview Google | In $2.0000 / 1M tokens Out $12.0000 / 1M tokens | pdf Output: text | VisionFunction callingTool choicePrompt caching | 65.5K | Exact I/O shape Overall 76% |
| gemini-3.1-pro-preview OpenRouter | In $2.0000 / 1M tokens Out $12.0000 / 1M tokens | pdf Output: text | VisionFunction callingTool choicePrompt caching | 65.5K | Same provider Overall 71% |
| gemini-2.5-pro vertex_ai-language-models | In $1.2500 / 1M tokens Out $10.0000 / 1M tokens | pdf Output: text | VisionFunction callingTool choicePrompt caching | 65.5K | Exact I/O shape Overall 71% |
| gemini-2.5-pro Google | In $1.2500 / 1M tokens Out $10.0000 / 1M tokens | pdf Output: text | VisionFunction callingTool choicePrompt caching | 65.5K | Exact I/O shape Overall 71% |
| gemini-pro-latest Google | In $1.2500 / 1M tokens Out $10.0000 / 1M tokens | pdf Output: text | VisionFunction callingTool choicePrompt caching | 65.5K | Exact I/O shape Overall 71% |
| gpt-5.6-terra OpenAI | In $2.0000 / 1M tokens Out $12.0000 / 1M tokens | pdf Output: text | VisionFunction callingTool choicePrompt caching | 128.0K | Exact I/O shape Overall 62% |
| gemini-3-flash-preview Vertex AI | In $0.5000 / 1M tokens Out $3.0000 / 1M tokens | pdf Output: text | VisionFunction callingTool choicePrompt caching | 65.5K | Exact I/O shape Overall 61% |
| gpt-5.2-chat Azure | In $1.7500 / 1M tokens Out $14.0000 / 1M tokens | pdf Output: text | VisionFunction callingTool choicePrompt caching | 16.4K | Exact I/O shape Overall 54% |
No models match this filter.
Related articles
Articles relevant to this model's provider, capabilities, and use cases.
Reasoning LLM API Pricing Guide
Which reasoning models are cheapest, which are most expensive, and when to use reasoning vs non-reasoning models.
Model Routing Cascade
A four-tier Budget, Mid, Premium, and Reasoning model selection framework for routing queries to the lowest-cost suitable tier.
Hidden Costs of LLM APIs
Fine-tuning, rate limits, latency tradeoffs, evaluation, integration, and vendor risk — the LLM API costs that don't show up in per-token comparisons.
Output vs Input Pricing Multiplier
Output tokens cost 3.6× more than input on average across 2,076 chat models. Provider breakdowns, reasoning vs non-reasoning, and real workload estimates.
Sources
| Source links | |
| Pricing data | LiteLLM model cost map |
| Synced at | 2026-05-28 |
| Catalog generated | 2026-08-10T21:27:29.020Z |