gemini-flash-latest
Google · Chat model
gemini-flash-latest is listed here as a chat model from Google. This page shows simple API pricing, token limits, and capability flags so you can compare it with similar options.
Provider and model identifiers are kept in their original form for accuracy.
gemini-gemini-gemini-flash-latest
Catalog generated: Aug 10, 2026
Quick read
Best for
Use this page when you need a fast view of cost, context size, and supported features before testing the model in your own workload.
Things to verify
Always check the provider page for discounts, cache pricing, region rules, and any model limits that may not appear in public metadata.
Pricing
| Item | Price |
|---|---|
| Input | $0.3000 / 1M tokens |
| Output | $2.5000 / 1M tokens |
| Cached input | $0.0750 / 1M tokens |
| Reasoning output | $2.5000 / 1M tokens |
| Audio input | $1.0000 / 1M tokens |
| Embedding | $0.3000 / 1M tokens |
Limits
Capabilities
| Capability | Supported |
|---|---|
| Vision | Supported |
| Function calling | Supported |
| Parallel function calling | Supported |
| Tool choice | Supported |
| Prompt caching | Supported |
| Reasoning | Supported |
| Response schema | Supported |
| System messages | Supported |
| Audio input | - |
| Audio output | - |
| Web search | Supported |
| PDF input | Supported |
| Video input | - |
| Native streaming | - |
| Computer use | - |
| Assistant prefill | - |
| Structured output | - |
| Output config | - |
| URL context | Supported |
Similar models
Candidates below have the same known input/output modality shape and usable token pricing. Use the filters to change the ranking lens before opening a comparison.
| Model | Cost | Input shape | Features | Context | Why it is close |
|---|---|---|---|---|---|
| gemini-flash-latest Google | In $0.3000 / 1M tokens Out $2.5000 / 1M tokens | pdfurl Output: text | VisionFunction callingParallel function callingTool choice | 65.5K | Current model Reference row |
| Model | Cost | Input shape | Features | Context | Why it is close |
|---|---|---|---|---|---|
| gemini-2.5-flash vertex_ai-language-models | In $0.3000 / 1M tokens Out $2.5000 / 1M tokens | pdfurl Output: text | VisionFunction callingParallel function callingTool choice | 65.5K | Exact I/O shape Overall 76% |
| gemini-2.5-flash-preview-09-2025 vertex_ai-language-models | In $0.3000 / 1M tokens Out $2.5000 / 1M tokens | pdfurl Output: text | VisionFunction callingParallel function callingTool choice | 65.5K | Exact I/O shape Overall 76% |
| Gemini 2.5 Flash Google | In $0.3000 / 1M tokens Out $2.5000 / 1M tokens | pdfurl Output: text | VisionFunction callingParallel function callingTool choice | 65.5K | Same provider Overall 76% |
| gemini-2.5-flash-preview-09-2025 Google | In $0.3000 / 1M tokens Out $2.5000 / 1M tokens | pdfurl Output: text | VisionFunction callingParallel function callingTool choice | 65.5K | Same provider Overall 76% |
| gemini-flash-latest Google | In $0.3000 / 1M tokens Out $2.5000 / 1M tokens | pdfurl Output: text | VisionFunction callingParallel function callingTool choice | 65.5K | Same provider Overall 76% |
| gemini-exp-1206 Google | In $0.3000 / 1M tokens Out $2.5000 / 1M tokens | pdfurl Output: text | VisionFunction callingParallel function callingTool choice | 65.5K | Same provider Overall 76% |
| gemini-3.5-flash-lite Vertex AI | In $0.3000 / 1M tokens Out $2.5000 / 1M tokens | pdfurl Output: text | VisionFunction callingParallel function callingTool choice | 65.5K | Exact I/O shape Overall 72% |
| gemini-3.5-flash-lite Google | In $0.3000 / 1M tokens Out $2.5000 / 1M tokens | pdfurl Output: text | VisionFunction callingParallel function callingTool choice | 65.5K | Same provider Overall 72% |
| gemini-3-flash-preview Google | In $0.5000 / 1M tokens Out $3.0000 / 1M tokens | pdfurl Output: text | VisionFunction callingParallel function callingTool choice | 65.5K | Same provider Overall 71% |
| gemini-2.5-flash-lite vertex_ai-language-models | In $0.1000 / 1M tokens Out $0.4000 / 1M tokens | pdfurl Output: text | VisionFunction callingParallel function callingTool choice | 65.5K | Exact I/O shape Overall 61% |
No models match this filter.
Related articles
Articles relevant to this model's provider, capabilities, and use cases.
LLM API Provider Pricing Comparison
Compare OpenAI, Anthropic, Google, Mistral, and DeepSeek across cheapest chat, mid-range, and reasoning model pricing.
Reasoning LLM API Pricing Guide
Which reasoning models are cheapest, which are most expensive, and when to use reasoning vs non-reasoning models.
Google Provider Guide
Gemini 2.5 Pro, Gemini 2.5 Flash, and Gemini 2.0 Flash use cases, pricing, and when to choose Google.
Model Routing Cascade
A four-tier Budget, Mid, Premium, and Reasoning model selection framework for routing queries to the lowest-cost suitable tier.
Sources
| Source links | |
| Pricing data | LiteLLM model cost map |
| Synced at | 2026-05-28 |
| Catalog generated | 2026-08-10T21:27:29.020Z |